HAL will be down for maintenance from Friday, June 10 at 4pm through Monday, June 13 at 9am. More information
Skip to Main content Skip to Navigation
Conference papers

Named and specific entity detection in varied data: the Quaero named entity baseline evaluation

Abstract : The Quaero program that promotes research and industrial innovation on technologies for automatic analysis and classification of multimedia and multilingual documents. Within its context a set of evaluations of Named Entity recognition systems was held in 2009. Four tasks were defined. The first two concerned traditional named entities in French broadcast news for one (a rerun of ESTER 2) and of OCR-ed old newspapers for the other. The third was a gene and protein name extraction in medical abstracts. The last one was the detection of references in patents. Four different partners participated, giving a total of 16 systems. We provide a synthetic descriptions of all of them classifying them by the main approaches chosen (resource-based, rules-based or statistical), without forgetting the fact that any modern system is at some point hybrid. The metric (the relatively standard Slot Error Rate) and the results are also presented and discussed. Finally, a process is ongoing with preliminary acceptance of the partners to ensure the availability for the community of all the corpora used with the exception of the non-Quæro produced ESTER 2 one.
Document type :
Conference papers
Complete list of metadata

Cited literature [15 references]  Display  Hide  Download

Contributor : Migration Prodinra Connect in order to contact the contributor
Submitted on : Wednesday, June 3, 2020 - 8:21:41 PM
Last modification on : Tuesday, March 15, 2022 - 3:22:20 AM
Long-term archiving on: : Friday, December 4, 2020 - 5:23:57 PM


Publisher files allowed on an open archive


  • HAL Id : hal-02754184, version 1
  • PRODINRA : 40517


Olivier Galibert, Ludovic Quintard, Sophie Rosset, Pierre Zweigenbaum, Claire Nédellec, et al.. Named and specific entity detection in varied data: the Quaero named entity baseline evaluation. 7. Conference on international language resources and evaluation, May 2010, Valletta, Malta. ⟨hal-02754184⟩



Record views


Files downloads