ECO-CollecTF: A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts
Author
Hobbs, Elizabeth TGoralski, Stephen M
Mitchell, Ashley
Simpson, Andrew
Leka, Dorjan
Kotey, Emmanuel
Sekira, Matt
Munro, James B
Nadendla, Suvarna
Jackson, Rebecca
Gonzalez-Aguirre, Aitor
Krallinger, Martin
Giglio, Michelle
Erill, Ivan
Date
2021-07-13Journal
Frontiers in Research Metrics and AnalyticsPublisher
Frontiers Media S.A.Type
Article
Metadata
Show full item recordAbstract
Analysis of high-throughput experiments in the life sciences frequently relies upon standardized information about genes, gene products, and other biological entities. To provide this information, expert curators are increasingly relying on text mining tools to identify, extract and harmonize statements from biomedical journal articles that discuss findings of interest. For determining reliability of the statements, curators need the evidence used by the authors to support their assertions. It is important to annotate the evidence directly used by authors to qualify their findings rather than simply annotating mentions of experimental methods without the context of what findings they support. Text mining tools require tuning and adaptation to achieve accurate performance. Many annotated corpora exist to enable developing and tuning text mining tools; however, none currently provides annotations of evidence based on the extensive and widely used Evidence and Conclusion Ontology. We present the ECO-CollecTF corpus, a novel, freely available, biomedical corpus of 84 documents that captures high-quality, evidence-based statements annotated with the Evidence and Conclusion Ontology.Rights/Terms
Copyright © 2021 Hobbs, Goralski, Mitchell, Simpson, Leka, Kotey, Sekira, Munro, Nadendla, Jackson, Gonzalez-Aguirre, Krallinger, Giglio and Erill.Identifier to cite or link to this item
http://hdl.handle.net/10713/16294ae974a485f413a2113503eed53cd6c53
10.3389/frma.2021.674205
Scopus Count
Collections
Related articles
- Concept annotation in the CRAFT corpus.
- Authors: Bada M, Eckert M, Evans D, Garcia K, Shipley K, Sitnikov D, Baumgartner WA Jr, Cohen KB, Verspoor K, Blake JA, Hunter LE
- Issue date: 2012 Jul 9
- New directions in biomedical text annotation: definitions, guidelines and corpus construction.
- Authors: Wilbur WJ, Rzhetsky A, Shatkay H
- Issue date: 2006 Jul 25
- The BioC-BioGRID corpus: full text articles annotated for curation of protein-protein and genetic interactions.
- Authors: Islamaj Dogan R, Kim S, Chatr-Aryamontri A, Chang CS, Oughtred R, Rust J, Wilbur WJ, Comeau DC, Dolinski K, Tyers M
- Issue date: 2017
- FoodBase corpus: a new resource of annotated food entities.
- Authors: Popovski G, Seljak BK, Eftimov T
- Issue date: 2019 Jan 1
- Gold-standard ontology-based anatomical annotation in the CRAFT Corpus.
- Authors: Bada M, Vasilevsky N, Baumgartner WA, Haendel M, Hunter LE
- Issue date: 2017 Jan 1