ECO-CollecTF: A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts

dc.contributor.authorHobbs, Elizabeth T.
dc.contributor.authorGoralski, Stephen M.
dc.contributor.authorMitchell, Ashley
dc.contributor.authorSimpson, Andrew
dc.contributor.authorLeka, Dorjan
dc.contributor.authorKotey, Emmanuel
dc.contributor.authorSekira, Matt
dc.contributor.authorMunro, James B.
dc.contributor.authorNadendla, Suvarna
dc.contributor.authorJackson, Rebecca
dc.contributor.authorGonzalez-Aguirre, Aitor
dc.contributor.authorKrallinger, Martin
dc.contributor.authorGiglio, Michelle
dc.contributor.authorErill, Ivan
dc.date.accessioned2021-07-29T13:28:38Z
dc.date.available2021-07-29T13:28:38Z
dc.date.issued2021-07-13
dc.description.abstractAnalysis of high-throughput experiments in the life sciences frequently relies upon standardized information about genes, gene products, and other biological entities. To provide this information, expert curators are increasingly relying on text mining tools to identify, extract and harmonize statements from biomedical journal articles that discuss findings of interest. For determining reliability of the statements, curators need the evidence used by the authors to support their assertions. It is important to annotate the evidence directly used by authors to qualify their findings rather than simply annotating mentions of experimental methods without the context of what findings they support. Text mining tools require tuning and adaptation to achieve accurate performance. Many annotated corpora exist to enable developing and tuning text mining tools; however, none currently provides annotations of evidence based on the extensive and widely used Evidence and Conclusion Ontology. We present the ECO-CollecTF corpus, a novel, freely available, biomedical corpus of 84 documents that captures high-quality, evidence-based statements annotated with the Evidence and Conclusion Ontology.en_US
dc.description.sponsorshipThis work was supported by the National Science Foundation, Division of Biological Infrastructure (1458400) and the National Institutes of Health (R01GM089636, U41HG008735), and by a management commission from Plan TL (Plan de Impulso de las Tecnologías del Lenguaje) of the Spanish Ministerio de Asuntos Económicos y Transformación Digital to BSC-CNS.en_US
dc.description.urihttps://www.frontiersin.org/articles/10.3389/frma.2021.674205/fullen_US
dc.format.extent8 filesen_US
dc.genrejournal articlesen_US
dc.identifierdoi:10.13016/m2wxxg-tv5a
dc.identifier.citationHobbs, Elizabeth T. et al.; ECO-CollecTF: A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts; Frontiers in Research Metrics and Analytics, 13 July, 2021; https://doi.org/10.3389/frma.2021.674205en_US
dc.identifier.urihttps://doi.org/10.3389/frma.2021.674205
dc.identifier.urihttp://hdl.handle.net/11603/22208
dc.language.isoen_USen_US
dc.publisherFrontiersen_US
dc.relation.isAvailableAtThe University of Maryland, Baltimore County (UMBC)
dc.relation.ispartofUMBC Biological Sciences Department Collection
dc.relation.ispartofUMBC Student Collection
dc.relation.ispartofUMBC Faculty Collection
dc.relation.ispartofUMBC Computer Science and Electrical Engineering Department
dc.rightsThis item is likely protected under Title 17 of the U.S. Copyright Law. Unless on a Creative Commons license, for uses protected by Copyright Law, contact the copyright holder or the author.
dc.rightsAttribution 4.0 International (CC BY 4.0)*
dc.rights.urihttps://creativecommons.org/licenses/by/4.0/*
dc.titleECO-CollecTF: A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscriptsen_US
dc.typeTexten_US

Files

Original bundle

Now showing 1 - 5 of 8
Loading...
Thumbnail Image
Name:
frma-06-674205.pdf
Size:
1.86 MB
Format:
Adobe Portable Document Format
Description:
ECO-CollecTF: A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts
Loading...
Thumbnail Image
Name:
DataSheet1_ECO-CollecTF_ A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts.PDF
Size:
398.58 KB
Format:
Adobe Portable Document Format
Description:
Datasheet 1
Loading...
Thumbnail Image
Name:
DataSheet2_ECO-CollecTF_ A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts.PDF
Size:
217.64 KB
Format:
Adobe Portable Document Format
Description:
Datasheet 2
Loading...
Thumbnail Image
Name:
DataSheet3_ECO-CollecTF_ A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts.PDF
Size:
372.37 KB
Format:
Adobe Portable Document Format
Description:
Datasheet 3
Loading...
Thumbnail Image
Name:
DataSheet4_ECO-CollecTF_ A Corpus of Annotated Evidence-Based Assertions in Biomedical Manuscripts.PDF
Size:
247.98 KB
Format:
Adobe Portable Document Format
Description:
Datasheet 4

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
2.56 KB
Format:
Item-specific license agreed upon to submission
Description: