Score: 1

EvidenceOutcomes: a Dataset of Clinical Trial Publications with Clinically Meaningful Outcomes

Published: June 3, 2025 | arXiv ID: 2506.05380v1

By: Yiliang Zhou , Abigail M. Newbury , Gongbo Zhang and more

Potential Business Impact:

Helps computers find important health results in studies.

Business Areas:

Clinical Trials Health Care

The fundamental process of evidence extraction and synthesis in evidence-based medicine involves extracting PICO (Population, Intervention, Comparison, and Outcome) elements from biomedical literature. However, Outcomes, being the most complex elements, are often neglected or oversimplified in existing benchmarks. To address this issue, we present EvidenceOutcomes, a novel, large, annotated corpus of clinically meaningful outcomes extracted from biomedical literature. We first developed a robust annotation guideline for extracting clinically meaningful outcomes from text through iteration and discussion with clinicians and Natural Language Processing experts. Then, three independent annotators annotated the Results and Conclusions sections of a randomly selected sample of 500 PubMed abstracts and 140 PubMed abstracts from the existing EBM-NLP corpus. This resulted in EvidenceOutcomes with high-quality annotations of an inter-rater agreement of 0.76. Additionally, our fine-tuned PubMedBERT model, applied to these 500 PubMed abstracts, achieved an F1-score of 0.69 at the entity level and 0.76 at the token level on the subset of 140 PubMed abstracts from the EBM-NLP corpus. EvidenceOutcomes can serve as a shared benchmark to develop and test future machine learning algorithms to extract clinically meaningful outcomes from biomedical abstracts.

EvidenceBench: A Benchmark for Extracting Evidence from Biomedical Papers

Computation and Language

Finds science facts in papers for researchers.

25 Apr 2025 1

85%

Query-driven Document-level Scientific Evidence Extraction from Biomedical Studies

Computation and Language

Finds best medical answers from many studies.

9 May 2025 1

85%

Integrating Misclassified EHR Outcomes with Validated Outcomes from a Non-probability Sample

Methodology

Improves health records using brain autopsy data.

3 Mar 2025 1

View PDF Login to Bookmark

Country of Origin

🇺🇸 United States

Repos / Data Links

github.com

Page Count

21 pages

EvidenceOutcomes: a Dataset of Clinical Trial Publications with Clinically Meaningful Outcomes

Helps computers find important health results in studies.

Technical Abstract

EvidenceBench: A Benchmark for Extracting Evidence from Biomedical Papers

Query-driven Document-level Scientific Evidence Extraction from Biomedical Studies

Integrating Misclassified EHR Outcomes with Validated Outcomes from a Non-probability Sample