FirsthandHealth
medRxiv PreprintsInternational8 October 2026

Large language model consensus for reliable research cohort construction from radiology reports: a retrospective cohort study

This is an official announcement record

Firsthand records what medRxiv Preprints announced and links to the original. The wording below is theirs, not ours.

Background Large-scale retrospective studies often require researchers to adjudicate outcomes or cohort eligibility from unstructured clinical records. Manual review can provide reliable labels but is difficult to scale. We developed and evaluated a consensus workflow using multiple large language models (LLMs) to automatically assign high-confidence outcome labels while deferring ambiguous cases for manual review. Methods We included 6,718 brain MRI reports from 5,856 patients. The expert-adjudicated reference set included 680 reports from 675 patients: 372 normal (age-appropriate without sig
— medRxiv Preprints
Read the official announcement

Opens www.medrxiv.org

More from medRxiv Preprints

This content is for informational purposes only and is not medical advice. It is not intended to diagnose, treat, cure, or prevent any disease. Consult a healthcare professional before starting any supplement, treatment, or program — especially if you are pregnant, nursing, taking medication, or managing a health condition.