By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
MedsparkMedsparkMedspark
  • Home
  • News & Alerts
    News & AlertsShow More
    Lilly’s TuneLab network adds two wet-lab data partners
    By
    msadmin
    September 18, 2026
    Altis Labs banks $25M as its imaging endpoint beats RECIST
    By
    msadmin
    September 18, 2026
    FDA backs LLM juries to grade AI-written radiology reports
    By
    msadmin
    September 18, 2026
    Mithrl raises $20M for a biomedical world model aimed at R&D
    By
    msadmin
    September 17, 2026
    Voice AI mispronounces one in three new drug names, benchmark finds
    By
    msadmin
    September 17, 2026
  • Spotlight
    SpotlightShow More
    Robinhood Ventures Fund II brings retail capital to healthcare AI
    By
    msadmin
    August 13, 2026
    Adialante brings accessible MRI-based cancer screening
    By
    msadmin
    August 11, 2026
    CellType models biology so AI can discover drugs
    By
    msadmin
    August 11, 2026
    Healthcare AI startups in Robinhood Ventures Fund II
    By
    msadmin
    August 11, 2026
    OpenAI launches GPT-Rosalind for life sciences research
    By
    msadmin
    July 17, 2026
  • Articles
    ArticlesShow More
    Proteomics model picks breast cancer drugs from biopsies
    By
    msadmin
    September 10, 2026
    CRISP model reads frozen sections to steer cancer surgery
    By
    msadmin
    September 10, 2026
    Consumer chatbots are building a medical system outside hospitals
    By
    msadmin
    August 21, 2026
    Reasoning gaps hold back AI agents in scientific discovery
    By
    msadmin
    August 11, 2026
    Benchmark scores can’t track real clinical LLM use, Stanford says
    By
    msadmin
    August 11, 2026
  • About
    • Mission
    • Services
    • Contact
  • Shop
    • All Items
    • By Category
    • Cart
  • Newsletter
Font ResizerAa
MedsparkMedspark
Font ResizerAa
  • Home
  • News & Alerts
  • Spotlight
  • Articles
  • About
  • Shop
  • Newsletter
  • Home
  • News & Alerts
  • Spotlight
  • Articles
  • About
    • Mission
    • Services
    • Contact
  • Shop
    • All Items
    • By Category
    • Cart
  • Newsletter
Follow US
News & Alerts

FDA backs LLM juries to grade AI-written radiology reports

A $1.29M FDA contract will test whether several language models, supervised by radiologists, can evaluate AI-generated reports across roughly a million exams.

MedSpark Staff
By
msadmin
MedSpark Staff
Bymsadmin
Medical, Healthcare, & Biotech/Pharma AI News
Follow:
Published: September 18, 2026
Share
2 Min Read
SHARE

A $1.29M research contract from the FDA will pay for a new method of judging radiology reports that AI writes. The recipient is Cognita Imaging, a subsidiary of Mosaic Clinical Technologies.

The problem is scale. Reader studies remain the gold standard for checking a medical AI system, but expert review cannot cover the range of cases a generative model will meet in practice. Cognita’s answer is what it calls LLMs-as-a-jury: several language models independently grade human-drafted and AI-generated reports, and radiologists adjudicate the disagreements that matter clinically.

Building and validating the framework is step one. Step two pushes it onto a scale reader studies cannot reach: roughly 1 million patient exams from a large, diverse US cohort. Patient groups, care settings, imaging equipment and rare findings all get checked. A second pass rebuilds smaller cohorts to show what a trimmed study would miss.

The project is led by Cognita co-founder Akshay Chaudhari, who teaches radiology at Stanford. Once a model starts writing the whole report, he said, the evaluation problem itself changes. A few hundred cases can show whether a system works narrowly, but not every way it can fail in practice.

Deliverables include software code, comparisons of large and small validation cohorts, and discrepancy studies reviewed by radiologists. Cognita will also hand over guidance for building jury frameworks. The work builds on GREEN, its open-source tool for flagging clinically meaningful differences between two reports.

Why it matters: generative reporting tools are arriving faster than the methods used to check them, and automated review may be the only way to keep pace.

TAGGED:Cognita ImagingFDAGenerative AILLM evaluationMedical ImagingPatient SafetyRadiology
SOURCES:Yahoo Finance / Business Wire
Share This Article
Facebook Copy Link Print
MedSpark Staff
Bymsadmin
Follow:
Medical, Healthcare, & Biotech/Pharma AI News

You Might Also Like

ArticlesNews & Alerts

Michigan Medicine neuroimaging model outperforms GPT in brain scan analysis

By
msadmin
July 12, 2026
News & Alerts

AWS hands Columbia researchers cloud credits for diagnostic AI

By
msadmin
September 14, 2026
News & Alerts

White House convenes clinical AI experts for standards sprint on evaluation

By
msadmin
July 27, 2026
News & Alerts

Lyric acquires Concert to expand AI precision medicine payments

By
msadmin
July 17, 2026

AI news, analysis, and insights for healthcare, biotech, and pharma.

Facebook Twitter Youtube Linkedin
Quick Links
  • News & Alerts
  • Articles
  • Spotlight
  • Events
About Medspark
  • Mission
  • Services
  • Contact

© Copyright 2026 MedSpark. All rights reserved.

Privacy Policy | Legal