By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
MedsparkMedsparkMedspark
  • Home
  • News & Alerts
    News & AlertsShow More
    Alibaba hands hospitals a free AI that reads 146 conditions
    By
    msadmin
    September 20, 2026
    Angle Health banks $600M to sell AI-native plans to small businesses
    By
    msadmin
    September 20, 2026
    Medtronic wins FDA nod to bring LigaSure onto the Hugo robot
    By
    msadmin
    September 20, 2026
    St. Luke’s hands 400,000 patient calls a month to AI agents
    By
    msadmin
    September 20, 2026
    Qure.ai wins Europe’s toughest CE class for a primary care AI
    By
    msadmin
    September 20, 2026
  • Spotlight
    SpotlightShow More
    Robinhood Ventures Fund II brings retail capital to healthcare AI
    By
    msadmin
    August 13, 2026
    Adialante brings accessible MRI-based cancer screening
    By
    msadmin
    August 11, 2026
    CellType models biology so AI can discover drugs
    By
    msadmin
    August 11, 2026
    Healthcare AI startups in Robinhood Ventures Fund II
    By
    msadmin
    August 11, 2026
    OpenAI launches GPT-Rosalind for life sciences research
    By
    msadmin
    July 17, 2026
  • Articles
    ArticlesShow More
    Proteomics model picks breast cancer drugs from biopsies
    By
    msadmin
    September 10, 2026
    CRISP model reads frozen sections to steer cancer surgery
    By
    msadmin
    September 10, 2026
    Consumer chatbots are building a medical system outside hospitals
    By
    msadmin
    August 21, 2026
    Reasoning gaps hold back AI agents in scientific discovery
    By
    msadmin
    August 11, 2026
    Benchmark scores can’t track real clinical LLM use, Stanford says
    By
    msadmin
    August 11, 2026
  • About
    • Mission
    • Services
    • Contact
  • Shop
    • All Items
    • By Category
    • Cart
  • Newsletter
Font ResizerAa
MedsparkMedspark
Font ResizerAa
  • Home
  • News & Alerts
  • Spotlight
  • Articles
  • About
  • Shop
  • Newsletter
  • Home
  • News & Alerts
  • Spotlight
  • Articles
  • About
    • Mission
    • Services
    • Contact
  • Shop
    • All Items
    • By Category
    • Cart
  • Newsletter
Follow US
News & Alerts

Stanford and Harvard launch MAST framework for tracking clinical AI progress

Stanford and Harvard researchers debut MAST v1.0, a new benchmark framework for measuring clinical AI performance across real healthcare tasks.

MedSpark Staff
By
msadmin
MedSpark Staff
Bymsadmin
Medical, Healthcare, & Biotech/Pharma AI News
Follow:
Published: July 30, 2026
Share
2 Min Read
SHARE

Stanford and Harvard researchers have launched the Medical AI Superintelligence Test, or MAST, a new framework designed to track how clinical AI models perform across meaningful healthcare tasks.

Published in Nature Medicine by the ARISE Healthcare Network, MAST v1.0 measures AI models across clinical domains like diagnosis, management, and safety, as well as general domains including radiology and multimodal reasoning. The goal is a shared infrastructure for rapid, high-quality AI benchmarking.

The researchers found that no single frontier model dominates. “The highest-performing models are closely clustered on the composite rankings, but the ordering changes across clinical dimensions,” the team reported. “There is no single model that dominates.”

ARISE, which stands for AI Research and Science Evaluation, was established by Stanford in 2024 as a collaborative network of academic medical centers. The group argues that existing benchmarks are misleading and insufficient for assessing AI safety in real clinical settings.

“A model can perform well on knowledge questions while failing to integrate that knowledge in complex clinical scenarios,” ARISE explained. “It can make the right diagnosis while recommending an unsafe plan.”

The launch version maintains evaluations with standardized model runs and public reporting across diagnosis, management, and other domains. Future iterations will cross-benchmark trait analysis and use evaluations as probes of underlying clinical behaviors, since traits like aggressiveness cannot be captured by a single test.

The framework builds on the NOHARM study, which evaluated clinical AI safety across 45 large language models and four clinical AI systems. ARPA-H recently awarded $3.8 million to Stanford and Beth Israel Deaconess researchers under its PACT project to build task-first AI benchmarks grounded in electronic health record data.

TAGGED:AI benchmarkingAI evaluationARISEClinical AIHarvardMASTNature MedicineStanford
SOURCES:Healthcare IT News
Share This Article
Facebook Copy Link Print
MedSpark Staff
Bymsadmin
Follow:
Medical, Healthcare, & Biotech/Pharma AI News

You Might Also Like

News & Alerts

Angle Health banks $600M to sell AI-native plans to small businesses

By
msadmin
September 20, 2026
News & Alerts

10x Genomics adopts Lunit’s H&E AI for biomarker discovery

By
msadmin
September 13, 2026
News & AlertsSpotlight

Pharma AI Alliance Expands: Owkin and AstraZeneca Deploy New Drug Discovery Models

By
msadmin
May 14, 2026
News & Alerts

Arintra nets $25M Series B to plug hospital revenue leaks

By
msadmin
August 27, 2026

AI news, analysis, and insights for healthcare, biotech, and pharma.

Facebook Twitter Youtube Linkedin
Quick Links
  • News & Alerts
  • Articles
  • Spotlight
  • Events
About Medspark
  • Mission
  • Services
  • Contact

© Copyright 2026 MedSpark. All rights reserved.

Privacy Policy | Legal