Clinical Benchmarks

MedScribe (Vals AI): current results

Vals AI (dataset with Protege) · 100 rubric-scored SOAP-note cases · index updated August 16, 2026

Claude Opus 5 holds the top current result on MedScribe (Vals AI), 90.99% as of 2026-08, per Vals AI MedScribe leaderboard. Clinical documentation support: quality of SOAP notes generated from clinical visits, scored against rubrics for documentation quality and compliance.

Current results

Result detail

#modelscoreas of
1Anthropic logoClaude Opus 5 Anthropic90.99%2026-08
2Meta logoMuse Spark 1.2 Meta
rank 2; exact value not displayed by the source
~902026-08
3Meta logoMuse Spark 1.1 Meta
rank 3; exact value not displayed by the source
~902026-08
4Anthropic logoClaude Fable 5 Anthropic
rank 4; exact value not displayed by the source
~902026-08
5OpenAI logoGPT 5.1 OpenAI
at-release leader (Feb 2026 blog)
88.09%2026-02
6Anthropic logoClaude Opus 4.6 Anthropic
at-release (Feb 2026 blog)
86.74%2026-02

Scores appear exactly as Vals AI MedScribe leaderboard publishes them (independently run). Vals AI self-runs; 84 models, last updated August 15, 2026. Top scores cluster near 90, so the leaders sit close to the ceiling, and exact decimals below first place are not always displayed.

About the benchmark

publisherVals AI (dataset with Protege)
categorydocumentation and coding benchmarks
released2026-02
size100 rubric-scored SOAP-note cases
scalepercentage accuracy 0-100, higher better
result basisindependently run
sourceVals AI MedScribe leaderboard
last frontier result2026-08

What is MedScribe (Vals AI)?

MedScribe (Vals AI) is a documentation and coding benchmark from Vals AI, released 2026-02: 100 rubric-scored SOAP-note cases, scored on a percentage accuracy 0-100 scale. Clinical documentation support: quality of SOAP notes generated from clinical visits, scored against rubrics for documentation quality and compliance.

Which model leads MedScribe (Vals AI)?

Claude Opus 5 (Anthropic) holds the top current result on MedScribe (Vals AI) at 90.99%, per Vals AI MedScribe leaderboard, as of 2026-08.

Where do the MedScribe (Vals AI) numbers come from?

From Vals AI MedScribe leaderboard (independently run). Vals AI self-runs; 84 models, last updated August 15, 2026. Top scores cluster near 90, so the leaders sit close to the ceiling, and exact decimals below first place are not always displayed.

The rest of the field is on the index, and how sources qualify is on the methodology page. Model names in the table link to cross-benchmark pages.