Medical Long Context Reasoning (MLCR-AA)

An open benchmark from Wisedocs, run by Artificial Analysis, measuring how well models reason over long, fragmented medical records with multi-document synthesis.

ModelsModels with a result.16
Top resultBest result on this test.71.1%Claude Fable 5.1
Top-3 spreadPoints from first to third.15.5 pts
YearYear of release.2026

Result and price

020406080100$0.3$1$3$10$30$100RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

16 results
#ModelResultThe score from the source.
06GLM-5.348.3%
08Kimi K338.3%
16MiniMax M317.2%