
MLCR v1.1: Calibrating LLM Judges for Long-Context Medical Reasoning
Earlier this summer, we released the Medical Long Context Reasoning benchmark, or MLCR, to measure how well large language models reason across long, fragmented medical records. Wisedocs is proud to announce we have now released MLCR with Artificial Analysis as MLCR-AA, an independently run evaluation of frontier closed- and open-weight models.

