docketrouter
Models / Mistral

Mistral Large

by Mistral · mistralai/mistral-large

This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

tool-usereleased 2024-02-26
Legal score · raw
-
not yet benchmarked
Context
128K
max output 102K
Input
$2
per 1M tokens
Output
$6
per 1M tokens
Suite cost
-
run the suite to see

Benchmark results

TaskCategoryRawJuicedCorrectLatencyCostRan
Hearsay IdentificationEvidence------
Bluebook Citation FormatResearch & Writing------
Federal Civil ProcedureProcedure------
Limitations ArithmeticProcedure------
Contract Clause ClassificationContracts------
Citation Hallucination ResistanceReliability------

Measured by DocketBuster

These numbers come from DocketBuster's own legal battery, not from DocketRouter's suite. Latest run per metric, with n and a 95% Wilson interval where the source reports one. See docketbuster.com/benchmarks.

MetricValuenIntervalMeasured
Statute pinpoint, exact section (no retrieval)20.7% (62/300)30095% CI 16.5% to 25.6%2026-08-22
Statute pinpoint, exact section (with DocketBuster retrieval)77.9% (239/307)30795% CI 72.9% to 82.1%2026-08-24
Say-nothing rate (declines to bluff when the answer is not in the record)100.0%18995% CI 98.0% to 100.0%2026-08-22
Abstained on statute pinpoint5.9% (18/307)307count, no interval reported2026-08-24
Coaching quality (GW-14x, 0 to 8)6.92 / 8-rubric mean, no interval reported2026-08-23

Source files: hard-llm-ortier-mistrallg.json, gw14x-ortier-mistrallg.json, hard-llm-level-mistrallg-statute_rag.json. Raw model name in source: mistralai/mistral-large-2512.