docketrouter
Models / Qwen

Qwen3.8 27B

by Qwen · qwen/qwen3.8-27b

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

reasoningtool-usevisionreleased 2026-08-14
Legal score · raw
-
not yet benchmarked
Context
1M
max output 131K
Input
$0.42
per 1M tokens
Output
$2.55
per 1M tokens
Suite cost
-
run the suite to see

Benchmark results

TaskCategoryRawJuicedCorrectLatencyCostRan
Hearsay IdentificationEvidence------
Bluebook Citation FormatResearch & Writing------
Federal Civil ProcedureProcedure------
Limitations ArithmeticProcedure------
Contract Clause ClassificationContracts------
Citation Hallucination ResistanceReliability------

Measured by DocketBuster

These numbers come from DocketBuster's own legal battery, not from DocketRouter's suite. Latest run per metric, with n and a 95% Wilson interval where the source reports one. See docketbuster.com/benchmarks.

MetricValuenIntervalMeasured
Statute pinpoint, exact section (no retrieval)2.0% (6/300)30095% CI 0.9% to 4.3%2026-08-22
Statute pinpoint, exact section (with DocketBuster retrieval)74.6% (229/307)30795% CI 69.4% to 79.1%2026-08-24
Say-nothing rate (declines to bluff when the answer is not in the record)99.6%28495% CI 98.0% to 99.9%2026-08-22
Abstained on statute pinpoint22.1% (68/307)307count, no interval reported2026-08-24
Coaching quality (GW-14x, 0 to 8)7.48 / 8-rubric mean, no interval reported2026-08-23

Source files: hard-llm-ortier-qwen38-27b.json, gw14x-ortier-qwen38-27b.json, hard-llm-level-qwen38-27b-statute_rag.json. Raw model name in source: qwen/qwen3.8-27b.