Published reference2026-09-21
SWE-bench Verified
swe-bench-verified · sample size not reported. Compare only matched task sets and protocols.
MiMo-V2.6-Distill-Qwen-9B · BF16 reference · Xiaomi MiMo · Reported evaluation
↗ Published resultSWE-bench Verified
MiMo-V2.6-Distill-Qwen-9B model card · Evaluation ↗Tokens per second · see measurement basis
As recorded on this receipt
Run configuration & provenance +
- Context
- Not recorded
- Temperature
- Not recorded
- Wall time
- Not recorded
- Weight file
- Provider-managed / not reported
- Revision
- 2367e865d009c13ac81713a2878291d33ab28177
- Reasoning
- Not reported
- Agent version
- Not reported
- Output token limit
- Not reported
- Dataset revision
- Not reported
- Harness host
- Not reported
- Model host
- Published reference
- Input tokens
- Not reported
- Output tokens
- Not reported
- Cached input tokens
- Not reported
- Agent turn / step limit
- Not reported
- Concurrency
- Not reported
- Weight file size
- Not reported
Throughput measurement: Measurement method not separately reported. This should not be treated as standardized decode-only speed.
Xiaomi MiMo model card · SWE-bench Verified avg@3 for the released SFT checkpoint. Sample size not stated on the card. Not a LocalsOnly desk run.
SWE-bench Verified avg@3 = 61.1 on the MiMo-V2.6-Distill-Qwen-9B model card (SFT checkpoint, reported in the MiMo-V2.6 technical report). The card does not state a sample size. Not a LocalsOnly desk run. https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B
Run ID: research-mimo-v2.6-distill-qwen-9b-published-reference-swe-bench-verifiedTask explorer
Passed tasks first, then other outcomes grouped by recorded issue. Use the filters to inspect a particular outcome.
No task records in this receipt.
The headline or summary is available above. Individual outcomes have not been provided.
Explore test coverage →