Qwen3-0.6B bf16frontier
0.6B params · Apache-2.0 · Qwen/Qwen3-0.6B · added Jul 23, 2026
| Benchmark (as printed) | Mode | Shots | Score | Metric |
|---|---|---|---|---|
| MMLU-Reduxheadline | no-think | — | 44.6 | as printed |
| MMLU-Redux | think | — | 55.6 | as printed |
| GPQA-Diamond | no-think | — | 22.9 | as printed |
| GPQA-Diamond | think | — | 27.9 | as printed |
| IFEval strict prompt | no-think | — | 54.5 | as printed |
| IFEval strict prompt | think | — | 59.2 | as printed |
| MATH-500 | no-think | — | 55.2 | as printed |
| MATH-500 | think | — | 77.6 | as printed |
- On-disk size
- 1.50 GBcitedHF file listing — bf16 safetensors, 1,503,300,328 bytes
- Intelligence / GB
- 29.7computed from: quality (MMLU-Redux) (external) ÷ on-disk size (external)
- On-device efficiency
- pendingdecode · prefill · TTFT · peak RAM · cold start · energy — all await Synthiq Labs device runs
- Reproduction config
- ships with the first self-run measurement for this artifact
Dual-mode model: thinking and non-thinking scores are published separately and shown separately here — never averaged. The headline uses the non-thinking mode (the default on-device chat configuration).