Rejected no longer in llama-swap.yaml — entry and/or weights removed after this verdict.
| Run | Score | tok/step | Log |
|---|---|---|---|
| No results-*.log found under this name. | |||
Why rejected: Dense 70B model needs extreme quant to fit 32GB VRAM, destroying quality.
Even basic tasks like 232 returned 2.0 instead of 512.0.
Also ~13.6 t/s (5-6x slower than MoE models) and produces 8-11K thinking tokens per task.
The existing MoE models (Ornith-1.0-35b, Agents-A1-35b, Qwen3.6-35b) all provide better quality
at 5-6x the speed with 256K context. No reason to keep this model.
Removed from: llama-swap.yaml (port 9105), models.json