Scaling Post-Training Ternarization to Qwen3-8B
The same pipeline scaled from Qwen3-4B to Qwen3-8B. The larger model keeps more of what it knew, 78.5% against 69.6%, so size itself makes a model more robust to aggressive conversion.
Technical research report · September 2026 · Malik, Devan, Mehra