Post
80
🚀 Gemma-4-A4B 98e v7-coder cohort — two 20.8B MoE coders (4B-active), fresh-map prunes of Gemma 4 26B-A4B. 30/128 experts dropped per layer from the rebuilt v7 competence maps (audited producers, 10 classes), generic-code 3× + LiveCodeBench 2× on a [24,40] floor, plus the mandatory shared-FFN α=1.2.
📊 Q6_K, llama.cpp, greedy, same host (from summary.json):
🧪 v7-coder — science-augmented (+ targeted_gpqa 1.5). A coder that kept ALL its science: GPQA-D 70.71 (+9.6pp over v6-coder, at PARITY with the unpruned 128e's 67.17), AIME 76.67, MATH-500 92.0, GSM8K 93.0, HE 98.78, HE+ 92.68, LCB-55 96.36, LCB-100 97.0, MultiPL-E 88.67, IFEval 95, ARC 94.8. Edges 128e on GPQA/AIME/GSM8K with no code regression — science recovered by a dedicated targeted_gpqa calibration class.
⚡ v7-coderx — code-maximal (no science term). The strongest coder in the cohort: LCB-med-55 98.18 + LCB-100 99.0 (highest of any Gemma-4 prune to date, +1.8/+2.0pp past the unpruned 128e), MultiPL-E 90.0, HE+ 92.68, HE 95.73, IFEval 95, GSM8K 91, MATH-500 89, AIME 70.0. Whole budget on code; the trade is graduate science (GPQA 48.48).
🎯 v7-coder for strong code + graduate science; v7-coderx for max code (~1.8pp LCB-55 buys ~+22pp GPQA between them). Successor to v6-coder, whose code profile led the 14–22B band (+9pp HE over Qwen2.5-Coder-14B, same-rig).
📦 Each ships bf16 · GGUF (29 tiers + ContribDynamic CD-* + F16 + imatrix + mmproj vision) · NVFP4A16 (native vLLM, ~13 GB) · Ollama (29 tiers + vision-* + :latest=Q4_K_M).
🔗 v7-coder: ManniX-ITA/gemma-4-A4B-98e-v7-coder-it (+ -it-GGUF, -NVFP4A16) · https://ollama.com/mannix/gemma4-98e-v7-coder
🔗 v7-coderx: ManniX-ITA/gemma-4-A4B-98e-v7-coderx-it (+ -it-GGUF, -NVFP4A16) · https://ollama.com/mannix/gemma4-98e-v7-coderx
📊 Q6_K, llama.cpp, greedy, same host (from summary.json):
🧪 v7-coder — science-augmented (+ targeted_gpqa 1.5). A coder that kept ALL its science: GPQA-D 70.71 (+9.6pp over v6-coder, at PARITY with the unpruned 128e's 67.17), AIME 76.67, MATH-500 92.0, GSM8K 93.0, HE 98.78, HE+ 92.68, LCB-55 96.36, LCB-100 97.0, MultiPL-E 88.67, IFEval 95, ARC 94.8. Edges 128e on GPQA/AIME/GSM8K with no code regression — science recovered by a dedicated targeted_gpqa calibration class.
⚡ v7-coderx — code-maximal (no science term). The strongest coder in the cohort: LCB-med-55 98.18 + LCB-100 99.0 (highest of any Gemma-4 prune to date, +1.8/+2.0pp past the unpruned 128e), MultiPL-E 90.0, HE+ 92.68, HE 95.73, IFEval 95, GSM8K 91, MATH-500 89, AIME 70.0. Whole budget on code; the trade is graduate science (GPQA 48.48).
🎯 v7-coder for strong code + graduate science; v7-coderx for max code (~1.8pp LCB-55 buys ~+22pp GPQA between them). Successor to v6-coder, whose code profile led the 14–22B band (+9pp HE over Qwen2.5-Coder-14B, same-rig).
📦 Each ships bf16 · GGUF (29 tiers + ContribDynamic CD-* + F16 + imatrix + mmproj vision) · NVFP4A16 (native vLLM, ~13 GB) · Ollama (29 tiers + vision-* + :latest=Q4_K_M).
🔗 v7-coder: ManniX-ITA/gemma-4-A4B-98e-v7-coder-it (+ -it-GGUF, -NVFP4A16) · https://ollama.com/mannix/gemma4-98e-v7-coder
🔗 v7-coderx: ManniX-ITA/gemma-4-A4B-98e-v7-coderx-it (+ -it-GGUF, -NVFP4A16) · https://ollama.com/mannix/gemma4-98e-v7-coderx