Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Qwen 3.6 35B A3B Q4_K_M quant evaluation

Via r/LocalLlama
Saturday, Apr 18, 2026 · 10:01AM
Summary

About the Model: 35B total parameters, 3B active (A3B) mixture of experts architecture. Evaluation approach taken: We took Q4_K_M quantized GGUF from Unsloth. Ran it on CPU via llama-cpp-python and tested on three standard benchmarks: - HumanEval (code generation), - HellaSwag (commonsense reasoning

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories