Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Qwen 3.6 35B crushes Gemma 4 26B on my tests

Via r/LocalLlama
Friday, Apr 17, 2026 · 8:15PM
Summary

I have a personal eval harness: A repo with around 30k lines of code that has 37 intentional issues for LLMs to debug and address through an agentic setup (I use OpenCode) A subset of the harness also has the LLM extract key information from reasonably large PDFs (40-60 pages), summarize and evaluat

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories