Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Newer Qwen models are worse at summarization?

Via r/LocalLlama
Tuesday, Jun 9, 2026 · 8:15PM
Summary

We have summaries annotated by real humans that we benchmark various models, using an LLM as a judge, we found that in the 30B params range, Qwen 3 tops it out, followed by Gemma 4. It feels like newer Qwens are optimized to perform agentic tasks?

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories