Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Tmax-27b - a Qwen3.6-27b terminal agent for small GPUs trained with DPPO (RL)

Via r/LocalLlama
Tuesday, Jun 23, 2026 · 7:05PM
Summary

What is Tmax-27B? Ai2 just released Tmax, a family of terminal-agent LLMs trained with DPPO (RL) on top of Qwen3.6. The 27B model hits ~43% on Terminal Bench 2.0 and ~69% on TB Lite. These are agentic benchmarks where the model navigates a shell, edits files, runs tests, and completes real dev tasks

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories