Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

how to preserve gemma 4 thinking trace

Via r/LocalLlama
Thursday, Apr 23, 2026 · 5:24AM
Summary

how can i prevent discarding the thinking trace? llama.cpp (b8858) serving gemma 4 31b (UD-Q6_K_XL), (almost) vanilla pi harness got some flags here and there on llama-server, nothing relevant, but adding --jinja and --chat-template-kwargs ‘{“preserve_thinking”: true}’ didn’t seem to change it

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories