Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Get 30K more context using Q8 mmproj with Gemma 4

Via r/LocalLlama
Monday, Apr 6, 2026 ยท 8:13AM
Summary

Hey guys, quick follow up to my post yesterday about running Gemma 4 26B. I kept testing and realized you can just use the Q8_0 mmproj for vision instead of F16. There is no quality drop, and it actually performed a bit better in a few of my tests (with --image-min-tokens 300 --image-max-tokens 512)

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories