Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

DifussionGemma 4 on 4x7900xtx

Via r/LocalLlama
Thursday, Jun 11, 2026 · 3:18PM
Summary

Just got 100 tps on generation, but in total time it around 45-60 t/s in case of prompt processing waiting. Available memory show: GPU KV cache size: 152,671 tokens Maximum concurrency for 131,072 tokens per request: 1.16x amd-smi monitor for this gpu: GPU XCP POWER GPU_T MEM_T GFX_CLK GFX% MEM% ENC

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories