Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

[P] Gemma 4 running on NVIDIA B200 and AMD MI355X from the same inference stack, 15% throughput gain over vLLM on Blackwell

Via r/MachineLearning
Thursday, Apr 2, 2026 ยท 6:01PM
Summary

Google DeepMind dropped Gemma 4 today: Gemma 4 31B: dense, 256K context, redesigned architecture targeting efficiency and long-context quality Gemma 4 26B A4B: MoE, 26B total / 4B active per forward pass, 256K context Both are natively multimodal (text, image, video, dynamic resolution). We got both

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories