Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Tried testing qwen 35b moe model on s26 ultra , without compromising on precision [R] ,[D]

Via r/MachineLearning
Saturday, Jul 18, 2026 · 1:40AM
Summary

Started testing a private qwen 35B moe capacity LLM runtime on s26 ultra, early testing shows that active model footprint can fit within the device’s memory limits.( not sharing the methods or architecture used) and results suggest roughly 90 input processing t/s achievable after optimisation and ou

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories