Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Is a 128 GB MacBook Pro M5 Max actually too slow for large-context local LLM coding workflows?

Via r/LocalLlama
Wednesday, May 27, 2026 · 3:34PM
Summary

People are warning me about the prompt-processing speed of a MacBook Pro M5 Max with 128 GB RAM. My main concern is prompt ingestion / prefill latency and large-context handling — not raw token generation speed (which I think is OK). I only plan to use Qwen 3.5 / 3.6 / 3.7 models or similar mostly c

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories