Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

OpenClaw + oMLX shows 0 cached tokens, but Hermes uses cache fine with the same local model, what am I missing?

Via r/LocalLlama
Monday, May 11, 2026 · 3:31AM
Summary

Hey everyone, I’m trying to debug a weird prompt cache issue with OpenClaw + oMLX, and I’d appreciate help from anyone running local agents on MLX/oMLX. The short version is this: I’m running oMLX v0.3.8 on my Mac, serving: Qwen3.6-35B-A3B-RotorQuant-MLX-4bit OpenClaw runs in Docker on my NAS and co

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories