Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Lemonade v10.8: auto memory management, cloud offload, Omni improvements, and call your local models as MCP tools

Via r/LocalLlama
Wednesday, Jun 17, 2026 · 7:42PM
Summary

v10.8 is out, so here's a project update on what landed. This was a 20-contributor release in just 7 days! Smarter memory and context management Dynamic VRAM management now auto-unloads idle models and downsizes their KV-cache to reclaim GPU memory on the fly, plus model pinning so the ones you want

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories