Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Krasis update: Qwen3.6-35B-A3B (Q4) at reading speed, 1x 8GB 3070 Mobile laptop (32GB RAM)

Via r/LocalLlama
Thursday, May 28, 2026 · 9:42AM
Summary

Context Krasis is an LLM runtime for running models that don't fit into VRAM. Krasis streams the model through VRAM from system RAM efficiently and handles prefill and decode as separate architectures and optimised usecases. Latest results (v1.0 release) 1x Laptop RTX 3070 Mobile 8GB, (35B param, Q4

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories