Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Running Minimax 2.7 at 100k context on strix halo

Via r/LocalLlama
Saturday, May 9, 2026 · 8:21PM
Summary

Just wanted to share because it took me a lot of tweaking to get here: llama-server -hf unsloth/MiniMax-M2.7-GGUF:UD-IQ3_XXS --temp 1.0 --top-k 40 --top-p 0.95 --host 0.0.0.0 --port 8080 -c 100000 -fa on -ngl 999 --no-context-shift -fit off --no-mmap -np 2 --kv-unified --cache-ram 0 -b 1024 -ub 1024

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories