Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Opinions/improvements for my Qwen3.6-35B-A3B-FP8 + Hermes Agent setup on NVIDIA DGX Spark?

Via r/LocalLlama
Wednesday, May 20, 2026 · 11:54PM
Summary

I’m running Hermes Agent on a single NVIDIA DGX Spark using vLLM with: docker run --gpus all \ --name qwen36-aggressive \ --restart unless-stopped \ -p 8000:8000 \ --ipc=host \ --ulimit memlock=-1 \ --ulimit stack=67108864 \ --shm-size=32g \ -v ~/.cache/huggingface:/root/.cache/huggingface \ -e VLLM

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories