Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

My GLM-5.2-FP8 HGX-H200 SGLang docker deploy config

Via r/LocalLlama
Wednesday, Jun 17, 2026 · 6:03PM
Summary

Halo lads. Name says it all. Right now, after 1-2 hours of experimenting, this is maximum i could squeeze out current hardware No, im not rich. Its my companies GPUs, just sharing my experience docker run -d \ --name glm-5.2-sglang \ --restart unless-stopped \ --gpus all \ --shm-size 32g \ --ipc=hos

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories