Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Power-limit vs TG/s for 2x3090

Via r/LocalLlama
Tuesday, Apr 28, 2026 · 4:48AM
Summary

Trying to find the sweet-spot to tradeoff between power and tg/s. 250W seems to be a sweet spot for Qwen3.6-27B. It's interesting that I got higher tg/s at 275W for 1 concurrent request VLLM-server-config from tedivm vllm serve /models/Qwen3.6-27B-int4-AutoRound --tensor-parallel-size 2 --reasoning-

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories