Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

PyTorch model running 170x slower on T4 vs A100. What could cause a bottleneck this extreme? [D]

Via r/MachineLearning
Wednesday, Jul 15, 2026 · 1:44PM
Summary

Hey everyone, Seeing a ~170× slowdown running a point-tracking model on an NVIDIA T4 compared to an A100. On A100 the tracker takes ~0.5 seconds per half-video. On T4 the same call takes ~85 seconds. Video is 47 frames at 256×256, batch 1. I expect a meaningful gap between these cards, but 170× feel

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories