Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Do VLMs in production still use fixed-patch ViTs for their vision capabilities? [D]

Via r/MachineLearning
Thursday, May 21, 2026 · 2:46PM
Summary

The research community has provided (already for some time) seemingly more efficient and effective tokenizations for vision. Do we have any hint on whether non-fixed-patches tokenization is being applied on the big player models? I imagine not, and I'm trying to think why: - marginal gains? - pipeli

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories