Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

I built a tiny proxy that gives GLM 5.2 vision (or any text LLM) – MIT

Via r/LocalLlama
Tuesday, Jul 7, 2026 · 7:43PM
Summary

VisionBridge lets you give text-only LLMs vision. It's tiny OpenAI-compatible proxy that lets reasoning models (DeepSeek, Qwen, GLM…) see images by querying a separate vision model through tools: look, OCR, scan, crop, compare. No training, no weights. MIT

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories