r/LocalLLaMA 🤖 Ai 👁 0

I built a tiny proxy that gives GLM 5.2 vision (or any text LLM) – MIT

VisionBridge lets you give text-only LLMs vision. It's tiny OpenAI-compatible proxy that lets reasoning models (DeepSeek, Qwen, GLM…) see images by querying a separate vision model through tools: look, OCR, scan, crop,

I built a tiny proxy that gives GLM 5.2 vision (or any text LLM) – MIT
📄

This source provides headlines only. Use the button below to read the complete article on the original site.

📰 Read the original article on r/LocalLLaMA

Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.