Gives text-only LLM coding agents vision by routing images to a multimodal model and returning detailed textual descrip…
Gives text-only LLM coding agents vision by routing images to a multimodal model and returning detailed textual descriptions. Supports local files, URLs, clipboard, base64, raw bytes, and multiple providers like OpenAI, Anthropic, and Gemini.
KuaaMU
mcp
free
No benchmark results have been added yet.