Recognize and analyze images through OpenAI-compatible vision APIs across multiple platforms.
Copy the install command and let the AI configure it · recommended for beginners
No copy-paste install info for "api-vision-mcp" yet — see the docs or source repo.
Use api-vision-mcp to analyze this product image, identify the brand, category, main colors, and packaging text, and return the result in JSON.
A structured image recognition result containing brand, category, colors, and visible text.
Use api-vision-mcp to analyze this app screenshot, list the buttons, input fields, and navigation areas, and point out possible usability issues.
A list of UI elements and a brief analysis of potential usability issues.
Using api-vision-mcp, design a workflow to batch-recognize dates, amounts, and merchant names from receipt images and output them in a standardized schema.
An automation-friendly recognition workflow and a standardized output schema.
Analyze images with AI for OCR, scene description, detection, and comparison.
Enable any LLM to describe images from paths, URLs, or base64.
Enable vision-less LLMs to understand screenshots and images through a vision proxy.
Generate text descriptions from images with fast fallback vision model support.
Analyze screenshots, text, and UI mockups through one vision MCP tool.
Add screenshot analysis and visual Q&A to OpenCode via OpenAI-compatible vision endpoints.