at-vision
SolidInspect an image, screenshot, photo, diagram, file path, or image URL for a non-vision main model. Prefer the inspect_image MCP tool; if MCP namespace tools are unsupported, use the installed local vision CLI fallback.
Install
Quality Score: 87/100
Skill Content
Details
- Author
- kairyou
- Repository
- kairyou/agent-tools
- Created
- 3 weeks ago
- Last Updated
- yesterday
- Language
- JavaScript
- License
- MIT
Integrates with
Similar Skills
Semantically similar based on skill content — not just same category
screen-vision
Screenshot and read on-screen UI elements/buttons with pixel coordinates (accessibility tree + OCR); optional click. For desktop/native/game GUIs beyond the browser.
oatda-vision-analysis
Use when the user wants to analyze images using vision-capable AI models through OATDA's unified API. Supports OpenAI GPT-4o, Anthropic Claude, and Google Gemini vision models.
fal-ai-media
Use when generating images, video, or audio via the fal.ai MCP server. Covers MCP config, the tool set, model app_ids with per-model parameter schemas (Nano Banana, Seedance, Kling, Veo 3, CSM-1B, ThinkSound), and cost-estimate call shapes.