Call vision models (Doubao, Qwen, OpenAI) to analyze images. Use when you need to understand screenshots, UI layouts, diagrams, or any image content. Supports png/jpg/webp/gif.
At a glance
Manual install
Install
git clone --depth 1 https://github.com/xiincs/claude-code-vision-skill
cp -r claude-code-vision-skill/vision ~/.claude/skills/vision
Can use
Not declared by the author