Skip to main content

stepfun-vision-skill

Read and understand images for the user, but ONLY when the main model is DeepSeek (model is deepseek-v4-flash or deepseek-v4-pro in config.toml). Use this skill whenever the user sends or pastes an image, attaches a screenshot, references an image file ("看下这张图", "read this image", "screenshot shows..."), or when a message contains the placeholder "image content omitted because you do not support image input". Also use it when you need to inspect image content (OCR, screenshots, diagrams, photos) but your current model cannot process image input. Do not use this skill when the main model is any other provider (e.g. gpt-5.6* on the relay, Gemini, etc.) — those models can see images directly.

Jump to install

Source facts

Repository
jwangkun/stepfun-vision-skill
Last source activity
August 3, 2026 at 10:37
Detected SKILL.md language
English
Stars
16
Forks
1

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.