Skip to main content

stepfun-vision-skill

Read and understand images for the user, but ONLY when the main model is DeepSeek (model is deepseek-v4-flash or deepseek-v4-pro in config.toml). Use this skill whenever the user sends or pastes an image, attaches a screenshot, references an image file ("看下这张图", "read this image", "screenshot shows..."), or when a message contains the placeholder "image content omitted because you do not support image input". Also use it when you need to inspect image content (OCR, screenshots, diagrams, photos) but your current model cannot process image input. Do not use this skill when the main model is any other provider (e.g. gpt-5.6* on the relay, Gemini, etc.) — those models can see images directly.

Zur Installation springen

Quellinformationen

Repository
jwangkun/stepfun-vision-skill
Letzte Quellaktivität
3. August 2026 um 10:37
Erkannte Sprache von SKILL.md
Englisch
Sterne
16
Forks
1

Installationsoptionen

Standardmäßig ist der Prompt ausgewählt, der zuerst die Quelle prüft. Sie können zu einem direkten Befehl wechseln oder eine lokale Kopie herunterladen.

Quelldateien prüfen

Lesen Sie SKILL.md und alle von SkillsMP angezeigten Begleitdateien, bevor Sie sich für eine Installation entscheiden.