Skip to main content

stepfun-vision-skill

Read and understand images for the user, but ONLY when the main model is DeepSeek (model is deepseek-v4-flash or deepseek-v4-pro in config.toml). Use this skill whenever the user sends or pastes an image, attaches a screenshot, references an image file ("看下这张图", "read this image", "screenshot shows..."), or when a message contains the placeholder "image content omitted because you do not support image input". Also use it when you need to inspect image content (OCR, screenshots, diagrams, photos) but your current model cannot process image input. Do not use this skill when the main model is any other provider (e.g. gpt-5.6* on the relay, Gemini, etc.) — those models can see images directly.

Ir a la instalación

Datos de origen

Repositorio
jwangkun/stepfun-vision-skill
Última actividad en el origen
3 de agosto de 2026 a las 10:37
Idioma detectado de SKILL.md
inglés
Estrellas
16
Forks
1

Opciones de instalación

De forma predeterminada está seleccionado el prompt que primero revisa el origen. Puedes cambiar a un comando directo o descargar una copia local.

Revisa los archivos de origen

Lee SKILL.md y los archivos complementarios que muestra SkillsMP antes de decidir si quieres instalarlo.