Skip to main content

vision

Analyze images using a Vision-Language Model through the wu CLI. Use when the user wants to describe an image, caption a picture, ask questions about visual content, inspect renders or screenshots, compare visual evidence, or produce text or JSON answers with public VLM backends.

跳到安装

来源信息

仓库
NVIDIA-Omniverse/usd-content-agents
最近来源活动
2026年7月30日 21:34
检测到的 SKILL.md 语言
英语
星标
188
分支
23

安装方式

默认使用会先检查来源的 Prompt;你也可以切换为直接命令,或下载本地副本。

检查来源文件

决定是否安装前,请先阅读 SKILL.md,以及 SkillsMP 当前展示的配套文件。