ソース情報
- リポジトリ
- NeuralBlitz/Mito
- ソースの最終更新活動
- 2026年3月22日 13:29
- 検出された SKILL.md の言語
- 英語
- スター
- 0
- フォーク
- 0
インストール方法
デフォルトでは、最初にソースを確認する Prompt が選択されています。直接コマンドに切り替えるか、ローカルコピーをダウンロードすることもできます。
ソースファイルを確認
インストールを決める前に、SKILL.md と SkillsMP に表示されている付属ファイルをお読みください。
SKILL.md を表示中
SKILL.md
ソースの指示 · 読み取り専用プレビュー- name
- vision-language-models
- description
- Vision-language models, multimodal AI, and visual reasoning
- license
- MIT
- compatibility
- opencode
- metadata
- {"audience":"researchers","category":"machine-learning"}
## What I do
- Build vision-language models
- Work with CLIP, BLIP, LLaVA
- Perform visual question answering
- Create multimodal chatbots
## When to use me
When building systems that combine vision and language.
## Key Concepts
- CLIP
- BLIP and BLIP-2
- LLaVA
- Visual instruction tuning
- Cross-modal attention
- Image-text matching
- Multimodal reasoning
GitHubで見る