| name | venice-ai |
| description | Complete Venice AI platform — text generation, web search, embeddings, TTS, speech-to-text, image generation, video creation, upscaling, and AI editing. Private, uncensored AI inference for everything. |
| homepage | https://venice.ai |
| metadata | {"clawdbot":{"emoji":"🏛️","requires":{"bins":["python3"],"env":["VENICE_API_KEY"]},"primaryEnv":"VENICE_API_KEY","notes":"Full Venice AI platform — text, images, video, audio, embeddings. One skill for everything.","install":[{"id":"python-brew","kind":"brew","formula":"python","bins":["python3"],"label":"Install Python (brew)"}]}} |
Venice AI — Complete Platform Skill
What is Venice AI?
Venice AI is the AI platform for people who don't want Big Tech watching over their shoulder. Unlike OpenAI, Anthropic, and Google — where every prompt is logged, analyzed, and potentially used to train future models — Venice offers true privacy with zero data retention on private models. Your conversations stay yours.
Venice is also uncensored: no content filters, no refusals, no "I can't help with that." The AI answers what you actually ask, whether that's creative fiction, security research, controversial topics, or anything else mainstream AI platforms won't touch.
Why agents should care:
- Private inference — Models marked "Private" have zero data retention. Process sensitive data without it leaking into training sets.
- Uncensored responses — No guardrails blocking legitimate use cases. Generate any content your workflow needs.
- OpenAI-compatible API — Drop-in replacement. Same API format, just change the base URL.
- 30+ models — From tiny efficient models to Claude Opus 4.5, GPT-5.2, and Venice's own uncensored models.
- Built-in web search — LLMs can search the web and cite sources in a single API call.
- Image & video generation — Flux, Sora, Runway, WAN models for visual content.
This skill gives you the complete Venice platform: text generation, web search, embeddings, TTS, speech-to-text, image generation, video creation, upscaling, and AI editing.
⚠️ API changes: If something doesn't work as expected, check docs.venice.ai — the API specs may have been updated since this skill was written.
Prerequisites
Setup
Get Your API Key
- Create account at venice.ai
- Go to venice.ai/settings/api
- Click "Create API Key" → copy the key (starts with
vn_...)
Configure
Option A: Environment variable
export VENICE_API_KEY="vn_your_key_here"
Option B: Clawdbot config (recommended)
// ~/.clawdbot/clawdbot.json
{
skills: {
entries: {
"venice-ai": {
env: { VENICE_API_KEY: "vn_your_key_here" }
}
}
}
}
Verify
python3 {baseDir}/scripts/venice.py models --type text
Scripts Overview
| Script | Purpose |
|---|
venice.py | Text generation, models, embeddings, TTS, transcription |
venice-image.py | Image generation (Flux, etc.) |
venice-video.py | Video generation (Sora, WAN, Runway) |
venice-upscale.py | Image upscaling |
venice-edit.py | AI image editing |
Part 1: Text & Audio
Model Discovery & Selection
Venice has a huge model catalog spanning text, image, video, audio, and embeddings.
Browse Models
python3 {baseDir}/scripts/venice.py models --type text
python3 {baseDir}/scripts/venice.py models --type image
python3 {baseDir}/scripts/venice.py models --type text,image,video,audio,embedding
python3 {baseDir}/scripts/venice.py models --filter llama
Model Selection Guide
| Need | Recommended Model | Why |
|---|
| Cheapest text | qwen3-4b ($0.05/M in) | Tiny, fast, efficient |
| Best uncensored | venice-uncensored ($0.20/M in) | Venice's own uncensored model |
| Best private + smart | deepseek-v3.2 ($0.40/M in) | Great reasoning, efficient |
| Vision/multimodal | qwen3-vl-235b-a22b ($0.25/M in) | Sees images |
| Best coding | qwen3-coder-480b-a35b-instruct ($0.75/M in) | Massive coder model |
| Frontier (budget) | grok-41-fast ($0.50/M in) | Fast, 262K context |
| Frontier (max quality) | claude-opus-4-6 ($6/M in) | Best overall quality |
| Reasoning | kimi-k2-5 ($0.75/M in) | Strong chain-of-thought |
| Web search | Any model + enable_web_search | Built-in web search |
Text Generation (Chat Completions)
Basic Generation
python3 {baseDir}/scripts/venice.py chat "What is the meaning of life?"
python3 {baseDir}/scripts/venice.py chat "Explain quantum computing" --model deepseek-v3.2
python3 {baseDir}/scripts/venice.py chat "Review this code" --system "You are a senior engineer."
echo "Summarize this" | python3 {baseDir}/scripts/venice.py chat --model qwen3-4b
python3 {baseDir}/scripts/venice.py chat "Write a story" --stream
Web Search Integration
python3 {baseDir}/scripts/venice.py chat "What happened in tech news today?" --web-search auto
python3 {baseDir}/scripts/venice.py chat "Current Bitcoin price" --web-search on --web-citations
python3 {baseDir}/scripts/venice.py chat "Summarize: https://example.com/article" --web-scrape
Uncensored Mode
python3 {baseDir}/scripts/venice.py chat "Your question" --model venice-uncensored
python3 {baseDir}/scripts/venice.py chat "Your prompt" --no-venice-system-prompt
Reasoning Models
python3 {baseDir}/scripts/venice.py chat "Solve this math problem..." --model kimi-k2-5 --reasoning-effort high
python3 {baseDir}/scripts/venice.py chat "Debug this code" --model qwen3-4b --strip-thinking
Advanced Options
python3 {baseDir}/scripts/venice.py chat "Be creative" --temperature 1.2 --max-tokens 4000
python3 {baseDir}/scripts/venice.py chat "List 5 colors as JSON" --json
python3 {baseDir}/scripts/venice.py chat "Question" --cache-key my-session-123
python3 {baseDir}/scripts/venice.py chat "Hello" --show-usage
Embeddings
Generate vector embeddings for semantic search, RAG, and recommendations:
python3 {baseDir}/scripts/venice.py embed "Venice is a private AI platform"
python3 {baseDir}/scripts/venice.py embed "first text" "second text" "third text"
python3 {baseDir}/scripts/venice.py embed --file texts.txt
python3 {baseDir}/scripts/venice.py embed "some text" --output json
Model: text-embedding-bge-m3 (private, $0.15/M tokens)
Text-to-Speech (TTS)
Convert text to speech with 60+ multilingual voices:
python3 {baseDir}/scripts/venice.py tts "Hello, welcome to Venice AI"
python3 {baseDir}/scripts/venice.py tts "Exciting news!" --voice af_nova
python3 {baseDir}/scripts/venice.py tts --list-voices
python3 {baseDir}/scripts/venice.py tts "Some text" --output /tmp/speech.mp3
python3 {baseDir}/scripts/venice.py tts "Speaking slowly" --speed 0.8
Popular voices: af_sky, af_nova, am_liam, bf_emma, zf_xiaobei (Chinese), jm_kumo (Japanese)
Model: tts-kokoro (private, $3.50/M characters)
Speech-to-Text (Transcription)
Transcribe audio files to text:
python3 {baseDir}/scripts/venice.py transcribe audio.wav
python3 {baseDir}/scripts/venice.py transcribe recording.mp3 --timestamps
python3 {baseDir}/scripts/venice.py transcribe --url https://example.com/audio.wav
Supported formats: WAV, FLAC, MP3, M4A, AAC, MP4
Model: nvidia/parakeet-tdt-0.6b-v3 (private, $0.0001/audio second)
Check Balance
python3 {baseDir}/scripts/venice.py balance
Part 2: Images & Video
Pricing Overview
| Feature | Cost |
|---|
| Image generation | ~$0.01-0.03 per image |
| Image upscale | ~$0.02-0.04 |
| Image edit | $0.04 |
| Video (WAN) | ~$0.10-0.50 |
| Video (Sora) | ~$0.50-2.00 |
| Video (Runway) | ~$0.20-1.00 |
Use --quote with video commands to check pricing before generation.
Image Generation
python3 {baseDir}/scripts/venice-image.py --prompt "a serene canal in Venice at sunset"
python3 {baseDir}/scripts/venice-image.py --prompt "cyberpunk city" --count 4
python3 {baseDir}/scripts/venice-image.py --prompt "portrait" --width 768 --height 1024
python3 {baseDir}/scripts/venice-image.py --list-models
python3 {baseDir}/scripts/venice-image.py --list-styles
python3 {baseDir}/scripts/venice-image.py --prompt "fantasy" --model flux-2-pro --style-preset "Cinematic"
python3 {baseDir}/scripts/venice-image.py --prompt "abstract" --seed 12345
Key flags: --prompt, --model (default: flux-2-max), --count, --width, --height, --format (webp/png/jpeg), --resolution (1K/2K/4K), --aspect-ratio, --negative-prompt, --style-preset, --cfg-scale (0-20), --seed, --safe-mode, --hide-watermark, --embed-exif
Image Upscale
python3 {baseDir}/scripts/venice-upscale.py photo.jpg --scale 2
python3 {baseDir}/scripts/venice-upscale.py photo.jpg --scale 4 --enhance
python3 {baseDir}/scripts/venice-upscale.py photo.jpg --enhance --enhance-prompt "sharpen details"
python3 {baseDir}/scripts/venice-upscale.py --url "https://example.com/image.jpg" --scale 2
Key flags: --scale (1-4, default: 2), --enhance (AI enhancement), --enhance-prompt, --enhance-creativity (0.0-1.0), --url, --output
Image Edit
AI-powered image editing:
python3 {baseDir}/scripts/venice-edit.py photo.jpg --prompt "add sunglasses"
python3 {baseDir}/scripts/venice-edit.py photo.jpg --prompt "change the sky to sunset"
python3 {baseDir}/scripts/venice-edit.py photo.jpg --prompt "remove the person in background"
python3 {baseDir}/scripts/venice-edit.py --url "https://example.com/image.jpg" --prompt "colorize"
Note: The edit endpoint uses Qwen-Image which has some content restrictions.
Video Generation
python3 {baseDir}/scripts/venice-video.py --quote --model wan-2.6-image-to-video --duration 10s
python3 {baseDir}/scripts/venice-video.py --image photo.jpg --prompt "camera pans slowly" --duration 10s
python3 {baseDir}/scripts/venice-video.py --image photo.jpg --prompt "cinematic" \
--model sora-2-image-to-video --duration 8s --aspect-ratio 16:9 --skip-audio-param
python3 {baseDir}/scripts/venice-video.py --video input.mp4 --prompt "anime style" \
--model runway-gen4-turbo-v2v
python3 {baseDir}/scripts/venice-video.py --list-models
Key flags: --image or --video, --prompt, --model (default: wan-2.6-image-to-video), --duration, --resolution (480p/720p/1080p), --aspect-ratio, --audio/--no-audio, --quote, --timeout
Models:
- WAN — Image-to-video, configurable audio, 5s-21s
- Sora — Requires
--aspect-ratio, use --skip-audio-param
- Runway — Video-to-video transformation
Tips & Ideas
🔍 Web Search + LLM = Research Assistant
Use --web-search on --web-citations to build a research workflow. Venice searches the web, synthesizes results, and cites sources — all in one API call.
🔓 Uncensored Creative Content
Venice's uncensored models work for both text AND images. No guardrails blocking legitimate creative use cases.
🎯 Prompt Caching for Agents
If you're running an agent loop that sends the same system prompt repeatedly, use --cache-key to get up to 90% cost savings.
🎤 Audio Pipeline
Combine TTS and transcription: generate spoken content with tts, process audio with transcribe. Both are private inference.
🎬 Video Workflow
- Generate or find a base image
- Use
--quote to estimate video cost
- Generate with appropriate duration/model
- Videos take 1-5 minutes depending on settings
Troubleshooting
| Problem | Solution |
|---|
VENICE_API_KEY not set | Set env var or configure in ~/.clawdbot/clawdbot.json |
Invalid API key | Verify at venice.ai/settings/api |
Model not found | Run --list-models to see available; use --no-validate for new models |
| Rate limited | Check --show-usage output |
| Video stuck | Videos can take 1-5 min; use --timeout 600 for long ones |
Resources