| name | ai-image-gen |
| description | AI 圖像生成工具的完整參考技能。涵蓋 ChatGPT/DALL-E/gpt-image, Nano Banana Pro/Gemini Imagen, Grok Aurora, Midjourney v7, Stable Diffusion, FLUX, Ideogram, Recraft, Luma Photon, Adobe Firefly, Leonardo AI, Playground AI 等平台的功能比較、API 用法、 定價、prompt 撰寫指南與範例庫。 當使用者提到任何 AI 圖像生成、text-to-image、image generation、 AI 繪圖、AI 畫圖、生圖、prompt 撰寫、Midjourney、DALL-E、 Stable Diffusion、FLUX、Imagen、AI art、AI 插畫、AI 設計、 圖片 prompt 等關鍵字時觸發此技能。 也適用於使用者需要選擇圖像生成平台、比較 AI 繪圖工具、 撰寫圖像 prompt、學習 negative prompt、 或需要 API 整合指引的場景。
|
AI Image Generation Skill
涵蓋 12+ 主流 AI 圖像生成平台的完整參考,含功能比較、prompt 撰寫、API 整合
資料截至:2025-07
如何使用此技能
| 需求 | 讀哪裡 |
|---|
| 選擇適合的平台 | 下方「快速選擇指南」或 references/comparison-matrix.md |
| 撰寫圖像 prompt | 下方「Prompt 黃金公式」或 references/prompt-templates.md |
| 查詢特定平台細節 | 對應的 references/platform-*.md |
| 查看 prompt 範例 | references/prompt-examples-*.md |
| 學習跨平台 prompt 差異 | references/cross-platform-prompts.md |
| 學習 Negative Prompt | references/negative-prompt-guide.md |
| API 整合開發 | references/api-quick-reference.md |
| 平台能力雷達圖數據 | references/radar-scores.json |
快速選擇指南
根據需求推薦最佳平台:
| 需求場景 | 首選 | 備選 | 原因 |
|---|
| 攝影寫實 | FLUX.2 [max] | Midjourney V7, GPT Image 1.5 | 極致細節與自然光線 |
| 藝術風格多樣性 | Midjourney V7 | FLUX.2, Leonardo AI | 最豐富的風格表現力 |
| 文字渲染精準 | Ideogram 3.0 | Recraft V4, Playground v3 | 文字幾乎 100% 正確 |
| 向量/SVG 輸出 | Recraft V4 Pro | — | 唯一支援真正 SVG 的平台 |
| 角色一致性 | Midjourney (--cref) | FLUX.2 [max] (多圖ref) | 內建角色參考功能 |
| 本地部署/開源 | FLUX.2 [klein] / SD | — | Apache 2.0 可商用本地部署 |
| 商用版權安全 | Adobe Firefly | — | 完整 Content Credentials |
| 最簡單上手 | ChatGPT (GPT Image) | Grok Aurora, Gemini | 自然語言對話即可 |
| 性價比最高 | Luma Photon Flash | FLUX.1 schnell | $0.004/張最便宜 |
| 產品攝影 | FLUX.2 [max] | Firefly, Midjourney | 精確控制 + 商業品質 |
| 遊戲概念設計 | FLUX.2 [max] | Leonardo AI, Midjourney | 細節 + 風格可控 |
| 社群媒體內容 | GPT Image 1.5 | Ideogram 3.0 | 快速 + 文字精準 |
| Logo/品牌設計 | Recraft V4 Pro | Ideogram 3.0 | SVG 輸出 + 精確色彩 |
| 動漫/Anime | SD + AnimagineXL | Midjourney (--niji 6) | LoRA 提供精確風格控制 |
| 建築視覺化 | FLUX.2 [max] | Midjourney V7 | 結構準確性高 |
Prompt 黃金公式
[Subject 主體] + [Details 細節] + [Setting 場景] + [Composition 構圖] +
[Lighting 光線] + [Style 風格] + [Quality 品質修飾]
公式各段核心詞彙(摘要)
| 段落 | 核心詞彙範例 |
|---|
| Subject | 人物描述、物品材質、場景位置 |
| Composition | portrait, full body, bird's eye view, rule of thirds, 85mm |
| Lighting | golden hour, Rembrandt lighting, rim light, volumetric |
| Style | photorealistic, oil painting, anime, concept art, pixel art |
| Quality | highly detailed, 8K, masterpiece, sharp focus |
完整詞彙表見 references/prompt-templates.md
平台語法差異速查
| 平台 | 語法特點 |
|---|
| ChatGPT / Gemini / Grok | 自然語言句子,不需關鍵字堆疊 |
| Midjourney | 逗號分隔 + --ar --v --s --cref 等參數 |
| SD / FLUX | 逗號分隔 + 正面 prompt + Negative prompt |
| Ideogram | 自然語言,善用文字渲染能力 |
| Recraft | 簡潔描述 + Style preset + Hex 色碼 |
完整跨平台對照見 references/cross-platform-prompts.md
平台狀態速查
| Platform | Provider | Status | Latest Model | Text Render | API | Open Source |
|---|
| ChatGPT | OpenAI | ✅ Active | GPT Image 1.5 | good | ✅ | ✗ |
| Gemini | Google | ✅ Active | Imagen 4 / Nano Banana | excellent | ✅ | ✗ |
| Grok | xAI | ✅ Active | Aurora | good | ✅ | ✗ |
| Midjourney | Midjourney | ✅ Active | V7 + V8 Alpha | fair→good | ⚠️ 無公開 API | ✗ |
| FLUX | BFL | ✅ Active | FLUX.2 [max]/[pro]/[klein] | excellent | ✅ | ✅ (klein) |
| Stable Diffusion | Stability AI | ✅ Active | SD3.5 Large | fair | ✅ | ✅ |
| Ideogram | Ideogram Inc. | ✅ Active | 3.0 | excellent | ✅ | ✗ |
| Recraft | Recraft Inc. | ✅ Active | V4 Pro | excellent | ✅ | ✗ |
| Adobe Firefly | Adobe | ✅ Active | Image5 | good | ✅ | ✗ |
| Leonardo AI | Leonardo/Canva | ✅ Active | Lucid Origin | good | ✅ | ✗ |
| Luma Photon | Luma AI | ✅ Active | Photon-1 | fair | ✅ | ✗ |
| Playground | Playground AI | ✅ Active | PGv3 | excellent | ⚠️ 有限 | ✗ |
API 整合入門
OpenAI GPT Image(最廣泛使用)
from openai import OpenAI
client = OpenAI()
response = client.images.generate(
model="gpt-image-1",
prompt="A cozy bookshop interior with warm lighting",
size="1536x1024",
quality="high",
)
image_url = response.data[0].url
FLUX via fal.ai(最佳品質)
import fal_client
result = fal_client.subscribe(
"fal-ai/flux-pro/v1.1-ultra",
arguments={
"prompt": "A cozy bookshop interior with warm lighting",
"aspect_ratio": "16:9",
},
)
image_url = result["images"][0]["url"]
Stability AI(開源生態)
curl -X POST "https://api.stability.ai/v2beta/stable-image/generate/sd3" \
-H "Authorization: Bearer $STABILITY_API_KEY" \
-H "Accept: image/*" \
-F 'prompt=A cozy bookshop interior with warm lighting' \
-F 'output_format=png' \
-o output.png
完整 API 參考(含 10 個平台的端點、認證、定價、程式碼)見 references/api-quick-reference.md
Negative Prompt 快速指南
支援 Negative Prompt 的平台: SD 全系列、FLUX、Leonardo AI、Playground AI
部分支援: Midjourney (--no)、ChatGPT/Gemini/Grok(自然語言否定)
不支援: Ideogram、Recraft、Firefly
萬用基礎 Negative(SD/FLUX)
lowres, worst quality, low quality, blurry, watermark, text, bad anatomy, extra fingers
場景選擇性添加
- 人物肖像 →
bad hands, deformed face, extra limbs, ugly
- 產品攝影 →
people, busy background, cheap looking
- 風景 →
people, buildings, power lines, urban elements
完整指南含錯誤避免和決策流程見 references/negative-prompt-guide.md
⚠️ 重要注意事項
版權與商用授權
- Adobe Firefly 是唯一提供完整 Content Credentials 的平台,商用最安全
- FLUX.2 [klein] 和 Stable Diffusion 使用 Apache 2.0 / 開源授權,可本地部署
- Midjourney 付費方案可商用,但 Free Trial 生成的圖片不可商用
- ChatGPT / Gemini / Grok 生成圖片的商用權利依各家 ToS,需自行確認
內容政策限制
- 所有主流平台禁止生成 NSFW、暴力、仇恨、兒童不當內容
- ChatGPT 和 Gemini 限制與真實公眾人物相似的圖片
- Stable Diffusion 本地部署無平台限制,但使用者需自負法律責任
- FLUX.2 API 有
safety_tolerance 參數可調整(1-6)
已知限制
- Midjourney 無公開 API,僅能透過 Discord 或 Web UI 使用
- Playground AI 的 API 存取較受限
- 所有 AI 模型的人物手指/手部仍可能出錯,需 Negative prompt 或多次重試
深度資料索引
平台詳細資料
| 平台 | 檔案路徑 |
|---|
| ChatGPT / GPT Image / DALL-E | references/platform-chatgpt-openai.md |
| Gemini / Imagen / Nano Banana | references/platform-gemini-imagen.md |
| Grok / Aurora | references/platform-grok-aurora.md |
| Midjourney | references/platform-midjourney.md |
| FLUX / BFL | references/platform-stablediffusion-flux.md |
| Stable Diffusion / Stability AI | references/platform-stablediffusion-flux.md |
| Ideogram | references/platform-ideogram.md |
| Recraft | references/platform-recraft.md |
| Adobe Firefly | references/platform-adobe-firefly.md |
| Leonardo AI | references/platform-leonardo-ai.md |
| Luma Photon | references/platform-luma-photon.md |
| Playground AI | references/platform-playground-ai.md |
Prompt 範例庫
| 類別 | 檔案路徑 | 範例數 |
|---|
| 通用模板與詞彙表 | references/prompt-templates.md | 10 templates |
| 寫實攝影 | references/prompt-examples-photorealistic.md | 10 examples |
| 藝術/插畫 | references/prompt-examples-artistic.md | 12 examples |
| 設計/商業應用 | references/prompt-examples-design.md | 12 examples |
| 角色/人物設計 | references/prompt-examples-character.md | 12 examples |
| 跨平台同場景對照 | references/cross-platform-prompts.md | 5 scenes × 6 platforms |
工具與參考
| 資源 | 檔案路徑 |
|---|
| Negative Prompt 完整指南 | references/negative-prompt-guide.md |
| API 快速參考 | references/api-quick-reference.md |
| 平台比較矩陣 | references/comparison-matrix.md |
| 能力雷達圖數據 | references/radar-scores.json |