agent-deployment
Production deployment workflow for agentic systems. Directs to RAG for implementation.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Production deployment workflow for agentic systems. Directs to RAG for implementation.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
This skill should be used when the user asks about "visual builder", "no-code agent builder", "drag and drop", "ThinkingConfig", "extended thinking", "chain of thought", "reasoning", or needs guidance on using ADK's visual development tools or configuring advanced reasoning capabilities.
This skill should be used when the user asks about "callbacks", "lifecycle hooks", "before_model_call", "after_tool_call", "plugins", "session state", "state management", "artifacts", "file uploads", "events", "EventActions", "human-in-the-loop", "confirmation", "memory", "MemoryService", "long-term memory", "remember across sessions", "RAG", "retrieval augmented generation", "grounding", "knowledge base", "vector search", or needs guidance on customizing agent behavior, intercepting execution, managing state across turns, implementing approval workflows, or implementing persistent memory or grounding agent responses in external knowledge.
This skill should be used when the user asks about "creating a new ADK project", "initializing ADK", "setting up Google ADK", "adk create command", "ADK project structure", "YAML agent configuration", "creating an agent", "LlmAgent", "BaseAgent", "custom agent", "agent with different model", "Claude with ADK", "OpenAI with ADK", "LiteLLM", "multi-model agent", or needs guidance on bootstrapping an ADK development environment, authentication setup, choosing between Python code and YAML-based agent definitions, agent configuration, model selection, system instructions, or extending the base agent class for non-LLM logic.
This skill should be used when the user asks about "multi-agent systems", "sub-agents", "delegation", "agent routing", "orchestration", "SequentialAgent", "ParallelAgent", "LoopAgent", "agent-to-agent", "A2A protocol", "agent hierarchy", "streaming", "real-time responses", "SSE", "server-sent events", "websocket", "bidirectional", "Live API", "voice", "audio", "video", "multimodal streaming", or needs guidance on building systems with multiple specialized agents working together or implementing real-time communication patterns.
This skill should be used when the user asks about "deploying", "production", "Agent Engine", "Vertex AI", "Cloud Run", "GKE", "Kubernetes", "hosting", "scaling", "guardrails", "safety", "content filtering", "input validation", "output validation", "authentication", "OAuth", "API keys", "credentials", "security plugins", "testing agents", "evaluation", "evals", "benchmarks", "tracing", "Cloud Trace", "logging", "observability", "AgentOps", "LangSmith", "user simulation", or needs guidance on deploying ADK agents to production environments, implementing safety measures, access control, secure authentication, testing, debugging, monitoring, or evaluating ADK agent quality.
This skill should be used when the user asks about "adding a tool", "FunctionTool", "creating tools", "MCP integration", "OpenAPI tools", "built-in tools", "google_search tool", "code_execution tool", "long-running tools", "async tools", "third-party tools", "LangChain tools", "computer use", or needs guidance on extending agent capabilities with custom functions, API integrations, or external tool frameworks.
| name | Agent Deployment |
| description | Production deployment workflow for agentic systems. Directs to RAG for implementation. |
| Framework | Primary Option | Alternative | RAG Query |
|---|---|---|---|
| ADK | Agent Engine (Vertex AI) | Cloud Run, GKE | "ADK deployment agent engine" |
| OpenAI | Any Python hosting | Serverless, Docker | "openai agents deployment" |
| LangChain | LangServe, Cloud Run | Docker, K8s | "langchain langserve deployment" |
| LangGraph | LangGraph Platform | Cloud Run | "langgraph platform deployment" |
| CrewAI | CrewAI Enterprise | Docker | "crewai deployment production" |
| Anthropic | Any Python hosting | Docker, Serverless | "anthropic agent deployment" |
RAG Query: mcp__agentic-rag__search("[framework] environment configuration", mode="explain")
Production differs from development:
RAG Query: mcp__agentic-rag__search("[framework] dockerfile", mode="build")
RAG Query: mcp__agentic-rag__search("[framework] [platform] deployment", mode="explain")
RAG Query: mcp__agentic-rag__search("[framework] monitoring observability", mode="explain")
| Metric | Alert Threshold | Why It Matters |
|---|---|---|
| Latency p95 | > 5s | User experience |
| Error rate | > 1% | Reliability |
| Token usage | Spike > 200% | Cost control |
| Tool failures | > 5% | Agent effectiveness |
| Routing accuracy | < 90% | Multi-agent health |
RAG Query: mcp__agentic-rag__search("agent input validation security", mode="explain")
RAG Query: mcp__agentic-rag__search("agent guardrails output filtering", mode="explain")
RAG Query: mcp__agentic-rag__search("[framework] secret management", mode="explain")
| Concern | Solution | RAG Query |
|---|---|---|
| Cold starts | Keep warm instances | "[framework] cold start" |
| Concurrent requests | Queue + workers | "[framework] scaling" |
| Token limits | Request batching | "[framework] rate limiting" |
| State persistence | External store | "[framework] state persistence" |