一键导入
local-tool-reliability
Use when local Copilot/Qwen tool calls fail, malformed tool JSON appears, tools are skipped, or LLM Gateway agent mode is flaky.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Use when local Copilot/Qwen tool calls fail, malformed tool JSON appears, tools are skipped, or LLM Gateway agent mode is flaky.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
Use when local Copilot/Qwen/LLM Gateway hits prompt bloat, context limits, slow prefill, timeouts, or 24GB GPU token-budget issues.
Use when local Qwen, llama-server, or ik_llama.cpp dies, stalls, times out, OOMs, cancels, or needs log-based failure diagnosis.
Use when a local-agent coding task needs minimal repo context, targeted file discovery, compact search, or a small implementation brief.
Use when a local Copilot/Qwen chat is too long and needs compaction, checkpointing, restart, or a handoff summary.
Use when local Qwen/Copilot should delegate isolated research, review, or summarization to subagents without bloating parent context.
| name | local-tool-reliability |
| description | Use when local Copilot/Qwen tool calls fail, malformed tool JSON appears, tools are skipped, or LLM Gateway agent mode is flaky. |
| argument-hint | [tool-call symptom or error message] |
Debug and stabilize local-model tool calling before changing application code.
llama-server is running and /v1/models responds.github.copilot.llm-gateway.agentTemperature: 0.0 for tool mode.parallelToolCalling: false for Qwen.enableToolCalling: true unless isolating a basic chat failure.requestTimeout for large local prompts.--jinja).preserve_thinking: true).qwen3_coder tool parser where available.q8_0 when reliability matters.Return: