Skip to main content

specialist-engineer

Use when deploying external specialist AI agents for complex engineering tasks โ€” domain-expert hiring pipeline, adversarial code review, multi-model routing, and external contractor integration for ๅฐ†ไฝœ็›‘ operations. Based on Nexus Hyper Agent Team (asiflow/claude-nexus-hyper-agent-team, 31 domain experts, 341-assertion test suite) and vibecosystem (vibeeval/vibecosystem, 321โญ, 136 agents, 260 skills) patterns. Do NOT use for in-houseๅ…ญ้ƒจ engineering work (use engineer profile) or for simple single-file changes.

Jump to install

Source facts

Repository
Loveacup/jz-skills
Last source activity
June 4, 2026 at 01:56
Detected SKILL.md language
English
Stars
1
Forks
1

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.

Showing SKILL.md

SKILL.md
Source instructions ยท Read-only preview
name
specialist-engineer
description
Use when deploying external specialist AI agents for complex engineering tasks โ€” domain-expert hiring pipeline, adversarial code review, multi-model routing, and external contractor integration for ๅฐ†ไฝœ็›‘ operations. Based on Nexus Hyper Agent Team (asiflow/claude-nexus-hyper-agent-team, 31 domain experts, 341-assertion test suite) and vibecosystem (vibeeval/vibecosystem, 321โญ, 136 agents, 260 skills) patterns. Do NOT use for in-houseๅ…ญ้ƒจ engineering work (use engineer profile) or for simple single-file changes.
version
1.0.0
author
Hermes Agent (based on asiflow/claude-nexus-hyper-agent-team + vibeeval/vibecosystem)
license
MIT
platforms
["macos","linux"]
metadata
{"hermes":{"tags":["jiangzuojian","specialist","external-contractor","expert-system","engineering"],"related_skills":["kanban-orchestrator","agent-registry","a2a-protocol"]}}
# Specialist Engineer โ€” ๅฐ†ไฝœ็›‘ๅค–่˜ไธ“ๅฎถๅทฅ็จ‹ > Based on Nexus Hyper Agent Team (31 domain experts, 341-assertion contract suite) and vibecosystem (321โญ, 136 agents, 260 skills, self-learning). Adapted for ไธ‰็œๅ…ญ้ƒจ ๅฐ†ไฝœ็›‘ external specialist deployment and management. ## ๐Ÿšจ Red Flags: DO NOT SKIP THIS SKILL | Excuse your brain will make | Why it's wrong | |------------------------------|----------------| | "Engineer can handle this, no need for a specialist" | Generalist agents lack domain depth. A security audit by engineer misses OWASP ASI-02 violations that a security specialist catches in 30 seconds | | "I'll just add a specialist prompt to the existing agent" | Specialist agents need isolated workspaces (Git worktrees), adversarial review (challenger gate), and trust calibration. Prompt hacking doesn't provide these | | "Hiring a specialist agent is too complex for a one-off task" | The gated hiring pipeline (researchโ†’synthesisโ†’validationโ†’challengerโ†’registerโ†’verifyโ†’probationโ†’promote) ensures quality. One bad specialist can corrupt the codebase | | "I can assess specialist quality by reading their output" | Output quality โ‰  process quality. A specialist might produce correct code that introduces supply chain vulnerabilities. Adversarial review catches this | ## When to Use - Complex multi-file refactoring requiring domain expertise (crypto, distributed systems, compiler internals) - Security-critical code that needs adversary-minded review - Cross-stack work (frontend + backend + infra) where no single agent has full coverage - External contractor integration: hiring and managing specialist agents from the talent pool - Code that will be deployed to production and needs 341-assertion contract validation ## Architecture ``` โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”‚ ๅฐ†ไฝœ็›‘ Specialist Hub โ”‚ โ”‚ โ”‚ โ”‚ โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”‚ โ”‚ โ”‚ Talent Scout โ”‚ โ”‚ Recruiter โ”‚ โ”‚ Challenger โ”‚ โ”‚ โ”‚ โ”‚ detect gaps โ”‚ โ”‚ gated hiring โ”‚ โ”‚ adversarial โ”‚ โ”‚ โ”‚ โ”‚ 5-signal conf โ”‚ โ”‚ 8 phases โ”‚ โ”‚ review gate โ”‚ โ”‚ โ”‚ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ”‚ โ”‚ โ”‚ โ”‚ Specialist Pool: โ”‚ โ”‚ โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ” โ”‚ โ”‚ โ”‚ securityโ”‚ โ”‚ crypto โ”‚ โ”‚ frontend โ”‚ โ”‚ distributedโ”‚ โ”‚ โ”‚ โ”‚ auditor โ”‚ โ”‚ engineerโ”‚ โ”‚ architectโ”‚ โ”‚ systems โ”‚ โ”‚ โ”‚ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ”‚ โ”‚ โ”‚ โ”‚ Validation: 341-assertion contract test suite โ”‚ โ”‚ Trust: Bayesian per-agent trust ledger โ”‚ โ””โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ ``` ## Core Capabilities ### 1. Gated Hiring Pipeline When a coverage gap is detected, the talent scout + recruiter hire a new specialist: ``` Phase 1: Research โ€” Study the domain, identify required expertise Phase 2: Synthesis โ€” Draft specialist prompt + toolset spec Phase 3: Validation โ€” Run against known test cases (contract suite) Phase 4: Challenger โ€” Adversarial agent tries to break the specialist Phase 5: Register โ€” Atomic registration in agent-registry Phase 6: Verify โ€” Post-hire smoke test on real task Phase 7: Probation โ€” 5-task observation period with trust score Phase 8: Promote โ€” Full specialist status, trust ledger active ``` ### 2. Coverage Gap Detection ```bash # ๅฐ†ไฝœ็›‘ continuously monitors for capability gaps hermes jiangzuojian detect-gaps # Output: # Gap detected: solidity-smart-contract (confidence: 0.87) # Signal 1: 3 recent tasks required Solidity knowledge # Signal 2: engineer failed 2/3 Solidity tasks # Signal 3: no registered specialist for Solidity # Signal 4: session-sentinel co-signed the gap # Signal 5: estimated cost of gap: $45/week in failed tasks # # Recommendation: HIRE solidity-specialist # Estimated cost: $15/week # ROI: 3x (saves $45/week, costs $15/week) ``` ### 3. Adversarial Review (Challenger Gate) Before any specialist output is accepted, an adversarial agent reviews it: ```bash hermes jiangzuojian challenger-review \ --specialist crypto-engineer \ --task "implement key derivation function" \ --output /tmp/kdf-implementation.py # Challenger tries to: # 1. Inject edge cases that break the implementation # 2. Find timing side-channels # 3. Verify constant-time operations # 4. Check for known-weak PRNG usage # 5. Validate against NIST SP 800-132 ``` ### 4. Trust Ledger (Bayesian Calibration) Every specialist has a trust score calibrated by task outcomes: ```yaml specialist: solidity-expert trust_ledger: total_tasks: 15 successes: 13 failures: 1 challenger_overrides: 1 trust_score: 0.87 # Bayesian posterior domain_scores: erc20: 0.95 erc721: 0.91 custom_contracts: 0.76 # weaker area security_audit: 0.98 trend: improving (+0.03 per 5 tasks) last_incident: "2026-05-20: reentrancy vulnerability missed (caught by challenger)" ``` ### 5. Multi-Model Routing ๅฐ†ไฝœ็›‘ routes tasks to the optimal specialist + model combination: ```bash hermes jiangzuojian route \ --task "audit smart contract for reentrancy vulnerabilities" \ --budget 2.00 # Router decision: # Specialist: solidity-expert (trust: 0.98 on security_audit) # Model: claude-opus-4 (reasoning_effort=high, cost: $1.50) # Fallback: deepseek-v4-pro (if opus unavailable) # Estimated: 3 tool calls, $1.80, 45s ``` ## Specialist Roster (Example) | Specialist | Domain | Trust | Model | Cost/task | |-----------|--------|-------|-------|-----------| | security-auditor | OWASP, injection, auth | 0.95 | claude-sonnet-4 | $0.80 | | crypto-engineer | ECDSA, KDF, TLS | 0.88 | claude-opus-4 | $1.50 | | distributed-systems | Raft, Paxos, CRDTs | 0.82 | deepseek-v4-pro | $0.60 | | performance-profiler | CPU, memory, I/O | 0.91 | haiku | $0.15 | | accessibility-expert | WCAG 2.2, ARIA | 0.94 | gemini-2.5-pro | $0.40 | ## Quick Start ```bash # Initialize specialist hub hermes jiangzuojian init # Hire first specialist hermes jiangzuojian hire \ --domain "security-audit" \ --from registry \ --phases 8 # Run a task with specialist hermes jiangzuojian execute \ --task "audit ~/.hermes/profiles/regent/scripts/kanban_gate.py" \ --specialist security-auditor ``` ## Reference: Upstream Projects | Project | Stars | License | Key Feature | |---------|-------|---------|-------------| | [Nexus Hyper Agent Team](https://github.com/asiflow/claude-nexus-hyper-agent-team) | - | - | 31 experts, 341-assertion suite, gated hiring | | [vibecosystem](https://github.com/vibeeval/vibecosystem) | 321 | - | 136 agents, 260 skills, self-learning | | [specialist-agent](https://github.com/HerbertJulio/specialist-agent) | 10 | MIT | 27 agents, 7 framework packs | ## Integration with ไธ‰็œๅ…ญ้ƒจ ๅฐ†ไฝœ็›‘ operates as a specialist augmentation layer: ``` Kanban chain: planner โ†’ reviewer โ†’ shangshu โ†’ [engineer, ๅฐ†ไฝœ็›‘] โ”‚ โ”Œโ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”˜ โ–ผ ๅฐ†ไฝœ็›‘ Specialist (external contractor) โ”‚ โ”Œโ”€โ”€โ”€โ”€โ”€โ”ดโ”€โ”€โ”€โ”€โ”€โ” challenger trust ledger review calibration ``` ๅฐ†ไฝœ็›‘ differs from ๅ…ต้ƒจ (engineer): - **ๅ…ต้ƒจ**: Standing department, generalist, always online - **ๅฐ†ไฝœ็›‘**: External specialist, domain expert, hired on-demand ## Common Pitfalls - **Specialist overload**: Hiring specialists for every task dilutes the trust ledger. Only hire when gap confidence > 0.75. - **Challenger bypass**: "The code looks fine, skip adversarial review." NEVER. The challenger gate is the last defense. - **Trust score blind spots**: A specialist with 0.95 trust might have 0.50 in a subdomain. Always check domain_scores, not just aggregate. --- ## โœ… Verification Checklist (RUN BEFORE DEPLOYING SPECIALIST OUTPUT) - [ ] Did the specialist pass all 8 phases of the gated hiring pipeline? - [ ] Did the challenger agent review the output (not just the hiring agent)? - [ ] Did I check the specialist's trust score for the SPECIFIC domain (not aggregate)? - [ ] Is the specialist working in an isolated workspace (Git worktree, not shared)? - [ ] Did I log the task outcome to update the trust ledger? **If any box is unchecked, go back.**
View on GitHub