Delegate a bounded implementation task to the fixed Grok worker (grok-4.5) under workspace sandbox and headless execution. Use when a host agent or user needs Grok to implement, fix, or verify a scoped coding task; default effort high unless explicitly supplied.
Create and validate risk-based Playwright browser tests for changed or reported user flows, including real linked UI/backend proof, exact route and network assertions, permissions, persistence, and optional local video evidence. Use when asked for Playwright, browser E2E, UI regression tests, cross-browser flows, route or CORS verification, or explicitly requested PR/MR test evidence.
Record and validate a human-paced MP4 walkthrough of a completed task, feature, or bug fix using the real runnable UI, with acceptance-criterion traceability, privacy checks, deterministic media validation, and optional external publication. Use when asked for a demo video, proof video, walkthrough, screen recording, MP4, or explicitly requested PR/MR video attachment.
Consult Codex on gpt-5.6-sol as a read-only advisor for a focused assumption, tradeoff, architecture question, security concern, difficult diagnosis, or high-risk decision. Use when the user requests a Codex second opinion or when another AI agent needs a bounded independent advisory pass; default reasoning effort to high unless the user explicitly supplies an effort.
Consult Claude Fable as a read-only advisor for a focused assumption, tradeoff, architecture question, security concern, difficult diagnosis, or high-risk decision. Use when the user requests a Fable second opinion or when another AI agent needs a bounded independent advisory pass; default effort to high unless the user explicitly supplies an effort.
Coordinate complex work as a controller-only orchestrator using cost- and risk-aware capability routing, explicit task dependencies and ownership, skeptical proof verification, and an independent review gate. Use when the user asks for delegated execution, parallel agents, controller-only operation, or rigorous multi-agent delivery.
Create or update an evidence-backed description for a pull request, merge request, or equivalent change proposal using the actual base, immutable branch diff, complete file coverage, verified tests and safety claims, and deterministic validation. Use when asked to draft, rewrite, standardize, or remotely update a proposal description.
Run an evidence-backed review of local files, staged or unstaged work, branches, commits, diff ranges, pull requests, merge requests, or equivalent remote proposals, with immutable diff coverage, calibrated tests, a validated decision, and optional explicitly authorized publication. Use when asked to review, audit, inspect, approve, request changes, comment on, or publish feedback for code changes.