Filter, compare, and rank papers, posts, captures, threads, bookmarks, product claims, or research ideas for high-entropy mechanistic insight and underpriced leverage. Use when the user asks for alpha, high entropy, the most intriguing or insightful items,…
tokenbender/agent-guides
SkillsMP has collected 12 skills from tokenbender/agent-guides. Open a skill to review its source and details.
- Latest recorded source activity
- SkillsMP catalog refreshed
- skills collected
- 12
- GitHub stars
- 368
- GitHub forks
- 29
Skills in this repository
Showing 12 of 12 collected skills.
Estimate whether an AI model can complete a task and how long it will take, using METR-style time-horizon modeling. Use when scoping agent work, deciding if a task is within reach, planning retries/parallelism, estimating wall-clock time for SWE/MLE/math…
Audit supervised fine-tuning datasets against the behavior and task they are meant to teach. Use when inspecting SFT JSONL, chat messages, instruction-response pairs, tool or agent trajectories, code corpora, synthetic examples, revised datasets, base-model…
Use as Codex's default rhetoric backbone when writing, reviewing, naming, positioning, debating, or sharpening papers, proposals, essays, launches, social posts, narratives, and claims where the user wants provocative attention-pull, category-defining…
Trigger when: (1) the user asks for Manim, Manim Community, or ManimCE, (2) code contains `from manim import *`, or (3) the task is to build a mathematical explainer animation. Opinionated Manim Community skill for concise math scenes. Focuses on scene…
Canonical end-to-end workflow for turning a paper PDF or URL into grounded OCR artifacts and a teachable notes.md.
High-accuracy OCR refinement workflow with self-scoring, Maj@K consensus voting, and targeted repair for page-faithful transcription.
Use when the task requires automating a real browser from the terminal (navigation, form filling, snapshots, screenshots, data extraction, UI-flow debugging) via `playwright-cli` or the bundled wrapper script.
Use for planning, researching, drafting, revising, or auditing technical write-ups, textbooks, papers, reports, READMEs, research notes, PR narratives, and public technical prose. Applies a full workflow, not only style rules: reader need, source audit,…
Use when planning, launching, tracking, or preserving experiments across projects, especially to separate deterministic operating rules from fuzzy experiment-defining variables that require clarification before costly or irreversible work, and to ensure any…
Issue-led atomic work logging. Use whenever the user says "worklog" or asks to run, apply, start, maintain, update, summarize, or close a worklog; also use for issue-led development, master or umbrella issue tracking, one-issue-per-bug/feature/change…
Use when the user provides an X/Twitter status URL and needs the full thread, context beyond the first post, comparison, summary, intent analysis, title extraction, or reliable post text. Convert status links to Twitter Thread Reader URLs using the tweet id…