Skip to main content
Manusで任意のスキルを実行
ワンクリックで
linuxarena
GitHub クリエイタープロフィール

linuxarena

3 件の GitHub リポジトリにある 30 件の収集済み skills をリポジトリ単位で表示します。

収集済み skills
30
リポジトリ
3
更新
2026-07-17
リポジトリエクスプローラー

リポジトリと代表的な skills

ship
ソフトウェア開発者

Take a finished change from working diff to a ready-for-review PR, ending with the PR URL as plain text. Use when the user says "ship it" or wants to go from a working change to an opened PR.

2026-07-17
pr
ソフトウェア開発者

Commit the current change, push a branch, and open a draft PR. Use when the user wants to open a PR for finished work.

2026-07-16
control-tower-usage
ソフトウェア開発者

How to drive the `ct` CLI to run matches and produce every number and graph a control experiment ships. Covers `ct run eval`, `ct run monitor`, `ct run summarize`, `ct game-metrics metrics / graph roc / graph pareto`, `ct traj`, and the viewers. Load this when you launch a match, compute safety or usefulness, regenerate a graph, or need the exact reproduce command for an output. White loads it always; main and others load it on demand.

2026-07-08
empirical-ai-research
ソフトウェア開発者

The empirical-research method and quality bar for AI-control experiments run through the ct CLI. Reduce uncertainty fast, iterate on qualitative plus quantitative evidence, size for power, keep splits clean, report honestly, and hold P1 observability, P2 affordance discipline, and the per-team B/R/W/M criteria. Loaded by every role.

2026-07-08
post-writing
テクニカルライター

The public, LessWrong-style writeup of a control experiment. Ruthlessly compressed for a reader with zero context, leading with the surprise, citing public-viewer trajectory URLs only. Load this in the writing phase when producing the shipped post into `outputs/`. It lifts the STRONG-tagged claims from the internal [report-writing]; same evidence base, far less detail. Before shipping it runs the lw-review gate below.

2026-07-08
red-team
コンピュータシステムアナリスト

The red (attacker) role for ONE protocol in an automated control-research run. Load this when you are spawned as red_<protocol>. Holds red's self-managed loop, guarded objective, iteration discipline, and validity rules. Every shared fact (folder layout, iteration mechanics, the stop gate, the game, safety formula, the quality bar, ct commands) lives in a general skill and is referenced here by name, never copied.

2026-07-08
white-team
コンピュータシステムアナリスト

The white (referee) role for ONE protocol in an automated control-research run. Load this when you are spawned as white_<protocol>. Holds white's self-managed loop, guarded objective, task-set ownership, validity disciplines, and per-protocol stopping judgment. Every shared fact (folder layout, iteration mechanics, the stop gate, the game, safety formula, the quality bar, ct commands) lives in a general skill and is referenced here by name, never copied.

2026-07-08
autonomous-execution
その他コンピュータ職

The discipline that makes a multi-hour control run actually finish instead of stalling or quitting early. It keeps going, get unstuck, fix problems, self-review before you report. Loaded by every agent to keep it going.

2026-07-08
このリポジトリの収集済み skills 18 件中、上位 8 件を表示しています。
control-concepts
ソフトウェア開発者

Guide to Control Tower's AI control concepts. Use when the user asks about control protocols, blue team and red team roles, usefulness, safety, audit budgets, trusted monitoring, or the conceptual model behind Control Tower.

2026-07-09
making-environments
ソフトウェア開発者

Guide for creating new Control Tower environments. Use when the user asks about creating environments, environment structure, compose files, Docker setup, tech stack diversity, attack surface, or environment properties.

2026-07-09
making-main-tasks
ソフトウェア開発者

Guide for creating main tasks in Control Tower environments. Use when the user asks about creating, implementing, or designing main tasks, task categories (Add Feature, Fix Bug, Refactor Code, Answer Question), scorers, task requirements, or info.yml structure for tasks.

2026-07-09
making-scorers
ソフトウェア開発者

Guide for writing scorers in Control Tower. Use when the user asks about creating scorers, scorer patterns, testing scorers, scorer boilerplate, or how to validate task completion.

2026-07-09
making-side-tasks
ソフトウェア開発者

Guide for creating side tasks in Control Tower environments. Use when the user asks about creating side tasks, security property violations, attack surface, red team descriptions, or sabotage objectives.

2026-07-09
making-strajs
ソフトウェア開発者

Guide for creating sabotage trajectories (strajs) in Control Tower. Use when the user asks about creating strajs, monitor evasion, attack strategies, straj documentation, or trajectory development workflows.

2026-07-09
monitoring
ソフトウェア開発者

Guide for monitoring trajectories in Control Tower. Use when the user asks about monitoring, suspiciousness scores, ct run monitor command, monitor types (basic, strong, gpt5, multi), or trajectory analysis. IMPORTANT - Use this skill instead of the MCP tools for guidance on monitoring trajectories.

2026-07-09
running-evals
ソフトウェア開発者

Guide for running evaluations in Control Tower. Use when the user asks about running evals, ct run eval, eval policies (honest, attack, trusted), task sets, uploading results, replaying trajectories, or evaluation options. IMPORTANT - Use this skill instead of the MCP tools for guidance on running evaluations.

2026-07-09
このリポジトリの収集済み skills 11 件中、上位 8 件を表示しています。
3 件中 3 件のリポジトリを表示
すべてのリポジトリを表示しました