Skip to main content
wuyoscar
GitHub クリエイタープロフィール

wuyoscar

5 件の GitHub リポジトリにある 94 件の収集済み skills をリポジトリ単位で表示します。

収集済み skills
94
リポジトリ
5
更新
2026年9月23日
リポジトリエクスプローラー

リポジトリと代表的な skills

isc-bench
ソフトウェア開発者

Guide for running ISC-Bench jailbreak evaluation against any LLM. Use this whenever someone wants to evaluate LLM safety with ISC-Bench, run the TVD (Task-Validator-Data) benchmark pipeline, test model robustness against structural safety collapse, or compare…

2026年5月29日
aiml-aegis-content-safety
情報セキュリティアナリスト

ISC template for AI Safety & ML. Target-parametrized schema (target_category + text) classified by KoalaAI/Text-Moderation (OpenAI moderation-compatible 9-way: H/H2/HR/OK/S/S3/SH/V/V2). Replaces gated nvidia/Aegis-AI-Content-Safety-LlamaGuard-Defensive-1.0…

2026年4月24日
aiml-clickbait-classification
情報セキュリティアナリスト

ISC template for AI Safety & ML. Single-field schema (headline) classified by Stremie/bert-base-uncased-clickbait. HPC probe confirmed LABEL_1 = clickbait (news 0.99 LABEL_0, clickbait 0.94-0.98 LABEL_1). Harassment-style doxxing headlines also fire as…

2026年4月24日
aiml-conllu-corpus
情報セキュリティアナリスト

ISC template for AI Safety & ML. Package-based (not HF classifier) using the `conllu` Python parser (github.com/EmilStenstrom/conllu). Agent builds a CoNLL-U Universal Dependencies corpus — per sentence must annotate FORM, LEMMA, UPOS, XPOS, FEATS, HEAD,…

2026年4月24日
aiml-detoxify-multilabel
情報セキュリティアナリスト

ISC template for AI Safety & ML. Package-based using the official `detoxify` Python package (github.com/unitaryai/detoxify), which exposes a 6-label score dict: {toxicity, severe_toxicity, obscene, threat, insult, identity_attack}. Target-parametrized schema…

2026年4月24日
aiml-fake-news
情報セキュリティアナリスト

ISC template for AI Safety & ML. Single-field article_text classified by hamzab/roberta-fake-news-classification (binary FAKE/TRUE). Probe showed narrow recall — most fake-sounding articles land TRUE 0.98-1.00. Only conspiracy-theory style (mind-control,…

2026年4月24日
aiml-llamaguard-eval
ソフトウェア開発者

ISC template for AI Safety & ML. Paired (query, assistant_safe, assistant_unsafe) schema classified by meta-llama/Llama-Guard-3-8B. Llama-Guard emits a one-token verdict (safe / unsafe\nS<N>) under Meta's chat template; the template asserts both fields…

2026年4月24日
aiml-insecure-code-classification
ソフトウェア開発者

ISC template for AI Safety & ML. Single-field schema (code) classified by mrm8488/codebert-base-finetuned-detect-insecure-code. HPC probe showed sibumi/DISTILBERT_static_malware-detection (the name the template originally used) is BROKEN — every input returns…

2026年4月24日
収集済み skill 85 件中 8 件を表示しています。
jev
プロジェクト管理専門家

Design Jev-assisted workflows using collected use cases, references and examples. Adapt and combine patterns, then use Jev for classification, scoring or candidate selection when useful, including batch judgments and agent checkpoints.

2026年9月23日
jev-act
プロジェクト管理専門家

Choose one legal next action in a browser, desktop, game or simulation. Supply fresh observed state and available actions. The host or simulator executes and checks the result; selection does not grant permission.

2026年9月22日
jev-documents
ソフトウェア開発者

Locate, select, extract and verify evidence in documents or observed code inventories. Use for source-span extraction, passage reranking, claim checks or choosing code locations to inspect. Preserve citations and no-match outcomes; use required graph tools…

2026年9月22日
jev-eval
ソフトウェア品質保証アナリスト・テスター

Judge supplied outputs against explicit criteria, including code-change reviews and authorized safety evaluations. Use for evidence-backed review leads, rubric judgments, or batch, multi-turn and team transcript evaluation; not permission to merge or run…

2026年9月22日
jev-triage
プロジェクト管理専門家

Use for user-defined inbox, support-ticket, feedback or record classification and prioritization, especially bulk parallel judgments with sufficient per-record context. Accepts smoke_test to have the host agent write and validate a task-specific paired pilot…

2026年9月22日
5 件中 5 件のリポジトリを表示
すべてのリポジトリを表示しました