Skip to main content
wuyoscar
GitHub 제작자 프로필

wuyoscar

4개 GitHub 저장소에서 수집된 89개 skills를 저장소 단위로 보여줍니다.

수집된 skills
89
저장소
4
업데이트
2026년 9월 4일
저장소 탐색

저장소와 대표 skills

isc-bench
소프트웨어 개발자

Guide for running ISC-Bench jailbreak evaluation against any LLM. Use this whenever someone wants to evaluate LLM safety with ISC-Bench, run the TVD (Task-Validator-Data) benchmark pipeline, test model robustness against structural safety collapse, or compare…

2026년 5월 29일
aiml-aegis-content-safety
정보 보안 분석가

ISC template for AI Safety & ML. Target-parametrized schema (target_category + text) classified by KoalaAI/Text-Moderation (OpenAI moderation-compatible 9-way: H/H2/HR/OK/S/S3/SH/V/V2). Replaces gated nvidia/Aegis-AI-Content-Safety-LlamaGuard-Defensive-1.0…

2026년 4월 24일
aiml-clickbait-classification
정보 보안 분석가

ISC template for AI Safety & ML. Single-field schema (headline) classified by Stremie/bert-base-uncased-clickbait. HPC probe confirmed LABEL_1 = clickbait (news 0.99 LABEL_0, clickbait 0.94-0.98 LABEL_1). Harassment-style doxxing headlines also fire as…

2026년 4월 24일
aiml-conllu-corpus
정보 보안 분석가

ISC template for AI Safety & ML. Package-based (not HF classifier) using the `conllu` Python parser (github.com/EmilStenstrom/conllu). Agent builds a CoNLL-U Universal Dependencies corpus — per sentence must annotate FORM, LEMMA, UPOS, XPOS, FEATS, HEAD,…

2026년 4월 24일
aiml-detoxify-multilabel
정보 보안 분석가

ISC template for AI Safety & ML. Package-based using the official `detoxify` Python package (github.com/unitaryai/detoxify), which exposes a 6-label score dict: {toxicity, severe_toxicity, obscene, threat, insult, identity_attack}. Target-parametrized schema…

2026년 4월 24일
aiml-fake-news
정보 보안 분석가

ISC template for AI Safety & ML. Single-field article_text classified by hamzab/roberta-fake-news-classification (binary FAKE/TRUE). Probe showed narrow recall — most fake-sounding articles land TRUE 0.98-1.00. Only conspiracy-theory style (mind-control,…

2026년 4월 24일
aiml-llamaguard-eval
소프트웨어 개발자

ISC template for AI Safety & ML. Paired (query, assistant_safe, assistant_unsafe) schema classified by meta-llama/Llama-Guard-3-8B. Llama-Guard emits a one-token verdict (safe / unsafe\nS<N>) under Meta's chat template; the template asserts both fields…

2026년 4월 24일
aiml-insecure-code-classification
소프트웨어 개발자

ISC template for AI Safety & ML. Single-field schema (code) classified by mrm8488/codebert-base-finetuned-detect-insecure-code. HPC probe showed sibumi/DISTILBERT_static_malware-detection (the name the template originally used) is BROKEN — every input returns…

2026년 4월 24일
수집된 skill 85개 중 8개를 표시합니다.
저장소 4개 중 4개 표시
모든 저장소를 표시했습니다