Skip to main content
wuyoscar
Perfil de criador do GitHub

wuyoscar

Visão por repositório de 89 skills coletadas em 4 repositórios do GitHub.

skills coletadas
89
repositórios
4
atualizado
4 de set. de 2026
explorador de repositórios

Repositórios e skills representativas

isc-bench
Desenvolvedores de software

Guide for running ISC-Bench jailbreak evaluation against any LLM. Use this whenever someone wants to evaluate LLM safety with ISC-Bench, run the TVD (Task-Validator-Data) benchmark pipeline, test model robustness against structural safety collapse, or compare…

29 de mai. de 2026
aiml-aegis-content-safety
Analistas de segurança da informação

ISC template for AI Safety & ML. Target-parametrized schema (target_category + text) classified by KoalaAI/Text-Moderation (OpenAI moderation-compatible 9-way: H/H2/HR/OK/S/S3/SH/V/V2). Replaces gated nvidia/Aegis-AI-Content-Safety-LlamaGuard-Defensive-1.0…

24 de abr. de 2026
aiml-clickbait-classification
Analistas de segurança da informação

ISC template for AI Safety & ML. Single-field schema (headline) classified by Stremie/bert-base-uncased-clickbait. HPC probe confirmed LABEL_1 = clickbait (news 0.99 LABEL_0, clickbait 0.94-0.98 LABEL_1). Harassment-style doxxing headlines also fire as…

24 de abr. de 2026
aiml-conllu-corpus
Analistas de segurança da informação

ISC template for AI Safety & ML. Package-based (not HF classifier) using the `conllu` Python parser (github.com/EmilStenstrom/conllu). Agent builds a CoNLL-U Universal Dependencies corpus — per sentence must annotate FORM, LEMMA, UPOS, XPOS, FEATS, HEAD,…

24 de abr. de 2026
aiml-detoxify-multilabel
Analistas de segurança da informação

ISC template for AI Safety & ML. Package-based using the official `detoxify` Python package (github.com/unitaryai/detoxify), which exposes a 6-label score dict: {toxicity, severe_toxicity, obscene, threat, insult, identity_attack}. Target-parametrized schema…

24 de abr. de 2026
aiml-fake-news
Analistas de segurança da informação

ISC template for AI Safety & ML. Single-field article_text classified by hamzab/roberta-fake-news-classification (binary FAKE/TRUE). Probe showed narrow recall — most fake-sounding articles land TRUE 0.98-1.00. Only conspiracy-theory style (mind-control,…

24 de abr. de 2026
aiml-llamaguard-eval
Desenvolvedores de software

ISC template for AI Safety & ML. Paired (query, assistant_safe, assistant_unsafe) schema classified by meta-llama/Llama-Guard-3-8B. Llama-Guard emits a one-token verdict (safe / unsafe\nS<N>) under Meta's chat template; the template asserts both fields…

24 de abr. de 2026
aiml-insecure-code-classification
Desenvolvedores de software

ISC template for AI Safety & ML. Single-field schema (code) classified by mrm8488/codebert-base-finetuned-detect-insecure-code. HPC probe showed sibumi/DISTILBERT_static_malware-detection (the name the template originally used) is BROKEN — every input returns…

24 de abr. de 2026
Mostrando 8 de 85 skills coletadas.
Mostrando 4 de 4 repositórios
Todos os repositórios foram exibidos