Skip to main content
wuyoscar
Profil créateur GitHub

wuyoscar

Vue par dépôt de 89 skills collectés dans 4 dépôts GitHub.

skills collectés
89
dépôts
4
mis à jour
4 sept. 2026
explorateur de dépôts

Dépôts et skills représentatifs

isc-bench
Développeurs de logiciels

Guide for running ISC-Bench jailbreak evaluation against any LLM. Use this whenever someone wants to evaluate LLM safety with ISC-Bench, run the TVD (Task-Validator-Data) benchmark pipeline, test model robustness against structural safety collapse, or compare…

29 mai 2026
aiml-aegis-content-safety
Analystes en sécurité de l'information

ISC template for AI Safety & ML. Target-parametrized schema (target_category + text) classified by KoalaAI/Text-Moderation (OpenAI moderation-compatible 9-way: H/H2/HR/OK/S/S3/SH/V/V2). Replaces gated nvidia/Aegis-AI-Content-Safety-LlamaGuard-Defensive-1.0…

24 avr. 2026
aiml-clickbait-classification
Analystes en sécurité de l'information

ISC template for AI Safety & ML. Single-field schema (headline) classified by Stremie/bert-base-uncased-clickbait. HPC probe confirmed LABEL_1 = clickbait (news 0.99 LABEL_0, clickbait 0.94-0.98 LABEL_1). Harassment-style doxxing headlines also fire as…

24 avr. 2026
aiml-conllu-corpus
Analystes en sécurité de l'information

ISC template for AI Safety & ML. Package-based (not HF classifier) using the `conllu` Python parser (github.com/EmilStenstrom/conllu). Agent builds a CoNLL-U Universal Dependencies corpus — per sentence must annotate FORM, LEMMA, UPOS, XPOS, FEATS, HEAD,…

24 avr. 2026
aiml-detoxify-multilabel
Analystes en sécurité de l'information

ISC template for AI Safety & ML. Package-based using the official `detoxify` Python package (github.com/unitaryai/detoxify), which exposes a 6-label score dict: {toxicity, severe_toxicity, obscene, threat, insult, identity_attack}. Target-parametrized schema…

24 avr. 2026
aiml-fake-news
Analystes en sécurité de l'information

ISC template for AI Safety & ML. Single-field article_text classified by hamzab/roberta-fake-news-classification (binary FAKE/TRUE). Probe showed narrow recall — most fake-sounding articles land TRUE 0.98-1.00. Only conspiracy-theory style (mind-control,…

24 avr. 2026
aiml-llamaguard-eval
Développeurs de logiciels

ISC template for AI Safety & ML. Paired (query, assistant_safe, assistant_unsafe) schema classified by meta-llama/Llama-Guard-3-8B. Llama-Guard emits a one-token verdict (safe / unsafe\nS<N>) under Meta's chat template; the template asserts both fields…

24 avr. 2026
aiml-insecure-code-classification
Développeurs de logiciels

ISC template for AI Safety & ML. Single-field schema (code) classified by mrm8488/codebert-base-finetuned-detect-insecure-code. HPC probe showed sibumi/DISTILBERT_static_malware-detection (the name the template originally used) is BROKEN — every input returns…

24 avr. 2026
Affichage de 8 skills collectés sur 85.
4 dépôts affichés sur 4
Tous les dépôts sont affichés