Skip to main content
wuyoscar
GitHub-Creator-Profil

wuyoscar

Repository-Ansicht von 89 gesammelten Skills in 4 GitHub-Repositories.

gesammelte Skills
89
Repositories
4
aktualisiert
9. Sept. 2026
Repository-Explorer

Repositories und repräsentative Skills

isc-bench
Softwareentwickler

Guide for running ISC-Bench jailbreak evaluation against any LLM. Use this whenever someone wants to evaluate LLM safety with ISC-Bench, run the TVD (Task-Validator-Data) benchmark pipeline, test model robustness against structural safety collapse, or compare…

29. Mai 2026
aiml-aegis-content-safety
Informationssicherheitsanalysten

ISC template for AI Safety & ML. Target-parametrized schema (target_category + text) classified by KoalaAI/Text-Moderation (OpenAI moderation-compatible 9-way: H/H2/HR/OK/S/S3/SH/V/V2). Replaces gated nvidia/Aegis-AI-Content-Safety-LlamaGuard-Defensive-1.0…

24. Apr. 2026
aiml-clickbait-classification
Informationssicherheitsanalysten

ISC template for AI Safety & ML. Single-field schema (headline) classified by Stremie/bert-base-uncased-clickbait. HPC probe confirmed LABEL_1 = clickbait (news 0.99 LABEL_0, clickbait 0.94-0.98 LABEL_1). Harassment-style doxxing headlines also fire as…

24. Apr. 2026
aiml-conllu-corpus
Informationssicherheitsanalysten

ISC template for AI Safety & ML. Package-based (not HF classifier) using the `conllu` Python parser (github.com/EmilStenstrom/conllu). Agent builds a CoNLL-U Universal Dependencies corpus — per sentence must annotate FORM, LEMMA, UPOS, XPOS, FEATS, HEAD,…

24. Apr. 2026
aiml-detoxify-multilabel
Informationssicherheitsanalysten

ISC template for AI Safety & ML. Package-based using the official `detoxify` Python package (github.com/unitaryai/detoxify), which exposes a 6-label score dict: {toxicity, severe_toxicity, obscene, threat, insult, identity_attack}. Target-parametrized schema…

24. Apr. 2026
aiml-fake-news
Informationssicherheitsanalysten

ISC template for AI Safety & ML. Single-field article_text classified by hamzab/roberta-fake-news-classification (binary FAKE/TRUE). Probe showed narrow recall — most fake-sounding articles land TRUE 0.98-1.00. Only conspiracy-theory style (mind-control,…

24. Apr. 2026
aiml-llamaguard-eval
Softwareentwickler

ISC template for AI Safety & ML. Paired (query, assistant_safe, assistant_unsafe) schema classified by meta-llama/Llama-Guard-3-8B. Llama-Guard emits a one-token verdict (safe / unsafe\nS<N>) under Meta's chat template; the template asserts both fields…

24. Apr. 2026
aiml-insecure-code-classification
Softwareentwickler

ISC template for AI Safety & ML. Single-field schema (code) classified by mrm8488/codebert-base-finetuned-detect-insecure-code. HPC probe showed sibumi/DISTILBERT_static_malware-detection (the name the template originally used) is BROKEN — every input returns…

24. Apr. 2026
Es werden 8 von 85 gesammelten Skills angezeigt.
4 von 4 Repositories angezeigt
Alle Repositories angezeigt