Skip to main content

hebrew-ml-datasets-navigator

Sterne11
Forks4
Aktualisiert18. Juni 2026 um 21:12

Navigate the fragmented landscape of Hebrew and Yiddish ML datasets and models. Covers ivrit.ai (22K+ hours of Hebrew audio, whisper-large-v3 ASR variants, Yiddish models), Dicta (DictaLM 3.0 LLM family, DictaBERT variants, HeQ reading comprehension), the Israeli National NLP Program / NNLP-IL (HebrewSentiment, HebNLI), AlephBERT, and Knesset Plenums. Helps researchers and ML engineers pick the right dataset for a task by use case, license (commercial vs research), Hebrew register coverage, and model-dataset pairing. Use when choosing training data for a Hebrew NLP or ASR project, verifying license compatibility for a commercial product, finding a baseline model for a Hebrew downstream task, or exploring Yiddish ML resources. Do NOT use for Arabic NLP datasets (a separate ecosystem), general HuggingFace dataset discovery (use HuggingFace Hub search), or Hebrew OCR dataset selection (use hebrew-ocr-forms). Activate for: ื“ืื˜ื”ืกื˜ ืขื‘ืจื™ืช, ืžืื’ืจ ื ืชื•ื ื™ื, ื ืชื•ื ื™ ืื™ืžื•ืŸ, ืžื•ื“ืœ ืขื‘ืจื™ืช, ืจื™ืฉื™ื•ืŸ ืžืกื—ืจื™, ืงื•ืจืคื•ืก ืขื‘ืจื™ืช, ื–ื™ื”ื•ื™ ื“ื™ื‘ื•ืจ

Installation

Mit Codex oder Claude installieren Kopieren Sie diesen Prompt, fรผgen Sie ihn in Codex, Claude oder einen anderen Assistant ein und lassen Sie die Skill-Seite prรผfen und installieren.

Datei-Explorer
10 Dateien
SKILL.md
readonly