Skip to main content
uiuc-kang-lab
ملف منشئ GitHub

uiuc-kang-lab

عرض على مستوى المستودعات لـ ٢٤ skills مجمعة عبر ١ مستودعات GitHub.

skills مجمعة
٢٤
مستودعات
١
محدث
١٢ مايو ٢٠٢٦
خريطة المستودعات

أين توجد skills

أهم المستودعات حسب عدد skills المجمعة، مع حصتها من كتالوج هذا المنشئ وانتشارها المهني.

مستكشف المستودعات

المستودعات و skills الممثلة

checkpoints
مطوّرو البرمجيات

Guide for checkpointing — saving, loading, and resuming training with CheckpointRecord. Use when the user asks about saving weights, resuming training, checkpoint management, or the checkpoint lifecycle.

١٢ مايو ٢٠٢٦
ci
محللو ضمان جودة البرمجيات والمختبرون

Guide for testing conventions and CI pipelines — unit tests, integration smoke tests, pytest markers, and GitHub Actions workflows. Use when the user asks about testing, CI, running tests, or adding tests for a recipe.

١٢ مايو ٢٠٢٦
completers
مطوّرو البرمجيات

Guide for using completers — TokenCompleter and MessageCompleter for text generation during RL rollouts and evaluation. Use when the user asks about generating text, completing messages, or using completers in RL environments.

١٢ مايو ٢٠٢٦
contributing
مطوّرو البرمجيات

Guide for contributing to the tinker-cookbook repo — development setup, code style, type checking, PR process, and design conventions. Use when the user asks about how to contribute, set up the dev environment, code style, or project conventions.

١٢ مايو ٢٠٢٦
datasets
مطوّرو البرمجيات

Guide for dataset construction — SupervisedDatasetBuilder, RLDatasetBuilder, ChatDatasetBuilder, and custom dataset creation from JSONL, HuggingFace, or conversation data. Use when the user asks about datasets, data loading, data preparation, or custom data…

١٢ مايو ٢٠٢٦
distillation
مطوّرو البرمجيات

Set up and run knowledge distillation (on-policy, off-policy, or multi-teacher) from a teacher model to a student model using the Tinker API. Use when the user wants to distill knowledge, compress models, or train a student from a teacher.

١٢ مايو ٢٠٢٦
dpo
علماء البيانات

Set up and run Direct Preference Optimization (DPO) training on preference datasets using the Tinker API. Use when the user wants to train with preference data, chosen/rejected pairs, or DPO.

١٢ مايو ٢٠٢٦
environments
علماء البيانات

Guide for defining RL environments — the Env protocol, EnvGroupBuilder, RLDataset, and custom environment creation. Use when the user asks about RL environments, reward functions, or how to define custom tasks for RL training.

١٢ مايو ٢٠٢٦
عرض 8 من أصل ٢٤ skills مجمعة.
عرض ١ من أصل ١ مستودعات
تم تحميل كل المستودعات