Skip to main content
uiuc-kang-lab
Perfil de creador de GitHub

uiuc-kang-lab

Vista por repositorio de 24 skills recopiladas en 1 repositorios de GitHub.

skills recopiladas
24
repositorios
1
actualizado
12 may 2026
mapa de repositorios

Dónde viven las skills

Repositorios principales por número de skills recopiladas, con su participación en este catálogo del creador y su variedad ocupacional.

explorador de repositorios

Repositorios y skills representativas

checkpoints
Desarrolladores de software

Guide for checkpointing — saving, loading, and resuming training with CheckpointRecord. Use when the user asks about saving weights, resuming training, checkpoint management, or the checkpoint lifecycle.

12 may 2026
ci
Analistas de garantía de calidad de software y probadores

Guide for testing conventions and CI pipelines — unit tests, integration smoke tests, pytest markers, and GitHub Actions workflows. Use when the user asks about testing, CI, running tests, or adding tests for a recipe.

12 may 2026
completers
Desarrolladores de software

Guide for using completers — TokenCompleter and MessageCompleter for text generation during RL rollouts and evaluation. Use when the user asks about generating text, completing messages, or using completers in RL environments.

12 may 2026
contributing
Desarrolladores de software

Guide for contributing to the tinker-cookbook repo — development setup, code style, type checking, PR process, and design conventions. Use when the user asks about how to contribute, set up the dev environment, code style, or project conventions.

12 may 2026
datasets
Desarrolladores de software

Guide for dataset construction — SupervisedDatasetBuilder, RLDatasetBuilder, ChatDatasetBuilder, and custom dataset creation from JSONL, HuggingFace, or conversation data. Use when the user asks about datasets, data loading, data preparation, or custom data…

12 may 2026
distillation
Desarrolladores de software

Set up and run knowledge distillation (on-policy, off-policy, or multi-teacher) from a teacher model to a student model using the Tinker API. Use when the user wants to distill knowledge, compress models, or train a student from a teacher.

12 may 2026
dpo
Científicos de datos

Set up and run Direct Preference Optimization (DPO) training on preference datasets using the Tinker API. Use when the user wants to train with preference data, chosen/rejected pairs, or DPO.

12 may 2026
environments
Científicos de datos

Guide for defining RL environments — the Env protocol, EnvGroupBuilder, RLDataset, and custom environment creation. Use when the user asks about RL environments, reward functions, or how to define custom tasks for RL training.

12 may 2026
Mostrando 8 de 24 skills recopiladas.
Mostrando 1 de 1 repositorios
Todos los repositorios cargados