Skip to main content
uiuc-kang-lab
Perfil de criador do GitHub

uiuc-kang-lab

Visão por repositório de 24 skills coletadas em 1 repositórios do GitHub.

skills coletadas
24
repositórios
1
atualizado
12 de mai. de 2026
mapa de repositórios

Onde as skills estão

Principais repositórios por número de skills coletadas, com sua participação neste catálogo do criador e sua distribuição ocupacional.

explorador de repositórios

Repositórios e skills representativas

checkpoints
Desenvolvedores de software

Guide for checkpointing — saving, loading, and resuming training with CheckpointRecord. Use when the user asks about saving weights, resuming training, checkpoint management, or the checkpoint lifecycle.

12 de mai. de 2026
ci
Analistas de garantia de qualidade de software e testadores

Guide for testing conventions and CI pipelines — unit tests, integration smoke tests, pytest markers, and GitHub Actions workflows. Use when the user asks about testing, CI, running tests, or adding tests for a recipe.

12 de mai. de 2026
completers
Desenvolvedores de software

Guide for using completers — TokenCompleter and MessageCompleter for text generation during RL rollouts and evaluation. Use when the user asks about generating text, completing messages, or using completers in RL environments.

12 de mai. de 2026
contributing
Desenvolvedores de software

Guide for contributing to the tinker-cookbook repo — development setup, code style, type checking, PR process, and design conventions. Use when the user asks about how to contribute, set up the dev environment, code style, or project conventions.

12 de mai. de 2026
datasets
Desenvolvedores de software

Guide for dataset construction — SupervisedDatasetBuilder, RLDatasetBuilder, ChatDatasetBuilder, and custom dataset creation from JSONL, HuggingFace, or conversation data. Use when the user asks about datasets, data loading, data preparation, or custom data…

12 de mai. de 2026
distillation
Desenvolvedores de software

Set up and run knowledge distillation (on-policy, off-policy, or multi-teacher) from a teacher model to a student model using the Tinker API. Use when the user wants to distill knowledge, compress models, or train a student from a teacher.

12 de mai. de 2026
dpo
Cientistas de dados

Set up and run Direct Preference Optimization (DPO) training on preference datasets using the Tinker API. Use when the user wants to train with preference data, chosen/rejected pairs, or DPO.

12 de mai. de 2026
environments
Cientistas de dados

Guide for defining RL environments — the Env protocol, EnvGroupBuilder, RLDataset, and custom environment creation. Use when the user asks about RL environments, reward functions, or how to define custom tasks for RL training.

12 de mai. de 2026
Mostrando 8 de 24 skills coletadas.
Mostrando 1 de 1 repositórios
Todos os repositórios foram exibidos