Skip to main content

jmagly/aiwg-training

SkillsMP는 jmagly/aiwg-training에서 15개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
15
GitHub 스타
1
GitHub 포크
1

이 저장소의 skills

직업 카테고리 2개 · 100% 분류됨

수집된 skill 15개 중 15개를 표시합니다.

직업 분류
데이터 과학자
설명

Generate Datasheet, Model Card, and Data Statement from a dataset manifest

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Deterministically rebuild a dataset from its manifest and verify fixity equivalence

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Create a versioned training dataset with manifest, fixity, provenance, and archive snapshot

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

End-to-end training dataset pipeline — acquire sources through publication

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Detect training-eval overlap against benchmark sets before dataset publication

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Generate SFT training examples from raw sources using Self-Instruct / Evol-Instruct / SQuAD / STaR patterns

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Convert canonical training examples to Alpaca format for training frameworks

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Convert canonical training examples to ChatML format for training frameworks

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Convert canonical training examples to JSONL format for training frameworks

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Convert canonical training examples to Parquet format for training frameworks

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Convert canonical training examples to ShareGPT format for training frameworks

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Generate preference pairs (chosen/rejected) for DPO/KTO/ORPO/SimPO training

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Large-scale synthetic training data generation with Model Collapse recursion guard

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Acquire a training data source with license validation and delegate ingest to the semantic memory kernel

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

GRADE quality assessment adapted for individual training examples

원문 언어: 영어

업데이트
수집된 skill 15개 중 15개를 표시합니다.