Skip to main content
xinyangli
Perfil de creador de GitHub

xinyangli

Vista por repositorio de 21 skills recopiladas en 1 repositorios de GitHub.

skills recopiladas
21
repositorios
1
actualizado
23 ago 2026
mapa de repositorios

Dónde viven las skills

Repositorios principales por número de skills recopiladas, con su participación en este catálogo del creador y su variedad ocupacional.

explorador de repositorios

Repositorios y skills representativas

batch-annotation-retry-recovery
sin clasificar

Recover and retry failed records from Gemini batch annotation jobs. Covers grid-level retry (no re-concat), full re-concat retry pipelines, and after concat — how to upload manifest to S3, split shard ranges across teammates, run with --use-concat,…

23 ago 2026
canva-model-weights-deploy
sin clasificar

Update Canva ML model weights through W&B registry upload, artifact mirroring, model lockfile PRs, and staged deployment PRs. Use when updating ingredient-generation/media-transformation model checkpoints, running arnold registry upload, following Canva…

23 ago 2026
dataset-parquet-packaging
sin clasificar

Package image datasets into bucketed Parquet shards with resolution and aspect ratio bucketing. Use when converting JSONL+image datasets to Parquet format, creating training-ready datasets, or when the user mentions parquet, packaging, bucketing, sharding, or…

23 ago 2026
dataset-pipeline-verifiers
sin clasificar

Run verifier gates for Core CN dataset processing stages: S3 counts, parquet schema, row alignment, metadata naming, bbox validity, OCR/caption validity, resized image audits, and final sidecar quality.

23 ago 2026
dataset-processing-pipeline
sin clasificar

Orchestrate Core CN image dataset processing from raw assets through CPU filtering, dedup, main/sidecar Parquet packaging, 512/1024 resized Parquet, layer detection, HunyuanOCR, JSON captioning, score enrichment, and verifier gates. Use when planning or…

23 ago 2026
distributed-concat-pipeline
sin clasificar

Build 2x2 grid concat images from S3 image datasets for batch Gemini captioning. Use when creating concat grids, running the annotation batch infer pipeline, generating concat manifests, managing the prepare→concat→worker→run→postprocess workflow, or…

23 ago 2026
fixed-resolution-sidecar
sin clasificar

Build 512-area and 1024-area resized image Parquet datasets, preserve row keys, audit valid/invalid rows, and feed crop metadata back into structured_description bbox columns.

23 ago 2026
fuck-yubikey
sin clasificar

Reduce YubiKey touch frequency on Canva Linux devboxes. Diagnoses where touches come from (tsh/kubectl, Python kubernetes clients like utp, AWS prod profile) and applies workarounds. Use when the user says "yubikey 触发太频繁", "kubectl 每次要 touch", "aws cli 一直要…

23 ago 2026
Mostrando 8 de 21 skills recopiladas.
Mostrando 1 de 1 repositorios
Todos los repositorios cargados