Skip to main content

tenstorrent/tt-inference-server

SkillsMP는 tenstorrent/tt-inference-server에서 4개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
4
GitHub 스타
71
GitHub 포크
34

이 저장소의 skills

직업 카테고리 1개 · 100% 분류됨

수집된 skill 4개 중 4개를 표시합니다.

직업 분류
소프트웨어 개발자
설명

Checklist for adding or modifying code in the workflow engine (run_workflows.py plus the llm_module/report_module/test_module/workflow_module packages) — new models, workflows, media runners, spec tests, report kinds, or CLI flags — and wiring it back to the…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Checklist for onboarding a new LLM to the cpp_server inference backend so it serves through Dynamo — registering the model type, fetching tokenizer files, tokenizer static data (eos/stop/think tokens), and Dynamo discovery (reasoning + tool-call parsers,…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Capture on-CPU and off-CPU flamegraphs of the running tt_media_server_cpp main (Drogon) and worker processes using Linux perf + Brendan Gregg's FlameGraph. Use when the user asks to profile the C++ server, find a performance bottleneck, identify slow code…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Trigger Tenstorrent Blaze media server and tt-metal upstream Docker image builds with gh. Use when the user asks to build Blaze and Metal Docker images, run the tt-shield media server image workflow, or build tt-metal images from the tt-llm-engine submodule…

원문 언어: 영어

업데이트
수집된 skill 4개 중 4개를 표시합니다.