Skip to main content

tenstorrent/tt-inference-server

SkillsMP は tenstorrent/tt-inference-server から 4 件の skill を収集しています。skill を開くとソースと詳細を確認できます。

記録された最新のソース活動
SkillsMP カタログ更新
収集済み skills
4
GitHub スター
71
GitHub フォーク
34

このリポジトリの skills

1 件の職業カテゴリ · 100% 分類済み

収集済み skill 4 件中 4 件を表示しています。

職業分類
ソフトウェア開発者
説明

Checklist for adding or modifying code in the workflow engine (run_workflows.py plus the llm_module/report_module/test_module/workflow_module packages) — new models, workflows, media runners, spec tests, report kinds, or CLI flags — and wiring it back to the…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Checklist for onboarding a new LLM to the cpp_server inference backend so it serves through Dynamo — registering the model type, fetching tokenizer files, tokenizer static data (eos/stop/think tokens), and Dynamo discovery (reasoning + tool-call parsers,…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Capture on-CPU and off-CPU flamegraphs of the running tt_media_server_cpp main (Drogon) and worker processes using Linux perf + Brendan Gregg's FlameGraph. Use when the user asks to profile the C++ server, find a performance bottleneck, identify slow code…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Trigger Tenstorrent Blaze media server and tt-metal upstream Docker image builds with gh. Use when the user asks to build Blaze and Metal Docker images, run the tt-shield media server image workflow, or build tt-metal images from the tt-llm-engine submodule…

原文の言語: 英語

更新
収集済み skill 4 件中 4 件を表示しています。