Skip to main content

awslabs/llm-evaluation-system

SkillsMP は awslabs/llm-evaluation-system から 3 件の skill を収集しています。skill を開くとソースと詳細を確認できます。

記録された最新のソース活動
SkillsMP カタログ更新
収集済み skills
3
GitHub スター
22
GitHub フォーク
3

このリポジトリの skills

1 件の職業カテゴリ · 100% 分類済み

収集済み skill 3 件中 3 件を表示しています。

職業分類
ソフトウェア開発者
説明

Port an existing third-party benchmark or eval into this repo as an Inspect AI task, faithfully. Use this whenever the task is "add benchmark X", "port this eval", "can we run <github repo> here", or adapting any external eval harness (aiewf-eval,…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Use when the user asks to create, update, or render an AWS architecture diagram (or a cloud/infrastructure diagram using AWS service icons) — e.g. "diagram our AWS setup", "architecture diagram with CloudFront/S3/EKS/RDS", "update docs/image.png", "make an…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Ship work in the llm-evaluation-system repo end-to-end — commit with conventional-commit messages, push to a feature branch (never directly to main), open a PR with the proper title format, and after merge either run `make release` to publish to PyPI or just…

原文の言語: 英語

更新
収集済み skill 3 件中 3 件を表示しています。