Skip to main content
braintrustdata
ملف منشئ GitHub

braintrustdata

عرض على مستوى المستودعات لـ ٦٣ skills مجمعة عبر ١٤ مستودعات GitHub.

skills مجمعة
٦٣
مستودعات
١٤
محدث
٢٧ أغسطس ٢٠٢٦
خريطة المستودعات

أين توجد skills

أهم المستودعات حسب عدد skills المجمعة، مع حصتها من كتالوج هذا المنشئ وانتشارها المهني.

#01
eval-library
٢٤ skills · ١٧ أغسطس ٢٠٢٦
التصنيف قيد الانتظار
٣٨٪؜الحصة
#02
braintrust-coding-agent-plugins
١٢ skills · ٢٧ أغسطس ٢٠٢٦
مطوّرو البرمجياتمحللو ضمان جودة البرمجيات والمختبرون
٢ فئات مهنية · ٩٢٪؜ مصنفة
١٩٪؜الحصة
#03
lingua
٧ skills · ٣١ يوليو ٢٠٢٦
محللو ضمان جودة البرمجيات والمختبرونمطوّرو البرمجيات
٢ فئات مهنية · ١٠٠٪؜ مصنفة
١١٪؜الحصة
#04
braintrust-sdk-python
٦ skills · ٣٠ يوليو ٢٠٢٦
مطوّرو البرمجياتمحللو ضمان جودة البرمجيات والمختبرون
٢ فئات مهنية · ١٠٠٪؜ مصنفة
٩٫٥٪؜الحصة
#05
braintrust-sdk-javascript
٣ skills · ٤ أغسطس ٢٠٢٦
محللو ضمان جودة البرمجيات والمختبرونمطوّرو البرمجيات
٢ فئات مهنية · ١٠٠٪؜ مصنفة
٤٫٨٪؜الحصة
#06
agentbehavior
٢ skills · ٢١ يوليو ٢٠٢٦
المحامونالمهن الحاسوبية الأخرى
٢ فئات مهنية · ١٠٠٪؜ مصنفة
٣٫٢٪؜الحصة
#07
braintrust-sdk-ruby
٢ skills · ٦ يناير ٢٠٢٦
مطوّرو البرمجيات
١ فئات مهنية · ١٠٠٪؜ مصنفة
٣٫٢٪؜الحصة
#08
braintrust-claude-plugin
١ skills · ٢ فبراير ٢٠٢٦
متخصصو دعم شبكات الحاسوب
١ فئات مهنية · ١٠٠٪؜ مصنفة
١٫٦٪؜الحصة
نعرض هنا أهم 8 مستودعات؛ تستمر القائمة الكاملة أدناه.
مستكشف المستودعات

المستودعات و skills الممثلة

braintrust-analyze-eval-experiment
غير مصنف

Analyze completed LLM or agent eval experiments using uncertainty-aware and decision-relevant methods. Use to audit run completeness and pairing, calculate confidence intervals, run paired comparisons, report wins, losses, and ties, incorporate run-to-run…

١٧ أغسطس ٢٠٢٦
braintrust-attribute-multi-variable-change
غير مصنف

Attribute an observed change when several things moved at once — model plus prompt plus tools, a provider migration, a framework upgrade, or a vendor swap that bundles serving stack with model. Use when asked which part of a change caused the result, when a…

١٧ أغسطس ٢٠٢٦
braintrust-build-eval-dataset
غير مصنف

Create, edit, audit, or compare eval datasets for LLM applications and agents, including target-population definition, case sourcing from production traces, stratified sampling, label provenance and label audits, expected values as constraints for open-ended…

١٧ أغسطس ٢٠٢٦
braintrust-define-eval-objective
غير مصنف

Create, edit, or audit an eval objective by working backward from a product decision to the target outcome, construct, population, intended claim, and verification-versus-validation questions. Use when a team is unsure what an eval should establish, asks…

١٧ أغسطس ٢٠٢٦
braintrust-define-eval-release-gate
غير مصنف

Create, edit, audit, or apply release gates for LLM applications and agents. Use to combine minimum meaningful improvement, statistical significance, regression rate, subgroup consistency, worst-run stability, all-attempts reliability, safety upper bounds,…

١٧ أغسطس ٢٠٢٦
braintrust-deploy-evaluator
غير مصنف

Take a validated scorer or classifier from definition to running instrument in Braintrust — scope selection, inline testing before saving, saving as an evaluator, attaching an online-scoring rule, activating it for new traffic, and backfilling history with a…

١٧ أغسطس ٢٠٢٦
braintrust-design-eval-experiment
غير مصنف

Design or audit controlled eval experiments for model, prompt, retrieval, tool, guardrail, or agent-architecture changes. Use before data collection to state directional and minimum-effect hypotheses, name independent, dependent, and control variables…

١٧ أغسطس ٢٠٢٦
braintrust-design-eval-instrumentation
غير مصنف

Design the trace and eval-dataset schema for an LLM app or agent, and wire the system to emit it. Use when deciding what to log, designing a trace schema, setting up tracing or observability before evals, or when failures cannot be debugged or sliced from…

١٧ أغسطس ٢٠٢٦
عرض 8 من أصل ٢٤ skills مجمعة.
troubleshoot-braintrust-mcp
غير مصنف

This plugin auto-configures a "braintrust" MCP server. If you can't see it or reach it, activate this skill

٢٧ أغسطس ٢٠٢٦
add-coding-agent-capture
مطوّرو البرمجيات

Add or review live coding-agent event capture through blocking command hooks or an in-process plugin such as a JavaScript adapter. Use when forwarding native lifecycle, model, tool, permission, or subagent events into the Braintrust daemon, or when fixing…

١٠ أغسطس ٢٠٢٦
add-coding-agent-import
مطوّرو البرمجيات

Add or review historical transcript import and live attach support for a coding agent, including session lookup, native parsing, synthetic lifecycle envelopes, shared translator use, incremental tailing, destination overrides, and import tests. Use when…

١٠ أغسطس ٢٠٢٦
add-coding-agent-integration
مطوّرو البرمجيات

Orchestrate a complete Braintrust coding-agent integration by delegating feasibility, daemon translation, event capture, setup, managed run, transcript import, verification, and shipping to the specialized repo-local skills. Use for end-to-end support for a…

١٠ أغسطس ٢٠٢٦
add-coding-agent-run
مطوّرو البرمجيات

Add or review invocation-local managed-run support for a coding agent, including executable dispatch, temporary hook or plugin injection, route isolation, duplicate-capture suppression, process behavior, trust safety, and run-command tests. Use when…

١٠ أغسطس ٢٠٢٦
add-coding-agent-setup
مطوّرو البرمجيات

Add or review persistent setup for a coding-agent tracing integration, including public CLI exposure, plugin or hook installation, non-secret route configuration, idempotent updates, disablement, and isolated setup tests. Use when implementing or fixing the…

١٠ أغسطس ٢٠٢٦
add-coding-agent-translator
مطوّرو البرمجيات

Implement or review a coding agent's daemon translator, including source registration, native event correlation, trace shape, deterministic recovery, routing metadata, and translator tests. Use when adding a new agent translator or changing how an existing…

١٠ أغسطس ٢٠٢٦
ship-coding-agent-integration
مطوّرو البرمجيات

Package and release a coding-agent tracing integration through its marketplace, package registry, or distribution repository, including manifests, builds, validation, versioning, CI, documentation, publishing automation, and installed-artifact smoke tests.…

١٠ أغسطس ٢٠٢٦
عرض 8 من أصل ١٢ skills مجمعة.
provider-type-update
مطوّرو البرمجيات

Audit, plan, implement, and verify updates to OpenAI, Anthropic, or Google generated provider types. Use when a provider specification or generated.rs diff may require capability, adapter, universal type, serializer, streaming, payload-case, or test changes.

٣١ يوليو ٢٠٢٦
add-provider
مطوّرو البرمجيات

Add support for a new LLM provider format to Lingua. Follow the test-first workflow with payload snapshots.

٢٠ أبريل ٢٠٢٦
coverage-report
محللو ضمان جودة البرمجيات والمختبرون

Run transformation coverage tests. Use compact mode when iterating on fixes, full mode when planning or documenting bugs.

٢٠ أبريل ٢٠٢٦
testing
محللو ضمان جودة البرمجيات والمختبرون

Lingua's testing tools for transformation coverage and end-to-end SDK validation.

٢٠ أبريل ٢٠٢٦
add-provider
مطوّرو البرمجيات

Add support for a new LLM provider format to Lingua. Follow the test-first workflow with payload snapshots.

١٨ فبراير ٢٠٢٦
coverage-report
محللو ضمان جودة البرمجيات والمختبرون

Run transformation coverage tests. Use compact mode when iterating on fixes, full mode when planning or documenting bugs.

١٨ فبراير ٢٠٢٦
testing
محللو ضمان جودة البرمجيات والمختبرون

Lingua's testing tools for transformation coverage and end-to-end SDK validation.

١٨ فبراير ٢٠٢٦
sdk-integrations
مطوّرو البرمجيات

Create or update Braintrust Python SDK integrations built on the integrations API under `py/src/braintrust/integrations/`. Use when adding a new integration package, extending an existing provider integration, changing patchers, tracing, manual `wrap_*()`…

٣٠ يوليو ٢٠٢٦
sdk-dependency-updates
مطوّرو البرمجيات

Review and refresh Braintrust Python SDK dependency update PRs, especially automated `chore(deps): weekly dependency update` PRs. Use when Codex needs to inspect `py/pyproject.toml` and `py/uv.lock`, reproduce the workflow's label decision, decide whether…

١ مايو ٢٠٢٦
sdk-benchmarking
مطوّرو البرمجيات

Run, compare, and extend Braintrust Python SDK pyperf benchmarks. Use when touching hot-path code in `py/src/braintrust/` such as serialization, deep-copy, span creation, or logging; when adding or updating files under `py/benchmarks/`; or when you need…

١٥ أبريل ٢٠٢٦
sdk-ci-triage
مطوّرو البرمجيات

Triage and reproduce Braintrust Python SDK CI failures. Use when asked why CI failed, to fix broken CI on a PR, to inspect a failing GitHub Actions job, or to map a failing matrix job back to the exact local nox session, provider version, or workflow step…

١٥ أبريل ٢٠٢٦
sdk-vcr-workflows
محللو ضمان جودة البرمجيات والمختبرون

Work with Braintrust Python SDK VCR and cassette-backed tests. Use when adding or updating cassette-backed provider tests, deciding whether to re-record cassettes, debugging VCR failures, converting mock-heavy coverage to VCR, or handling cassette hygiene for…

١٥ أبريل ٢٠٢٦
commit-message
مطوّرو البرمجيات

Suggest a Braintrust SDK repo-style commit message from the current diff and conversation. Use when asked to write, suggest, or generate a commit message for the current changes.

١٣ أبريل ٢٠٢٦
عرض ١٢ من أصل ١٤ مستودعات