Skip to main content
sahel-sh
ملف منشئ GitHub

sahel-sh

عرض على مستوى المستودعات لـ ٥ skills مجمعة عبر ١ مستودعات GitHub.

skills مجمعة
٥
مستودعات
١
محدث
١ يوليو ٢٠٢٦
خريطة المستودعات

أين توجد skills

أهم المستودعات حسب عدد skills المجمعة، مع حصتها من كتالوج هذا المنشئ وانتشارها المهني.

مستكشف المستودعات

المستودعات و skills الممثلة

evaluate-and-etc
علماء البيانات

Evaluate a DeepHone deep-research run with the gpt-oss-120b LLM-as-judge (accuracy / recall / calibration), aggregate token usage, and compute the Effective Token Cost (ETC) plus the paper's accuracy-vs-cost figures (Figures 1-3). Use after producing run…

١ يوليو ٢٠٢٦
extend-searcher-reranker
مطوّرو البرمجيات

Add a custom retriever (searcher) or reranker to DeepHone by implementing the BaseSearcher / BaseReranker interface and registering it in the SearcherType / RerankerType enum. Use when integrating a new retrieval or reranking method into the benchmark.

١ يوليو ٢٠٢٦
run-reranking-experiment
علماء البيانات

Run the paper's core experiments — one-shot reranking effectiveness (Table 1) and end-to-end deep-research with listwise or cross-encoder reranking at depth d in {10,20,50} and search reasoning in {low,medium,high} (Tables 2/4, Figures 2/3). Use to produce…

١ يوليو ٢٠٢٦
serve-vllm-models
مطوّرو البرمجيات

Launch the vLLM servers needed for DeepHone experiments — the gpt-oss search agent (with tool calling), the reranker (gpt-oss listwise or Qwen3-Reranker-0.6B cross-encoder), and the gpt-oss-120b LLM-as-judge. Use before running or evaluating deep-research…

١ يوليو ٢٠٢٦
setup-benchmark
مطوّرو البرمجيات

Set up the DeepHone / BrowseComp-Plus benchmark — install the uv environment (Python 3.10, Java 21, flash-attn), decrypt the dataset, and download or build the BM25 / Qwen3-Embedding-8B retrieval indexes. Use this before running any experiment or evaluation.

١ يوليو ٢٠٢٦
عرض ١ من أصل ١ مستودعات
تم تحميل كل المستودعات