Skip to main content

sahel-sh/DeepHone

SkillsMP는 sahel-sh/DeepHone에서 5개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
5
GitHub 스타
4
GitHub 포크
0

이 저장소의 skills

직업 카테고리 2개 · 100% 분류됨

수집된 skill 5개 중 5개를 표시합니다.

직업 분류
데이터 과학자
설명

Evaluate a DeepHone deep-research run with the gpt-oss-120b LLM-as-judge (accuracy / recall / calibration), aggregate token usage, and compute the Effective Token Cost (ETC) plus the paper's accuracy-vs-cost figures (Figures 1-3). Use after producing run…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Add a custom retriever (searcher) or reranker to DeepHone by implementing the BaseSearcher / BaseReranker interface and registering it in the SearcherType / RerankerType enum. Use when integrating a new retrieval or reranking method into the benchmark.

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Run the paper's core experiments — one-shot reranking effectiveness (Table 1) and end-to-end deep-research with listwise or cross-encoder reranking at depth d in {10,20,50} and search reasoning in {low,medium,high} (Tables 2/4, Figures 2/3). Use to produce…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Launch the vLLM servers needed for DeepHone experiments — the gpt-oss search agent (with tool calling), the reranker (gpt-oss listwise or Qwen3-Reranker-0.6B cross-encoder), and the gpt-oss-120b LLM-as-judge. Use before running or evaluating deep-research…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Set up the DeepHone / BrowseComp-Plus benchmark — install the uv environment (Python 3.10, Java 21, flash-attn), decrypt the dataset, and download or build the BM25 / Qwen3-Embedding-8B retrieval indexes. Use this before running any experiment or evaluation.

원문 언어: 영어

업데이트
수집된 skill 5개 중 5개를 표시합니다.