Skip to main content

wanshuiyin/Anti-Autoresearch

SkillsMP는 wanshuiyin/Anti-Autoresearch에서 12개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
12
GitHub 스타
149
GitHub 포크
8

이 저장소의 skills

직업 카테고리 2개 · 100% 분류됨

수집된 skill 12개 중 12개를 표시합니다.

직업 분류
기타 중등 후 교사
설명

End-to-end substantive-integrity forensic sweep of a research paper (especially autoresearch / AI-Scientist-style output). Orchestrates the whole pipeline: ingest (arxiv-id | pdf | dir → working dir + pdftotext for L0) → /evidence-ledger (artifact manifest +…

원문 언어: 영어

업데이트
직업 분류
기타 중등 후 교사
설명

Synthesize the single strongest EVIDENCE-BOUND reviewer case to reject a paper, built ONLY from the evidence ledger (claims.json) + the other auditors' confirmed findings — never free-floating LLM critique. Two fresh cross-model codex threads: an attack…

원문 언어: 영어

업데이트
직업 분류
기타 중등 후 교사
설명

Transparent, itemized impressions of AI-generated WRITING STYLE — the repo's ONLY non-integrity track. Two passes: a deterministic defensive-hedge density screen (tools/check_ai_style.py, AIS-DEFENSIVE-HEDGE) plus a fresh cross-model GROSS-cases-only semantic…

원문 언어: 영어

업데이트
직업 분류
기타 중등 후 교사
설명

Audit whether a paper's baseline comparisons are COMPLETE, FAIR, and SIGNIFICANT: a required recent SOTA baseline is missing while 'best/SOTA' is claimed (HP-MISSING-BASELINE); a baseline is undertuned / given less compute-tuning-data, run at a mismatched…

원문 언어: 영어

업데이트
직업 분류
기타 중등 후 교사
설명

Citation-integrity forensics: is every reference real, correctly attributed, and used in a context the cited work actually supports? Catches hallucinated references (no paper at the claimed arXiv id/DOI/venue, fabricated authors/year), metadata drift (wrong…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Flagship intra-paper self-consistency forensics: does the paper contradict ITSELF across abstract/intro/tables/body/appendix, and does the method DESCRIBED match the method EVALUATED? Needs no external ground truth — works PDF-only (L0). Runs a deterministic…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Audit whether a paper's EVALUATION DESIGN actually measures what it claims and whether its reporting is complete — the validity layer family D (experiment-forensics) cannot reach. Three patterns: train/test leakage means the reported score may not measure…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Build the deterministic evidence ledger (artifact_manifest.json + claims.json) that every other Anti-Autoresearch auditor reads. One pass inventories artifacts, derives the observability level (L0 PDF-only / L1 +LaTeX / L2 +repo+results) by fixed rule, and…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Audit experiment integrity against the evidence ledger. At L2 (repo + result files present) a fresh cross-model reviewer reads the eval code line-by-line for fake/derived ground truth, score self-normalization, phantom results (a paper number with no backing…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

MEMO-ONLY prior-work overlap advisory: surfaces the two ADVISORY taxonomy signals neither a tool nor a model can decide from the paper alone — ADV-TRIVIAL-COMBINATION (standard A+B+C / 缝合 stapling) and ADV-DUPLICATE-PUBLICATION (repackaged / duplicate…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Checkable-ish surface presentation signals a reviewer notices first — duplicate/near-identical tables, leftover pipeline/template strings, too-few or LLM-looking figures, and page-padding. AUXILIARY ONLY and weak by design: a deterministic pass…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Family-G proof & derivation integrity forensics: does a THIRD PARTY's written proof/derivation actually establish its theorem, or does it skip an obligation, assume its own conclusion, take an invalid step, drift a symbol's meaning, or smuggle an unstated…

원문 언어: 영어

업데이트
수집된 skill 12개 중 12개를 표시합니다.