Skip to main content

fetchlib

A compliance-first page-fetch waterfall for the repo's scraping skills: tiered backends that escalate only when the current tier is genuinely blocked — L1 curl_cffi (TLS/JA3 impersonation, no JS) → L2 r.jina.ai (JS render to clean markdown) → L3 pluggable nodriver/browser → L4 opt-in paid unblocker — wrapped in api-pacer + AIMD per-domain rate control, a per-domain circuit breaker, JSONL fetch telemetry (url/tier/status/blocked/bytes/ms), robots.txt honoring, and optional Thompson-sampling backend learning (learn=True) because the best tier varies by site AND by IP. Use when the user wants to fetch or scrape pages without getting 403'd, needs one shared fetch layer instead of ad-hoc curl in every script, or asks which fetch method to use for a specific hard site and whether it's time to pay for an unblocker. Triggers on "fetchlib", "抓取被封了", "403 抓不到", "换个方式抓这个站", "统一抓取层", "要不要上付费代理/unblocker", "curl_cffi 还是 Jina", "抓竞品页面老失败", "scrape without getting blocked", "fetch waterfall", "which fetcher for this site".

跳到安装

来源信息

仓库
noique/cross-border-ecommerce-skills
最近来源活动
2026年7月26日 17:01
检测到的 SKILL.md 语言
中文
星标
38
分支
6

安装方式

默认使用会先检查来源的 Prompt;你也可以切换为直接命令,或下载本地副本。

检查来源文件

决定是否安装前,请先阅读 SKILL.md,以及 SkillsMP 当前展示的配套文件。