Skip to main content

pp-scrape-creators

Every Scrape Creators endpoint across 28 platforms, with credit-aware comment mining and a local corpus no other Scrape Creators tool has. Trigger phrases: `find which platforms a creator is on`, `pull the comments and replies from this post`, `monitor a brand's ads`, `search creator transcripts for a keyword`, `how many credits would this sweep cost`, `use scrape creators`, `run scrape-creators`.

الانتقال إلى التثبيت

معلومات المصدر

المستودع
mvanhorn/printing-press-library
آخر نشاط في المصدر
١٣ أغسطس ٢٠٢٦ في ٠٤:٥٥
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
١٬٩١٨
التفرعات
٥٧٢

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.

عرض SKILL.md

SKILL.md
تعليمات المصدر · معاينة للقراءة فقط
name
pp-scrape-creators
description
Every Scrape Creators endpoint across 28 platforms, with credit-aware comment mining and a local corpus no other Scrape Creators tool has. Trigger phrases: `find which platforms a creator is on`, `pull the comments and replies from this post`, `monitor a brand's ads`, `search creator transcripts for a keyword`, `how many credits would this sweep cost`, `use scrape creators`, `run scrape-creators`.
author
Adrian Horning
license
Apache-2.0
argument-hint
<command> [args] | install cli|mcp
allowed-tools
Read Bash
metadata
{"openclaw":{"requires":{"bins":"[Truncated]"},"install":["[Truncated]"]}}
<!-- GENERATED FILE — DO NOT EDIT. This file is a verbatim mirror of library/developer-tools/scrape-creators/SKILL.md, regenerated post-merge by tools/generate-skills/. Hand-edits here are silently overwritten on the next regen. Edit the library/ source instead. See the repository agent guide, section "Generated artifacts: registry.json, cli-skills/". --> # Scrape Creators — Printing Press CLI ## Prerequisites: Install the CLI This skill drives the `scrape-creators-pp-cli` binary. **You must verify the CLI is installed before invoking any command from this skill.** If it is missing, install it first: 1. Install via the Printing Press installer. It defaults binaries to `$HOME/.local/bin` on macOS/Linux and `%LOCALAPPDATA%\Programs\PrintingPress\bin` on Windows: ```bash npx -y @mvanhorn/printing-press-library install scrape-creators --cli-only ``` 2. Verify: `scrape-creators-pp-cli --version` 3. Ensure the reported install directory is on `$PATH` for the agent/runtime that will invoke this skill. If the `npx` install fails (no Node, offline, etc.), fall back to a direct Go install (requires Go 1.26.5 or newer). This installs into `$GOPATH/bin` (default `$HOME/go/bin`), so add that directory to `$PATH` instead: ```bash go install github.com/mvanhorn/printing-press-library/library/developer-tools/scrape-creators/cmd/scrape-creators-pp-cli@latest ``` If `--version` reports "command not found" after install, the runtime cannot see the binary directory on `$PATH`. Do not proceed with skill commands until verification succeeds. The official CLI mirrors endpoints and the official skills describe curl workflows; neither remembers anything between runs. This CLI syncs profiles, posts, comments with their replies, transcripts, and ads into SQLite with FTS5 search, routes comment-thread fetches on credit economics (comments thread), audits reply completeness against ground truth (comments coverage), and gates expensive sweeps behind a pre-flight credit estimate (account estimate). ## When to Use This CLI Use this CLI when a task touches public social-media data at scale: mining comments and replies for a brand, qualifying creators across platforms, monitoring competitor ads, or searching transcript/comment corpora you have already synced. It is the right choice whenever credit economics matter — its thread routing, sweep budgets, and pre-flight estimates exist so agents never spend blind. ## Anti-triggers Do not use this CLI for: - Do not use this CLI to post, like, follow, or message on any platform — it is read-only public-data scraping - Do not use it for private/logged-in-only content; it sees what the public sees - Do not use it as a general web scraper for non-social sites; use a crawling tool instead ## Unique Capabilities These capabilities aren't available in any other tool for this API. ### Comment-thread completeness - **`comments thread`** — Fetch one post's complete comment threads, automatically picking the cheaper route between the 15-credit flat include_replies call and 1-credit per-comment reply calls (don't trust child_comment_count to decide: it's unreliable). _Reach for this when you need every reply on a post without doing credit arithmetic by hand._ ```bash scrape-creators-pp-cli comments thread https://www.instagram.com/reel/C8rKmYvsrck --agent ``` - **`comments coverage`** — Rank synced posts by how many comments the API reported versus how many actually landed in your local store — ground truth where the API's child_comment_count is unreliable as a thread filter. _Reach for this after a sweep to find which posts are silently missing their replies._ ```bash scrape-creators-pp-cli comments coverage bracken.design --agent ``` - **`comments sweep`** — Pull recent posts for a handle and their comments in one command, stopping cleanly at a credit budget you set. _The one-command version of a multi-hundred-post comment-mining ritual, budget-gated._ ```bash scrape-creators-pp-cli comments sweep bracken.design --since 7d --max-credits 200 --agent ``` ### Credit governance - **`account estimate`** — Project the credit cost of a planned run against your live balance and exit non-zero if it would exhaust the budget. _Run this before any bulk sweep so an agent never burns the balance mid-pipeline._ ```bash scrape-creators-pp-cli account estimate --posts 950 --with-replies flat --agent ``` - **`account budget`** — See how fast you're spending API credits and how many days remain at the current pace. _Check runway before committing to a new recurring pipeline._ ```bash scrape-creators-pp-cli account budget --agent ``` ### Cross-platform intelligence - **`creator find`** — Given one handle, see which of 12 creator platforms the creator is on with follower counts side-by-side. _Start any collab qualification here before pulling per-platform detail._ ```bash scrape-creators-pp-cli creator find mkbhd --agent ``` - **`creator compare`** — Compare two or more creators side-by-side on follower count, engagement rate, and content volume. _Strip vanity follower counts out of a collab decision._ ```bash scrape-creators-pp-cli creator compare mkbhd mrwhosetheboss --agent ``` - **`content spikes`** — Surface the videos that performed far above a creator's own baseline — the ones that actually went viral. _Find outlier content without eyeballing hundreds of posts._ ```bash scrape-creators-pp-cli content spikes mkbhd --platform youtube ``` - **`trends triangulate`** — Snapshot a hashtag or topic across platforms in one call to see which platform it is biggest on. _Decide where to publish before creating the content._ ```bash scrape-creators-pp-cli trends triangulate "matcha" --agent ``` ### Local state that compounds - **`transcripts search`** — FTS5 full-text search across every platform transcript you've synced — nine resource types spanning YouTube, TikTok, Instagram, Facebook, LinkedIn, Rumble, and more. _Search transcripts you already paid for instead of re-fetching them._ ```bash scrape-creators-pp-cli transcripts search "pricing objection" --limit 10 ``` - **`ads monitor`** — Snapshot a brand's live ads across Facebook, TikTok, Google, and LinkedIn ad libraries; on rerun, diff new ads versus ones that disappeared. _Rerun weekly and read only the delta of a competitor's ad activity._ ```bash scrape-creators-pp-cli ads monitor nike --agent ``` - **`comments search`** — Full-text search across every synced comment and reply, offline. _Mine questions and complaints from comments you already pulled without spending credits._ ```bash scrape-creators-pp-cli comments search "refund" --limit 20 ``` - **`creator track`** — Append a follower snapshot per run on a chosen platform, then read the growth trajectory over time. _Track a partner's growth on a schedule you control._ ```bash scrape-creators-pp-cli creator track mkbhd --platform instagram ``` - **`creator tagged`** — Snapshot the posts a creator or brand is tagged in and diff new mentions on rerun. _Weekly UGC check for a client brand without re-reading the full list._ ```bash scrape-creators-pp-cli creator tagged bracken.design --agent ``` ## Command Reference **account** — Manage account - `scrape-creators-pp-cli account list` — Returns the number of API credits remaining on your Scrape Creators account. - `scrape-creators-pp-cli account list-getapiusage` — Returns a paginated list of your API requests, including the endpoint called, status code, credits used, and timestamp. - `scrape-creators-pp-cli account list-getdailyusagecount` — Returns aggregated daily usage statistics for the last 30 days - `scrape-creators-pp-cli account list-getmostusedroutes` — Returns your top 20 most called API endpoints ranked by call count, along with total credits consumed per endpoint. **amazon** — Manage amazon - `scrape-creators-pp-cli amazon` — Scrapes a creator's Amazon Shop page by URL, returning their storefront profile and product collections. **apple-music** — Scrape Apple Music artists, songs, albums, and search results - `scrape-creators-pp-cli apple-music list` — Retrieves public Apple Music album details, including title, artist, artwork, release info, tracks - `scrape-creators-pp-cli apple-music list-applemusic` — Retrieves public Apple Music artist details, including artwork, editorial notes, top songs, albums, music videos - `scrape-creators-pp-cli apple-music list-applemusic-2` — Searches Apple Music and returns public result sections for artists, albums, songs, playlists, stations - `scrape-creators-pp-cli apple-music list-applemusic-3` — Retrieves public Apple Music song details by id or URL. Album track URLs with an i= song id are supported. **bluesky** — Get Bluesky posts and profile info - `scrape-creators-pp-cli bluesky list` — Fetches a single Bluesky post by URL, returning the post's record text, author info, embed content, replyCount - `scrape-creators-pp-cli bluesky list-profile` — Retrieves a Bluesky user's public profile including handle, displayName, avatar, description, followersCount - `scrape-creators-pp-cli bluesky list-user` — Fetches a paginated feed of posts from a Bluesky user, returning each post's uri, record text, author info **detect-age-gender** — Manage detect age gender - `scrape-creators-pp-cli detect-age-gender` — Uses AI to analyze a creator's profile photo and estimate their age and gender. **facebook** — Get public Facebook profiles and posts - `scrape-creators-pp-cli facebook create` — Fetches all ads currently running for a specific company from the Meta Ad Library. - `scrape-creators-pp-cli facebook create-adlibrary` — Searches the Meta Ad Library by keyword and returns matching ads. - `scrape-creators-pp-cli facebook list` — Get the events of a city. Check out this [link](https://www.facebook. - `scrape-creators-pp-cli facebook list-adlibrary` — Retrieves detailed information about a specific Facebook ad by its ID or URL. - `scrape-creators-pp-cli facebook list-adlibrary-2` — Retrieves a transcript for a single Facebook Ad Library video ad by ID or URL. - `scrape-creators-pp-cli facebook list-adlibrary-3` — Fetches all ads currently running for a specific company from the Meta Ad Library. - `scrape-creators-pp-cli facebook list-adlibrary-4` — Searches the Meta Ad Library by keyword and returns matching ads. - `scrape-creators-pp-cli facebook list-adlibrary-5` — Searches for companies by name in the Meta Ad Library and returns their page IDs for use with other ad library - `scrape-creators-pp-cli facebook list-event` — Get a specific event by its URL or id - `scrape-creators-pp-cli facebook list-events` — Search for events by name. - `scrape-creators-pp-cli facebook list-group` — Fetches the public information shown on a Facebook group's About page, including its description, privacy and visibility - `scrape-creators-pp-cli facebook list-group-2` — Fetches posts from a public Facebook group, limited to 3 posts per page due to API limitations. - `scrape-creators-pp-cli facebook list-marketplace` — Fetches details for a Facebook Marketplace item by item id or Marketplace item URL, including title, description, price - `scrape-creators-pp-cli facebook list-marketplace-2` — Searches Facebook Marketplace listings by keyword and lat/lng. Supports pagination with the returned cursor. - `scrape-creators-pp-cli facebook list-marketplace-3` — Searches Facebook Marketplace locations/cities and returns coordinates you can use with the Marketplace Search endpoint. - `scrape-creators-pp-cli facebook list-post` — Retrieves a single public Facebook post or reel by URL. - `scrape-creators-pp-cli facebook list-post-2` — Fetches comments from a Facebook post or reel with cursor-based pagination. - `scrape-creators-pp-cli facebook list-post-3` — Extracts the transcript text from a Facebook video post or reel. - `scrape-creators-pp-cli facebook list-post-4` — Get the replies to a comment. - `scrape-creators-pp-cli facebook list-profile` — Retrieves public Facebook page details including category, address, email, phone, website, services, priceRange, rating - `scrape-creators-pp-cli facebook list-profile-2` — Get the events of a public Facebook page - `scrape-creators-pp-cli facebook list-profile-3` — Fetches photos from a public Facebook page with pagination support. - `scrape-creators-pp-cli facebook list-profile-4` — Returns publicly visible Facebook profile posts, limited to 3 posts per page due to API limitations. - `scrape-creators-pp-cli facebook list-profile-5` — Fetches up to 10 reels per request from a public Facebook page. **github** — Scrape GitHub profiles, repositories, and public activity - `scrape-creators-pp-cli github list` — Retrieves public metadata for one GitHub repository, including owner, description, language, stars, forks, topics - `scrape-creators-pp-cli github list-trending` — Scrapes GitHub's public Trending developers page. - `scrape-creators-pp-cli github list-trending-2` — Scrapes GitHub's public Trending repositories page. - `scrape-creators-pp-cli github list-user` — Retrieves public GitHub user details including name, bio, avatar, company, location, blog, follower counts - `scrape-creators-pp-cli github list-user-2` — Retrieves GitHub profile contribution activity for a user from the public profile activity timeline. - `scrape-creators-pp-cli github list-user-3` — Retrieves the public GitHub contribution graph for a user and year - `scrape-creators-pp-cli github list-user-4` — Retrieves public GitHub followers for a user. Each follower includes login, avatar, user URL, type, and GitHub IDs. - `scrape-creators-pp-cli github list-user-5` — Retrieves public accounts followed by a GitHub user. - `scrape-creators-pp-cli github list-user-6` — Searches public GitHub pull requests authored by a user using GitHub's public search index. - `scrape-creators-pp-cli github list-user-7` — Retrieves a user's public repositories with repo metadata like description, language, stars, forks, topics, license **google** — Scrape Google search results - `scrape-creators-pp-cli google list` — Retrieves detailed information about a specific Google ad including advertiserId, creativeId, format, firstShown - `scrape-creators-pp-cli google list-adlibrary` — Searches the Google Ad Transparency Library for advertisers by name. - `scrape-creators-pp-cli google list-company` — Fetches public ads for a company from the Google Ad Transparency Library by domain or advertiser_id. - `scrape-creators-pp-cli google list-search` — Performs a Google search and returns organic results with url, title, and description for each result. **instagram** — Gets Instagram profiles, posts, and reels - `scrape-creators-pp-cli instagram list` — Fetches a lightweight Instagram profile summary by user ID, returning username, full name, biography - `scrape-creators-pp-cli instagram list-audio` — Fetches the reels Instagram exposes for an audio page like instagram.com/reels/audio/{audio_id}/. - `scrape-creators-pp-cli instagram list-media` — Generates an AI-powered speech-to-text transcription for an Instagram video post or reel. - `scrape-creators-pp-cli instagram list-post` — Fetches detailed metadata for a single Instagram post or reel by shortcode or URL. - `scrape-creators-pp-cli instagram list-post-2` — Retrieves comments on a public Instagram post or reel. - `scrape-creators-pp-cli instagram list-post-3` — Retrieves the public replies to a specific Instagram comment. - `scrape-creators-pp-cli instagram list-profile` — Retrieves public Instagram profile information including biography, bio links - `scrape-creators-pp-cli instagram list-reels` — Fetches trending reels from Instagram's public instagram.com/reels page. - `scrape-creators-pp-cli instagram list-reels-2` — Use this when you only want Google-indexed Instagram reels matching a keyword or phrase - `scrape-creators-pp-cli instagram list-search` — Use this for Instagram-native account, hashtag, or place lookup. - `scrape-creators-pp-cli instagram list-search-2` — Use this when you know the exact hashtag and want Google-indexed public Instagram posts or reels, optional date filters - `scrape-creators-pp-cli instagram list-search-3` — Use this to explore an Instagram topic and the posts Instagram curates for it. - `scrape-creators-pp-cli instagram list-search-4` — Use this for broad creator discovery from keywords found in Google-indexed Instagram profile pages, bios - `scrape-creators-pp-cli instagram list-user` — Returns the raw HTML embed snippet for an Instagram user's profile widget. - `scrape-creators-pp-cli instagram list-user-2` — Lists all story highlight albums for an Instagram user. - `scrape-creators-pp-cli instagram list-user-3` — Returns a paginated list of a user's public Instagram reels (short-form videos). - `scrape-creators-pp-cli instagram list-user-4` — Returns up to 10 public posts per page from an Instagram user's Tagged tab. - `scrape-creators-pp-cli instagram list-user-5` — Returns a paginated feed of a user's public Instagram posts, including reels, photos, videos, and carousels. - `scrape-creators-pp-cli instagram list-user-6` — Fetches the full contents of a specific Instagram story highlight album by its ID. **kick** — Scrape Kick clips - `scrape-creators-pp-cli kick` — Fetches detailed data for a Kick clip by URL, including video, metadata, and channel info. **komi** — Scrape Komi pages - `scrape-creators-pp-cli komi` — Scrapes a Komi page by URL, extracting the creator's profile, social links, and featured content. **kwai** — Scrape Kwai profiles, posts, and user feeds - `scrape-creators-pp-cli kwai list` — Fetches public Kwai post details including caption, media URLs, cover images, counts, author info, and music metadata. - `scrape-creators-pp-cli kwai list-profile` — Fetches public Kwai profile data including username, bio, avatar, verification status, gender, and public counts. - `scrape-creators-pp-cli kwai list-user` — Fetches a paginated list of public Kwai posts for a user, including captions, media URLs, covers, counts, author info **linkbio** — Scrape Linkbio (lnk.bio) pages - `scrape-creators-pp-cli linkbio` — Scrapes a Linkbio (lnk.bio) page by URL, extracting the creator's profile and all their links. **linkedin** — Scrape LinkedIn - `scrape-creators-pp-cli linkedin list` — Retrieves detailed information about a specific LinkedIn ad by URL. - `scrape-creators-pp-cli linkedin list-ads` — Searches the LinkedIn Ad Library by company name, keyword, or companyId with optional country and date filters. - `scrape-creators-pp-cli linkedin list-company` — Fetches a LinkedIn company page with details including name, description, logo, cover image, slogan, location - `scrape-creators-pp-cli linkedin list-company-2` — Retrieves paginated posts from a LinkedIn company page, including each post's URL, ID, publication date - `scrape-creators-pp-cli linkedin list-post` — Fetches a single LinkedIn post or article, returning the title, headline, full description text - `scrape-creators-pp-cli linkedin list-post-2` — Fetches the transcript from a LinkedIn post video when LinkedIn exposes one publicly. - `scrape-creators-pp-cli linkedin list-profile` — Retrieves a person's public LinkedIn profile data, including their name, photo, location, follower count (followers) - `scrape-creators-pp-cli linkedin list-search` — Finds public LinkedIn posts, feed updates, and Pulse articles by keyword using Google Search **linkme** — Get Linkme profile info - `scrape-creators-pp-cli linkme` — Retrieves a Linkme profile by URL, including identity, social links, and contact details. **linktree** — Scrape Linktree pages - `scrape-creators-pp-cli linktree` — Scrapes a Linktree page by URL, extracting the creator's profile and all their links. **pillar** — Scrape Pillar pages - `scrape-creators-pp-cli pillar` — Scrapes a Pillar page by URL, extracting the creator's profile, social links, and products. **pinterest** — Scrape Pinterest pins - `scrape-creators-pp-cli pinterest list` — Fetches a paginated list of pins from a Pinterest board by URL, returning each pin's id, description, title, images - `scrape-creators-pp-cli pinterest list-pin` — Fetches detailed information about a single Pinterest pin by URL, returning title, description, link, dominantColor - `scrape-creators-pp-cli pinterest list-search` — Searches Pinterest for pins matching a query, returning results with id, url, title, description, images, link, domain - `scrape-creators-pp-cli pinterest list-user` — Fetches a paginated list of boards for a Pinterest user, returning each board's name, url, description, pin_count **reddit** — Scrape Reddit posts and comments - `scrape-creators-pp-cli reddit create` — Retrieves comments and post details from a Reddit post by URL. - `scrape-creators-pp-cli reddit list` — Searches across all of Reddit for posts matching a query. - `scrape-creators-pp-cli reddit list-post` — Retrieves comments and post details from a Reddit post by URL. - `scrape-creators-pp-cli reddit list-post-2` — Gets the transcript from a Reddit video post or direct v.redd.it URL when Reddit exposes a VTT caption file. - `scrape-creators-pp-cli reddit list-subreddit` — Fetches posts from a subreddit with sorting and filtering options. - `scrape-creators-pp-cli reddit list-subreddit-2` — Retrieves metadata about a subreddit by name or URL. The subreddit name must be case-sensitive.
عرض على GitHub
ملف SKILL.md هذا كبير جدا، لذلك يعرض SkillsMP القسم الاول فقط هنا. عرض على GitHub