Apify Actor · No login required

Every public Substack, in clean JSON.

Post archives, full articles, comment threads with nested replies, author profiles, and category leaderboards with subscriber-count estimates — one actor, six modes, structured output you can use immediately.

From $0.30 / 1,000 records · all 6 data types in one actor · 95-test suite · never bypasses paywalls

comments mode — nested replies, flattened with depth
[
  {
    "postTitle": "Why we're leaving Substack",
    "authorName": "Gordon Strause",
    "depth": 0,
    "body": "Too bad — I think their policies are the right ones…",
    "reactionCount": 132,
    "childCount": 2
  },
  {
    "authorName": "Casey Newton",
    "depth": 1,          // ← a reply, one level deep
    "parentId": 47125539,
    "body": "Appreciate that, Gordon. We didn't decide lightly.",
    "reactionCount": 41,
    "isPinned": true
  }
]
One actor, six modes

Pull exactly the data you need

Switch modes with a single input field. Every mode returns flat, ready-to-use rows — no HTML soup, no pagination code, no login.

📚

Publication archive

A newsletter's full post list, newest-first, auto-paginated. Filter by date, audience, and type.

📄

Single post

Full public content of one post as clean HTML and auto-extracted plain text.

💬

Comments ★ rare

Every top-level comment plus nested replies, depth-flattened — at the same flat rate, no per-comment surcharge.

✍️

Author profile

Bio, publication details, custom domain, Twitter handle, and paid status — no auth needed.

🏆

Category leaderboard

Ranked publications in any of 30+ categories, with real subscriber estimates (1.1M+, 228K+).

🔎

Search

Keyword lookup across posts and publications for quick discovery.

How it compares

Six data types, one actor.

Most Substack actors do one thing — posts, or notes, or a leaderboard — and charge $1–$5 per 1,000. This one does all six in a single run, includes comment threads at the same flat rate, and comes in at the lowest price in the category.

ActorPrice / 1KCommentsData types
Substack Scraper$0.30✓ flat rate6 in one actor
sourabhbgp$0.30 + fees✓ +per-comment fee3
benthepythondev$1.001
easyapi$2.99–4.99 ×6 actors1 each
automation-labhigher1

Pure JSON API — no headless browser, no proxy, no JS rendering. Low cost to run, so the price stays low.

Pricing

Pay only for what you pull

$0.30 / 1,000 records
The lowest price of any all-in-one Substack scraper.

Pay-per-event: a $0.002 start fee plus $0.0003 per record. One flat rate for posts, comments, authors, and leaderboards — no per-comment surcharge, no subscription, no minimums. A 50-post pull costs under two cents.

  • Quick 50-post archive~$0.02
  • 500-post full archive~$0.15
  • 1,000-post AI dataset~$0.30
  • Top 200 newsletters (leaderboard)~$0.06
  • 500 comments on a viral post~$0.15
What people build with it

Research, leads, and AI datasets

Competitive intelligence

Pull a competitor's full archive and analyze posting cadence, topics, and engagement over time.

Lead generation

Top-N newsletters by subscriber estimate, with author handles — an outreach list in one run.

AI training data

Large batches of long-form text with bodyText already clean — no HTML stripping.

Audience sentiment

Every comment and reply for a post, with depth to rebuild the full conversation.

Turn any Substack into data in one run.

No login, no proxy setup, no scraping code. Point it at a publication and go.

Get it on Apify →