Geonode logo

Unlimited Reddit scraper for AI agents.

Reddit subreddit listings, posts, full comment trees, user profiles and search results. Unlimited requests for one flat bill from $50/mo, no credits, and a blocked request is free.

Start 48-hour free trial

No card. Unlimited pages. 5 threads. Failed requests $0

Unlimited web scraper
01

Images

02

Title

  • Title
  • Condition
03

Price & bidding

  • Price
  • Bids
  • Time left
  • Watchers
04

Seller

  • Seller rating
  • Sold listings
05

Specifics & shipping

  • Item specifics
  • Shipping
  • Item location

JavaScript rendered, parsed to JSON or Markdown, and a fetch that fails after three retries is never billed

Call it on Reddit

curl -X POST "https://scraper.geonode.io/v1/extract" \
  -H "X-Api-Key: $GEONODE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "url": "https://www.reddit.com/r/selfhosted/comments/1fz8k2q/what_do_you_use_for_home_backups_in_2026/",
  "formats": [
    "json",
    "markdown"
  ],
  "render_js": true,
  "proxy": {
    "country": "US"
  }
}'

You came for an Reddit scraper. The key opens everything.

100% first-pass success

you’re here

Subreddit listing: Sort tab (Hot, New, Top, Rising), Post cards, Flair chips, Subscriber and online counts, Sidebar rules and description, Pinned posts

Post header: Title, Author and author flair, Post flair, Score and upvote ratio, Comment count, Awards, Posted time, NSFW / Spoiler tags

Post body: Selftext markdown, Link URL and domain, Image / gallery, Video with audio track, Poll options and votes, Crosspost parent

Comment tree: Comment body, Depth and parent ID, Score and controversial flag, OP / mod distinguished badge, Collapsed and load-more stubs, Edited timestamp

User profile: Post and comment karma, Account age, Recent posts and comments, Trophies, Moderated subreddits, Verified email badge

One subscription, from $50 a month flat. Everything after the first line adds $0

The hard part is getting in. That's our part.

  • $50/moflat, priced by concurrency, not per request

  • 0credits, meters or per-page charges

  • $0for any Reddit request that still fails after 3 retries

  • 48 hfree trial on your own Reddit URLs

  • Cloudflare98%

  • Akamai97%

  • Imperva Incapsula97%

  • DataDome96%

  • PerimeterX / HUMAN95%

  • reCAPTCHA v395%

  • Kasada94%

  • hCaptcha93%

First-pass success by anti-bot vendor, measured by us over 30 days, so of course it flatters us. The 48-hour trial exists so you can count against your own URLs. Full benchmark

What unlimited changes for Reddit data

  • Sold-comps pricing

    Completed listings are the real price of anything. Pull them at scale and price with evidence.

  • Arbitrage sourcing

    Sweep categories for mispriced inventory continuously instead of when the credit budget allows

  • Demand research

    Bids, watchers and sell-through, across a whole category, on one flat bill

  • Collection tracking

    Every listing of the thing you care about, the moment it appears

Cheap is supposed to be the trade-off

  • Michael T.

    Switched from Bright Data. Same quality, fraction of the cost. Their wholesale model actually makes sense.

  • Kristina Halvorson

    Finally a proxy provider that actually feels built for developers

  • Edgar Weissnat

    We run multiple large-scale data collection pipelines across several regions, and proxy reliability has always been one of the biggest bottlenecks. After switching to Geonode we noticed two immediate improvements: connection stability and predictable pricing. Previously we had to constantly optimize traffic to avoid massive proxy bills. With Geonode’s wholesale pricing model we can focus on building products instead of worrying about proxy usage. Integration with our Python stack and Playwright automation was straightforward and took less than a day.

  • Gladys Paucek

    Best proxy infrastructure I've used. The MCP integration with our AI agents was seamless.

  • Gladys Paucek

    Finally something that doesn’t feel like enterprise sales software.

Pick your thread count. That's the only decision.

Nothing is metered. Plans price threads (how many requests run at once), and pages stay unlimited on every one. An Reddit page that needs an anti-bot bypass costs exactly what a plain page costs

Roughly what each plan clears in a month, by endpoint

Starter

2 threads

$50 /mo

~470K
~280K
~1.1M
~610K
~120K
~400K

Growth

5 threads

$100 /mo

Popular
~1.2M
~700K
~2.8M
~1.5M
~300K
~1M

Scale

25 threads

$400 /mo

~5.9M
~3.5M
~13.8M
~7.6M
~1.5M
~5M

Pro

100 threads

$1,250 /mo

~23.5M
~14M
~55M
~30M
~5.9M
~20M

Max

250 threads

$2,500 /mo

~59M
~35M
~137M
~76M
~15M
~50M

The 48-hour trial runs at five threads, no card. Custom terms past 250 threads.

Yeah, but...

In our benchmark data we do not yet publish a number for www.reddit.com, so no percentage here. The mechanism is what matters: the Reddit scraper reads the GraphQL gateway from a real browser session over residential IPs, resolves load-more stubs inside the same session, and falls back to old.reddit.com when the new shell is throttled. Reddit 429s and Cloudflare challenges are recognized as blocks, not content, and retried. A request that fails after three retries is never billed, on any plan, so a Reddit scraper run that hits a wall costs nothing.

Subreddit listings under any sort and time window, including flair filters. Individual Reddit posts with the full comment tree, any comment sort, and a parameter for how many load-more stubs to expand. Reddit user profiles with karma, account age, and the recent post and comment feeds. Reddit search, both site-wide and restricted to a subreddit. Subreddit sidebars, rules and moderator lists. Reddit chat, modmail, and anything behind a login such as private subreddits are out of scope for the Reddit scraper. Reddit Ads Manager data is also not read.

Because that stopped being a free lunch in 2023. Reddit now rate-limits the anonymous .json endpoints to roughly ten requests per minute per IP, blocks datacenter ranges, and serves the same comment-tree truncation the HTML does, so a .json Reddit scraper still needs the morechildren calls. The official Reddit API costs per thousand calls after the 2023 pricing change, which is what ended most third-party Reddit apps. The Reddit scraper gives you the same structured fields the .json route used to, at any volume, without an OAuth app or per-call pricing, and it keeps working when Reddit tightens the anonymous limit again.

Reddit serves about 200 top-level comments and their first replies, then stubs. The Reddit scraper expands stubs until it reaches your max_comments or the thread is exhausted, and a 5,000-comment Reddit megathread is one request billed once. Old.reddit.com is used as a fallback and can be forced with a parameter when you want server-rendered markup. NSFW and quarantined subreddits need a consent cookie that the Reddit scraper sets in-session. Reddit content is public and Reddit's terms govern redistribution, so store what your use requires and respect deletions on re-crawl; the Reddit scraper marks [deleted] and [removed] distinctly.

No request credits and no per-page meter. Plans are priced by concurrency, which is the number of requests in flight at once, from $50 a month. Run as many requests through that concurrency as you can, around the clock, for the same bill. A request that fails after three retries is not counted against anything.

Yes. Pass a country code and the request routes through a residential IP there, so marketplaces return local prices and currency, search engines return local results, and geo-restricted pages open. City and ASN targeting are available on the residential pool where a site is sensitive to it.

Typed JSON with the field table on this page, clean Markdown for feeding an LLM context window, or raw rendered HTML if you run your own parser. For pages without a dedicated parser you can pass a JSON schema and get AI-extracted fields back in that shape.

48 hours. Your worst Reddit URLs.

Five threads, no card, and the clock starts at your first API call. If it doesn’t earn the switch, don’t switch.