Scrape Reddit
Pull posts from a subreddit, or every post matching a search query, as rows.
- Actions
- 2
- Category
- Lead Sourcing — Scrape
- Authentication
- SyncGTM credits — no external API key
- Required input
- subreddit or search query
- Returns
- Posts with title, body, author, score and URL
Available scrapers
| Scraper | Input | What it returns |
|---|---|---|
| Scrape By Search Query | Keyword or phrase | Posts matching the term across Reddit |
| Scrape Community Posts | Subreddit | Top or recent posts from that subreddit |
Reddit is a SyncGTM native integration — it runs on SyncGTM credits, so there is no Reddit account or API key to connect.
What it is good for
Reddit is where people describe a problem before they start shopping for a solution. “Has anyone found a decent X?” is a buying intent post written by the buyer, unprompted, with the context attached.
The catch: Reddit gives you a username, not a company. These rows are best used for research and messaging — reading how your market phrases the problem — and for outreach only where the author identifies themselves.
How to use it
- Open the table and click Table → Import → Scrape Reddit
- Pick search query or community, and enter the term or subreddit
- Set a result limit and a sort order
- Run it, map title / body / author / URL to columns, and click Create Rows
- Point an AI agent at the body column to classify intent
Schedule Auto Import
Turn on Enable Auto Import in the import panel to re-run the search or community scrape on a frequency instead of only when you click it. Each run reuses the saved term, sort order and column mapping, and appends new posts to the same table.
- Run it manually once and confirm the mapping first — a schedule built on an unverified mapping repeats the same mistake every run
- Weekly, sorted by recency, is the intent play — a scheduled sweep beats one huge historical batch, because intent posts age out fast
- On a slow subreddit, a daily run mostly returns what you already have; match the frequency to how often the community actually posts
- Dedupe on post URL, and keep an AI agent scoring intent so nobody reads raw output
- Cap the result limit before you turn it on — every automatic run spends credits exactly as a manual one does
- Pair it with Auto Run so rows a scheduled run creates enrich themselves
Every schedule appears under Manage scheduled imports, where you can pause, edit, run now, or delete it.
Common use cases
- Harvest the exact phrasing your market uses. Copy written in the customer’s words converts better than copy written in yours.
- Find complaint threads about a competitor. A displacement list, self-assembled.
- Track a category subreddit for buying intent. Re-run weekly and filter on new posts asking for recommendations.
- Validate positioning. If nobody in the subreddit describes the problem your headline names, the headline is wrong.
Best practices
- Sort by recency for intent, by score for research — they are different jobs
- Filter out the perennial “what tool should I use” threads if you only want fresh posts
- Have an AI agent score each post for intent before a human reads any of them
- Do not cold-DM someone based on a post unless their profile invites it — it burns the account and the community
- Re-run weekly rather than pulling one huge historical batch
Where to next
- Reddit integration — the same scrapers as enrichment columns
- AI Agents — classify posts by intent at scale
- Web Scrapers — the rest of the catalogue and its legal obligations