Reddit Scraper — Posts, Comments & Discussion Insights
Pricing
from $2.00 / 1,000 results
Reddit Scraper — Posts, Comments & Discussion Insights
Extract Reddit posts, comment threads, user profiles and subreddit rules. Get source-linked media, engagement metrics and question insights from saved discussions. Search, filter by date and export clean JSON, CSV or Excel.
Pricing
from $2.00 / 1,000 results
Rating
0.0
(0)
Developer
Kelopr_bk
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
13 hours ago
Last modified
Categories
Share
Collect Reddit posts, comment threads, public profiles and community details in one place. Start with a subreddit, a search phrase or a permalink, then export source-linked results to JSON, CSV or Excel.
Choose your workflow
| Workflow | What you receive |
|---|---|
| Posts | Feed or search posts, text, scores, comment counts, flair and available media |
| Comments | Comment text, author, parent/post IDs, timestamps and thread structure |
| Users | Public profile, karma, profile description and available avatar links |
| Communities | Subreddit description, membership, posting settings and available rules |
| Discussion insights | Posts and comments plus a separate summary of questions and author participation |
| Custom | Combine result types and control profile histories with individual switches |
Start in three steps
- Choose a workflow. Add Reddit URLs, Subreddits or Search keywords. One search phrase belongs in each entry.
- Set Max results and the per-source/comment allowances. Add date filters if needed.
- Click Start. Open the result tabs while records arrive, then use Export for JSON, CSV or Excel.
Collect posts
{"mode":"POSTS","subreddits":["Python"],"maxItems":50,"maxPostsPerSource":50}
Explore a discussion
Choose Discussion insights, paste a post permalink, and set the number of comments to collect. The Discussion insights table summarizes the saved comment sample: question count, comments by the post author, author participation percentage and source links to sampled questions without saved replies. A missing reply in this sample does not prove the question has no replies on Reddit. These summaries are included; no separate analysis event is charged.
Find comment mentions
Choose Comments and enter a search phrase. This inspects comment text inside threads of matching posts; it does not search Reddit's global comment index. Up to ten matching posts and 100 comments per inspected thread are examined. Use direct post URLs to choose exact threads, or a public user URL to collect that user's comments.
Readable output
Results contains the complete collected records. Posts, Comments, Users and Communities provide focused tables without unrelated rows. Discussion insights contains the included sample analysis. Source issues and Run summary explain unavailable sources, applied filters, duplicate counts and why collection stopped.
A post can include:
- Identity and text:
id,parsedId,title,body,bodyHtml,url,contentUrl,flair. - Author and community:
author,authorId,subreddit,subredditId, available subscriber count. - Engagement:
score,upVotes,numComments,upvoteRatio, age and explicitly derived engagement ratios. - Provenance and media: publication/edit times, body availability status, available image/video links and outbound host.
Comments add parent/post IDs, author participation, available depth and question signals. User/community records include only the public fields supplied by Reddit. Missing fields stay absent; zero and false remain valid values. Enable Include original source data to attach the public source object for auditing.
Filters and data interpretation
Date filters accept ISO-8601 or Unix seconds/milliseconds. Use Publication cutoff for an ISO date, or the compatible numeric Only newer than field for Unix time. The cutoff is exclusive and applies to posts/comments, not profile creation dates. Date-only upper bounds include the entire UTC day. Invalid input returns a corrective run summary without result rows.
Reddit scores can change and are not a verified count of individual positive votes. Text stays in its original language. Optional sentiment is an English word-list heuristic, not an AI verdict; unsupported text is unclassified. Question signals use visible question marks. No fake-review, bot or demographic conclusions are inferred.
Media and exports
Original media URLs are included when available. Optional downloads retain up to three Reddit CDN files per saved post, up to 8 MB per file and 16 MB per run. Storage links follow the run's existing permissions and retention. Failed or oversized downloads keep the original URL and an explanation. User avatars share the same download allowance.
Pricing and result limits
The Store pricing panel shows the effective rate. Each successfully saved default-dataset record is one result event. Type-specific tables mirror these records without another result charge; summaries and issues are not result events. Actor Start remains unchanged. From September 23, 2026 at 16:59 UTC, platform usage is billed separately under the scheduled pricing configuration.
Max results applies across all sources and types. Filtered and duplicate candidates do not consume it. Zero selects the 100,000-result safety ceiling. Budget, time and source availability may stop a run earlier; the summary states the reason. Reddit listings are not a guaranteed complete historical archive. Private, deleted or unavailable content cannot be restored.
Common questions
Why did a successful run save nothing? Check Run summary. A recent cutoff, unavailable source or restrictive filter can legitimately yield zero results.
Can I collect only a profile? Select Users and enter a public profile URL. Custom offers separate switches for submitted posts and comment history.
Can I automate this through the API? Send the same JSON input to the Actor's Run endpoint from the API tab. Use the run's default dataset for all records or the named tables listed in OUTPUT for focused exports.