๐ How can you extract Reddit posts, comments, and replies from post URLs?
This template collects post and discussion data from submitted Old Reddit post URLs, including subreddit, post title and body, author, upvotes, images and body links, comment count, comment text and metadata, replies, and comment URLs and IDs. It is useful for social listening teams, academic researchers, and community analysts.
Data is collected from Old Reddit. Reddit is a community discussion platform organized into subreddits, posts, comments, and reply threads.
๐ฐ Pricing
Current price: $0.1/1,000 lines. Billing is based on the number of output lines produced by the task.
๐ฆ Output
Reddit Post & Comments Scraper returns these fields:
Input_URLSubRedditPost_TitlePost_BodyPost_UpvotePost_AuthorPost_ImagePost_Body_URLsComment_CountComment_bodyCommentorCommentor_pointsComment_TimeComment_URLComment_IDReplies_bodyReplierReplier_pointsReplied_TimeLoad_More_Comment_Countmsg_typecontentlink_idsortchildrenidlimit_childrenrrenderstyleHTML_SourceStatus
{
"Input_URL": null,
"SubReddit": null,
"Post_Title": null,
"Post_Body": null,
"Post_Upvote": null,
"Post_Author": null,
"Post_Image": null,
"Post_Body_URLs": null,
"Comment_Count": null,
"Comment_body": null,
"Commentor": null,
"Commentor_points": null,
"Comment_Time": null,
"Comment_URL": null,
"Comment_ID": null,
"Replies_body": null,
"Replier": null,
"Replier_points": null,
"Replied_Time": null,
"Load_More_Comment_Count": null,
"msg_type": null,
"content": null,
"link_id": null,
"sort": null,
"children": null,
"id": null,
"limit_children": null,
"r": null,
"renderstyle": null,
"HTML_Source": null,
"Status": null
}
๐ฏ Use Cases
- Analyze discussion themes using post bodies, comments, and replies.
- Measure engagement using post upvotes, comment counts, and commenter points.
- Study conversation structure using comment IDs, URLs, and reply fields.
- Build time-series discussion datasets using comment and reply timestamps.
๐ Why Octoparse
- Built for Old Reddit: The Python automation opens each Old Reddit post URL, parses the post, loads the comment tree, requests additional comment branches when present, and uploads structured post, comment, and reply records.
- Inputs match the workflow: The form uses Old Reddit Post URLs (up to 100,000 entries).
- Fields stay connected: Reddit Post & Comments Scraper returns
Post_Title,Input_URL,Comment_Count,Comment_Time, andPost_Bodyin the same structured dataset. - Ready for repeat use: After Reddit Post & Comments Scraper runs in the cloud, schedule eligible tasks and export the rows for spreadsheets, dashboards, APIs, or AI workflows.
๐ Input
Complete the following fields:
- Old Reddit Post URLs (Required) โ Old Reddit Post URLs. Up to 100,000 entries per run.
๐ How to Use
- Open Reddit Post & Comments Scraper and click Try it!.
- Complete Old Reddit Post URLs using the formats and limits shown in the Input section.
- Run it in the Octoparse cloud and start the task.
- Review fields such as
Post_Title,Input_URL, andComment_Count, then export the rows in the format you need.
๐ก Tips
- Test one representative value in Old Reddit Post URLs before submitting a large batch to Reddit Post & Comments Scraper.
- Use
Post_TitleandInput_URLwhen checking duplicates or comparing repeated exports.
โ FAQ
How many entries can Reddit Post & Comments Scraper accept in Old Reddit Post URLs per run?
Enter up to 100,000 values in Old Reddit Post URLs per run.
Which Old Reddit pages does Reddit Post & Comments Scraper process?
The workflow is configured for comment pages and returns fields such as Post_Title, Input_URL, Comment_Count, and Comment_Time.
What can I use data from Reddit Post & Comments Scraper for?
Analyze discussion themes using post bodies, comments, and replies.
๐ Related Templates
- Reddit Search Scraper โ Use Reddit Search Scraper for a complementary workflow on the same source site.
- Reddit Trending Scraper โ Use Reddit Trending Scraper for a complementary workflow on the same source site.
- Reddit Post Scraper (by Keywords) โ Use Reddit Post Scraper (by Keywords) for a complementary workflow on the same source site.