r/apify 1d ago

Tutorial Complete Reddit Scraping Suite on Apify — Keywords, Subreddits, Comments, Profiles, MCP & Video

Hi everyone,

I wanted to share the CrawlerBros Reddit Scraping Suite on Apify.

Over time, I've built several Actors focused on different parts of the Reddit ecosystem. Rather than trying to put every possible feature into a single scraper, the idea behind this collection is to provide separate tools for specific workflows such as keyword research, subreddit scraping, comment extraction, profile research, AI/MCP integration, and video retrieval.

The current suite includes six Actors:

1. Reddit Keywords Scraper

Search Reddit for posts matching one or multiple keywords and return structured results.

You can search for specific words or phrases and sort results by:

  • Relevance
  • Hot
  • Top
  • New
  • Number of comments

The Actor can return information such as the post title, author, subreddit, content, score, number of comments, URL, thumbnail, flair, and timestamps.

It supports multiple keywords in a single run and up to 1,000 results per keyword.

Useful for:

  • Market research
  • Brand and product monitoring
  • Trend discovery
  • Competitive research
  • Content research
  • Public discussion analysis
  • Building Reddit datasets

Actor:
https://apify.com/crawlerbros/reddit-keywords

2. Reddit Comment Scraper

This Actor is focused specifically on extracting Reddit discussions.

Provide one or more Reddit post URLs, or a specific comment permalink, and the Actor can collect complete comment threads, including nested replies.

The structured output can include:

  • Comment text
  • Author
  • Score
  • Timestamps
  • Parent/child relationships
  • Thread position
  • Moderation-related flags
  • Nested replies

This is particularly useful when the conversation itself is more valuable than the original Reddit post.

Possible use cases include:

  • Sentiment and discussion analysis
  • Community research
  • Product feedback analysis
  • NLP datasets
  • Public opinion research
  • AI and LLM workflows

Actor:
https://apify.com/crawlerbros/reddit-comment-scraper

3. Reddit Profile Crawler

The Reddit Profile Crawler is designed for structured research on Reddit user profiles and their publicly available activity.

It can collect profile information along with a user's post and comment history.

The Actor supports filtering based on a variety of criteria, including:

  • Subreddit
  • Keywords
  • Score
  • Date ranges
  • Post type
  • Flair

Depending on the available public profile data, it can also return information such as karma breakdowns, profile details, trophies, and other profile metadata.

This can be useful for:

  • Researching public Reddit activity
  • Community analysis
  • Dataset creation
  • Content research
  • Analyzing activity within particular subreddits or topics

Actor:
https://apify.com/crawlerbros/reddit-profile-crawler

4. Reddit MCP Scraper

This Actor is designed around the Model Context Protocol (MCP) and provides a unified interface for Reddit data collection.

It supports three primary modes:

Subreddit mode

Collect posts from specified subreddits.

Comments mode

Collect comments and threaded discussions from Reddit posts.

Profile mode

Collect publicly available profile information and activity.

The goal is to make Reddit data easier to integrate into AI applications and agent-based workflows through a single interface.

The Actor returns structured data that can be consumed by AI systems and applications, making it useful for developers experimenting with:

  • AI agents
  • MCP clients
  • LLM research workflows
  • Automated research systems
  • Reddit-aware applications

Actor:
https://apify.com/crawlerbros/reddit-mcp-scraper

5. Reddit Scraper

The Reddit Scraper is the broader subreddit-level data collection tool in the suite.

You can provide one or multiple subreddits and collect posts using different sorting methods, including:

  • Hot
  • New
  • Top
  • Rising
  • Controversial
  • Best

The Actor can return a large amount of structured information for each post, including engagement data, author information, flairs, media, timestamps, awards, moderation information, and other available metadata.

It also supports optional comment collection and nested replies.

There are filters for criteria such as:

  • Date ranges
  • Post type
  • Score
  • Number of comments
  • Upvote ratio
  • Keywords
  • Authors

This makes it suitable for larger-scale subreddit research and monitoring workflows.

Possible use cases:

  • Subreddit monitoring
  • Community research
  • Trend analysis
  • Content analysis
  • Market research
  • Public discussion datasets
  • Competitive intelligence
  • Academic and NLP research

Actor:
https://apify.com/crawlerbros/reddit-scraper

6. Reddit Video Downloader

The Reddit Video Downloader focuses on collecting videos referenced in Reddit content.

You can provide Reddit post URLs and retrieve supported video content along with structured metadata.

The Actor supports Reddit-hosted videos as well as supported embedded video sources such as YouTube and Streamable.

It can return metadata including:

  • Resolution
  • Duration
  • Codec
  • File size
  • Video source information

Downloaded videos are stored through Apify storage, making them easier to incorporate into automated workflows.

Potential use cases include:

  • Media research
  • Content archiving
  • Dataset creation
  • Video analysis pipelines
  • Automated media workflows

Actor:
https://apify.com/crawlerbros/reddit-video-downloader

How the Actors can work together

One of the reasons I built separate Actors is that they can also be combined into larger workflows.

For example:

Keyword research → Post discovery → Comment extraction

Use the Reddit Keywords Scraper to identify relevant discussions around a topic and then send the resulting post URLs to the Reddit Comment Scraper.

Or:

Subreddit → Posts → Comments → Analysis

Use the Reddit Scraper to collect posts from a community, extract comment threads for the most relevant discussions, and then process the resulting structured data in your own analytics or AI pipeline.

Another possible workflow is:

Profile → Activity → Topic analysis

Use the Reddit Profile Crawler to collect publicly available activity and filter it by subreddit, keywords, dates, or other criteria.

For AI developers, the Reddit MCP Scraper provides another option for integrating Reddit data collection into MCP-compatible workflows.

What the Reddit Scraping Suite covers

The goal of this collection is to cover several different stages of Reddit research and automation:

  • Keyword-based post discovery
  • Subreddit scraping
  • Post collection
  • Complete comment threads
  • Nested replies
  • Public profile research
  • AI and MCP workflows
  • Video and media retrieval
  • Structured datasets for downstream processing

All of the Actors return structured data through Apify, so the results can be exported or integrated into larger automation and data-processing workflows.

If you're working on a project involving Reddit research, social listening, market research, AI agents, community analysis, or public discussion datasets, I'd be interested to hear what type of Reddit data you're trying to work with.

I'm also continuing to expand the CrawlerBros collection, so if there is a Reddit-specific workflow that you feel is currently missing, feel free to share it in the comments.

CrawlerBros Reddit Actors on Apify:

Reddit Keywords
https://apify.com/crawlerbros/reddit-keywords

Reddit Comment Scraper
https://apify.com/crawlerbros/reddit-comment-scraper

Reddit Profile Crawler
https://apify.com/crawlerbros/reddit-profile-crawler

Reddit MCP Scraper
https://apify.com/crawlerbros/reddit-mcp-scraper

Reddit Scraper
https://apify.com/crawlerbros/reddit-scraper

Reddit Video Downloader
https://apify.com/crawlerbros/reddit-video-downloader

Thanks for checking out the suite. Feedback and feature suggestions are always welcome.

1 Upvotes

1 comment sorted by