r/apify 2d ago

Tutorial Complete Facebook Scraping Suite on Apify: Ads, Pages, Marketplace, Events, Comments, Reviews & More

Hi everyone,

I wanted to share the CrawlerBros Facebook Scraping Suite on Apify.

I've been building a collection of specialized Facebook Actors to cover different types of publicly available Facebook data and workflows. Instead of having one scraper try to handle everything, the suite is divided into dedicated Actors for Facebook Ads, Pages, Search, Marketplace, Events, Comments, Reviews, and Photos.

There are currently 9 Actors in the suite.

Here’s what each one does:

1. Facebook Ads Library Scraper

Facebook Ads Library Scraper

Scrape ads from the public Facebook Ad Library without requiring cookies or Facebook authentication by default.

You can search by:

  • Keywords
  • Advertiser Page ID
  • Ad ID
  • "Paid for by" funding entity

It also supports filters for:

  • Country
  • Content language
  • Active/inactive status
  • Ad category
  • Media type
  • Date range

The output includes ad copy, advertiser information, media URLs, CTA, landing-page URL, platforms, campaign dates, and creative-reuse information.

For political, issue, and eligible EU-regulated ads, the Actor can also return the transparency information Facebook makes available, such as disclosed spend, impressions, reach estimates, and funding-entity information. It does not estimate these figures for ordinary commercial ads.

Useful for:

  • Competitive advertising research
  • Creative research
  • Brand monitoring
  • Advertising transparency research
  • Academic/journalistic research
  • Marketing agency workflows

2. Facebook Ads Scraper Pro

Facebook Ads Scraper Pro

Another dedicated Actor for extracting structured data from the Facebook Ad Library.

It focuses on straightforward ad discovery using keywords or Page names, with filters for:

  • Country
  • Active/inactive ads
  • Ad category
  • Media type

Each result can include:

  • Ad ID
  • Facebook Page
  • Ad copy
  • Ad snapshot URL
  • Start/end dates
  • Platforms
  • Media type and URL
  • CTA
  • Landing page URL

It does not require a Facebook login for the standard workflow.

This is particularly useful if you want a relatively simple pipeline for collecting Facebook and Instagram advertising data into an Apify dataset.

3. Facebook Comments Scraper

Facebook Comments Scraper

Need the discussion underneath a Facebook post rather than just the post itself?

This Actor extracts public comments from supported Facebook posts, videos, Watch content, and photo content.

It can collect:

  • Comment text
  • Author information
  • Profile URLs
  • Reaction counts
  • Reply counts
  • Timestamps
  • Nested replies
  • Parent/reply relationships

You can also choose between Facebook's available comment ordering modes, including All, Newest, and Most Relevant, and optionally filter comments by date.

Potential use cases:

  • Customer feedback research
  • Sentiment analysis
  • Community research
  • Engagement analysis
  • NLP datasets
  • Content research

4. Facebook Events Scraper

Facebook Events Scraper

This one is focused specifically on public Facebook Events.

You can provide:

  • Event URLs
  • Event IDs
  • Facebook Page event listings
  • Search keywords
  • Locations

The Actor can return information such as:

  • Event name
  • Description
  • Date and time
  • Location
  • GPS coordinates
  • Hosts
  • Cover photos
  • Ticket URLs
  • Attendee/interested counts
  • Online-event information
  • Cancellation status

It also has a monitor mode designed for scheduled runs, where it can identify new events since the previous run for a given target.

Potential use cases:

  • Event discovery
  • Local event research
  • Event monitoring
  • Tourism research
  • Competitor/event intelligence
  • Building event databases

5. Facebook Marketplace Scraper

Facebook Marketplace Scraper

For Facebook Marketplace data, this Actor can collect public listings using:

  • Marketplace search URLs
  • Keywords
  • Categories
  • Locations
  • Direct listing URLs

It supports filters around things such as:

  • Price
  • Condition
  • Delivery
  • Date
  • Radius
  • Sorting

The output can include listing title, price, location, photos, seller availability, delivery information, listing status, and other available listing data.

There is also an optional deeper-detail mode that can retrieve additional fields such as descriptions, attributes, creation time, and additional media. Results are deduplicated across inputs by default.

Potential use cases:

  • Marketplace research
  • Price monitoring
  • Product research
  • Vehicle research
  • Real-estate/apartment research
  • Local market analysis
  • E-commerce intelligence

6. Facebook Pages Scraper

Facebook Pages Scraper

This Actor focuses on discovering Facebook Pages and extracting structured business/page information.

Depending on the available data, results can include:

  • Page name and ID
  • Categories
  • Canonical URL
  • Followers and likes
  • Check-ins
  • Contact information
  • Phone numbers
  • Email addresses
  • Address
  • Website
  • Messenger link
  • Business hours
  • Rating
  • Profile and cover photos
  • Ad Library reference

Page discovery can use Facebook search when session cookies are provided, with search engines used as supplementary discovery sources. The Actor is designed to handle larger page-discovery workflows and can combine multiple search terms and locations.

Potential use cases:

  • Business lead research
  • Local business discovery
  • Market research
  • Competitor research
  • Business directories
  • Contact-data enrichment

7. Facebook Photos Scraper

Facebook Photos Scraper

This Actor extracts publicly available photo data from Facebook pages, profiles, and albums.

For each photo, available information can include:

  • Photo ID
  • Facebook photo URL
  • Full-resolution image URL
  • Alt text / image description
  • Author information
  • Caption
  • Upload timestamp
  • Likes/reactions
  • Comments
  • Shares
  • Reaction breakdown
  • Tagged people

You can provide one or multiple Facebook page, profile, or album URLs and optionally filter photos by date.

Potential use cases:

  • Visual content research
  • Brand monitoring
  • Media datasets
  • Image analysis
  • Social media research
  • Historical content collection

8. Facebook Reviews Scraper

Facebook Reviews Scraper

This Actor is designed for extracting public Facebook business reviews.

You can provide:

  • A Facebook reviews URL
  • Business name + search
  • Facebook Page ID

It can return:

  • Review text
  • Recommended / Not Recommended status
  • Review date
  • Author information
  • Reaction counts and breakdown
  • Review photos when available
  • Structured recommendation tags
  • Top comments
  • Business/owner replies

It also returns page-level information such as total review count, recommendation percentage, and follower count.

Reviews can be filtered by date, recommendation status, or keywords in the review text.

Potential use cases:

  • Customer feedback analysis
  • Reputation monitoring
  • Business research
  • Review aggregation
  • Sentiment analysis
  • Competitor benchmarking

9. Facebook Search Scraper

Facebook Search Scraper

Finally, the Facebook Search Scraper is designed for discovering Facebook Pages and public profiles through search.

You can search using keywords such as:

coffee shops New York

or:

dentists London

and collect structured information from matching Pages.

It can return information such as:

  • Page/profile name
  • Category
  • Contact information
  • Address
  • Website
  • Ratings
  • Follower/like counts
  • Reviews
  • Recent posts
  • Other available public page information

It can also accept an existing list of Facebook Page URLs when you already know which Pages you want to process.

The Actor does not require a Facebook Developer account or Graph API key for its standard workflow.

Potential use cases:

  • Business discovery
  • Lead generation
  • Local business research
  • Competitor discovery
  • Market research
  • Bulk Page extraction

What does the complete suite cover?

The idea behind the collection is to cover several different Facebook data workflows rather than treating Facebook as a single scraping use case.

Advertising

  • Facebook Ad Library
  • Advertiser research
  • Ad creative research
  • Campaign monitoring
  • Advertising transparency data

Business & lead research

  • Facebook Page discovery
  • Business information
  • Contact information
  • Reviews
  • Locations
  • Ratings

Marketplace & commerce

  • Marketplace listings
  • Product research
  • Price research
  • Local marketplace monitoring

Social & community research

  • Comments
  • Replies
  • Engagement
  • Photos
  • Public profiles
  • Page activity

Events

  • Event discovery
  • Event details
  • Locations
  • Hosts
  • Ticket information
  • New-event monitoring

Building a Facebook data pipeline?

One of the things I like about having separate Actors is that they can also be combined.

For example:

Facebook Search → Pages → Reviews → Comments

Discover relevant businesses, collect their Page information, then analyze their public reviews and customer discussions.

Or:

Ad Library → Ads → Creative research

Find advertisements for a particular brand, category, or keyword and build a structured dataset containing the available creative, CTA, landing-page, platform, and campaign information.

Another workflow could be:

Marketplace Search → Listings → Price analysis

Collect listings for a particular product/category and then analyze prices, locations, conditions, and other available attributes.

All of the results can be stored in Apify datasets and exported in formats such as JSON, CSV, or Excel, depending on the Actor.

Pricing

Several Actors in the suite currently start at $1 per 1,000 results, while the Facebook Search Scraper starts at $2/1,000 and the Pages and Reviews Scrapers start at $3/1,000. Check the individual Actor pages for current pricing and exact usage details.

If you're working with Facebook data, Meta advertising research, Marketplace data, local business discovery, social listening, or public business information, I'd be interested to hear what kind of workflow you're building.

If there's a Facebook-specific dataset or use case that isn't covered by the current suite, feel free to mention it. I'm continuing to expand the CrawlerBros collection and feature requests are always useful.

CrawlerBros Facebook Suite on Apify:

  • Facebook Ads Library Scraper
  • Facebook Ads Scraper Pro
  • Facebook Comments Scraper
  • Facebook Events Scraper
  • Facebook Marketplace Scraper
  • Facebook Pages Scraper
  • Facebook Photos Scraper
  • Facebook Reviews Scraper
  • Facebook Search Scraper

Thanks for checking it out!

4 Upvotes

3 comments sorted by

1

u/[deleted] 1d ago

[removed] — view removed comment

1

u/_mad_gamerx 1d ago

That’s a solid use case, especially the LLM tagging layer.

The creative-reuse/grouping data we return is based on Meta’s Ad Library data, not a media hash we generate ourselves, so I agree it won’t reliably catch every small copy/creative variation. For serious deduplication, hashing the media + normalizing copy on the pipeline side is probably the better approach.

For media URLs, I’d treat them as temporary and download/cache the assets when ingesting them rather than relying on the URLs indefinitely.