Skip to content

Latest commit

 

History

1 Commit

Folders and files

NameName
Last commit message
Last commit date
 
 

Repository files navigation

Instagram Audience Scraper

Extract structured data from instagram.com — Audit any public instagram.com account's audience without a login. Samples the accounts that actually engage with its posts and returns one scored row per engager: bot score, engagement rate and the reason tags behind every verdict.

Instagram Audience Scraper on Apify →


🚀 How to use this actor

💚 $5 free Apify credits — every month

No credit card required. No commitment. Cancel anytime.

  1. Click sign up — pick GitHub, Google, or email; takes ~30 seconds
  2. Open this actor — input is pre-filled with a working example
  3. Click Start — export results as JSON, CSV, or Excel

Your $5 monthly platform credit is enough to run this actor right away — and again every month — scraping typically several hundred to several thousand results per run, depending on your input.

Key features

Search with filters — Search by keyword. Filter by description format, and more.

Detail enrichment — Fetch full descriptions, structured metadata for each listing.

Incremental mode — Only get new or changed listings since your last run. Content hash per listing — no duplicates, no re-processing.

Change classification — Track unchanged across runs. Build audit trails of how listings evolve over time.

Compact output — Emit core fields only (AI-agent / MCP-friendly). Keeps response size small for LLM workflows.

Description truncation — Cap description length per ${NOUN} to control output size and cost.

Result cap — Stop after N listings (up to 1.000). Set to 0 for the full catalog.

Export anywhere — Download as JSON, CSV, or Excel. Stream via Apify API, webhooks, or integrations with Make, Zapier, Airbyte, Keboola.

Structured data — Every listing returns the same schema with consistent field naming. All fields always present — null when unavailable, never omitted.


Use cases

Data pipeline automation Integrate with your ETL pipeline to collect structured listings from instagram.com on a schedule. Export to CSV, JSON, or directly to your database. Use compact mode to control output size.

Market research Monitor listings, track trends, and analyze market dynamics with structured, deduplicated data from instagram.com.

Change monitoring Run daily or hourly in incremental mode to capture only new, updated, or expired ${NOUNS}. Perfect for price-tracking, churn analysis, and alerting pipelines.

AI / LLM training data Structured JSON per listing is ready for RAG pipelines, embeddings, and agent workflows. Compact mode trims tokens for LLM context windows.


Quick start

{
  "query": "natgeo",
  "maxResults": 25,
  "includeDetails": true
}

Input parameters

Parameter Type Default Description
query string — Instagram account(s) to audit, without the @. Paste a JSON array to audit several in one run (e.g. ["natgeo","cristiano"]). Private accounts return a single row explaining why they cannot be sampled.
startUrls array [] Instagram profile URLs to audit, one per line (https://www.instagram.com/nasa/). Post and reel URLs are not accepted — this audits accounts, not posts. When provided, these replace the username field; results are deduped by account across all inputs.
maxResults integer 200 How many accounts to score in total, split evenly across the accounts you audit — 200 with two handles samples 100 each. There is no unlimited setting: an audience is sampled from recent activity, so asking for more than that activity holds cannot produce more.
includeDetails boolean true Look up each sampled account's own profile — posts, follower/following ratio, bio and avatar. That is where the bot signal actually lives. Off returns only what the sample itself carries, which systematically under-reports.
descriptionMaxLength integer 0 Truncate each account's bio to N characters. 0 = keep the full bio.
compact boolean false Core fields only (for AI-agent/MCP workflows).
incrementalMode boolean false Compare against the previous run and label each account NEW, UPDATED or UNCHANGED. stateKey is optional — runs with different inputs never share state.
stateKey string — Optional. Stable identifier for the tracked search universe. Leave empty to auto-generate from search inputs.
telegramToken string — Telegram bot token (from @BotFather). Required for Telegram notifications.
telegramChatId string — Telegram chat or channel ID (e.g. "-100123456789"). Required when telegramToken is set.
discordWebhookUrl string — Discord incoming webhook URL. Server Settings → Integrations → Webhooks → New Webhook.
slackWebhookUrl string — Slack incoming webhook URL. api.slack.com/messaging/webhooks.
notificationLimit integer 5 Maximum number of listings included in each notification message (1–20).
notifyOnlyChanges boolean false When Incremental Mode is on, only send notifications for NEW and UPDATED listings. Has no effect outside incremental mode.
whatsappAccessToken string — WhatsApp Cloud API permanent access token (System User token from Meta Business). Recipient must have messaged the business number within the last 24h (service-conversation window — free since Nov 2024).
whatsappPhoneNumberId string — Your WhatsApp Business phone-number ID (numeric, from Meta dashboard). Required when whatsappAccessToken is set.
whatsappTo string — Recipient phone in E.164 format without + (e.g. "436641234567"). Recipient must have messaged your business number within last 24h.
webhookUrl string — Receives a JSON POST with {metadata, items} after each run. Universal escape hatch for n8n / Make / Zapier / custom backends.
webhookHeaders object — Optional JSON object of custom headers (e.g. {"Authorization":"Bearer ..."}).
appConnector string — Optional. Pick a connected app under Settings → API & Integrations to receive your results (including any contact details). Best-effort across MCP connectors as Apify expands its catalog.
mcpIssueTeam string — Only when the connected app is an issue tracker: the team (name or ID) the summary issue is created under, if that app requires one.
descriptionFormat enum "all" Pick a single description representation. all keeps every variant; text / html / markdown drop the others.
excludeEmptyFields boolean false Drop null, empty-string, and empty-array fields from each record before push. Smaller payloads for AI agents and dashboards.
emitUnchanged boolean false When incremental mode is on, also emit listings whose content has not changed since the last run.

Output fields

Every listing returns the same 32-field schema. Missing values are null — never omitted.

  • listingId
  • username
  • profileUrl
  • portalUrl
  • status
  • displayName
  • bio
  • followers
  • following
  • postsCount
  • isPrivate
  • isVerified
  • hasDefaultAvatar
  • profilePicUrl
  • recentPostsSampled
  • avgLikes
  • avgComments
  • engagementRate
  • followerFollowingRatio
  • botScore
  • riskLevel
  • confidence
  • reasonTags
  • sampledAt
  • errorMessage
  • searchQuery
  • contentQuality
  • detailFetched
  • scrapedAt
  • source
  • contentHash
  • changeType

Sample output

One object per listing. Here is a real example from a production run:

{
  "listingId": "b18ac5c89291ca1a6928f38252d1e67eabe95d762b87c48e10e3c07ee298fdf8",
  "username": "zuck",
  "profileUrl": "https://www.instagram.com/zuck/",
  "portalUrl": "https://www.instagram.com/zuck/",
  "status": "ok",
  "displayName": "Mark Zuckerberg",
  "bio": "I build stuff",
  "followers": 17017285,
  "following": 622,
  "postsCount": 439,
  "isPrivate": false,
  "isVerified": true
}

Truncated — full records contain 32 fields. See Output fields for the complete schema.

Try Instagram Audience Scraper now — $5 free credit, no credit card →


Pricing

Pay only for what you extract. No subscription required — Apify's free $5 credit covers thousands of results.

Event Price (USD)
Actor Start $0.00005
Result $0.002

See the actor on Apify for current pricing.


FAQ

How do I scrape instagram.com? Use this actor on Apify to extract structured data from instagram.com. Configure your search query and filters in the input, then click Start — no coding required.

How do I get instagram.com data as JSON, CSV, or Excel? The actor writes each listing to Apify's dataset. Download as JSON, CSV, or Excel from the Console, stream via the API, or push to Make, Zapier, Airbyte, or Keboola.

Is it legal to scrape instagram.com? Web scraping of publicly available data is generally legal. This actor only accesses publicly visible information. Always check instagram.com's terms of service for your specific use case.

How much does it cost? Pay-per-event pricing — you only pay for ${NOUNS} extracted. Apify's free $5 credit is enough to run thousands of results before you pay anything.

How does incremental mode work? Each listing gets a content hash. On subsequent runs, only new or changed listings are emitted — saving time, compute, and storage. Expired listings can be tracked separately.

Do I need an API key or credentials? No. Just sign up for Apify, paste your input, and click Start. No credit card required.


Related products by Black Falcon Data

Browse all Black Falcon Data actors →


Getting started with Apify

New to Apify? Create a free account with $5 credit — no credit card required.

  1. Sign up — $5 platform credit included
  2. Open Instagram Audience Scraper and configure your input
  3. Click Start — export results as JSON, CSV, or Excel

Need more later? See Apify pricing.


About Black Falcon Data

Black Falcon Data builds production-grade web scrapers for job boards and marketplace data. Browse our full actor catalog at www.blackfalcondata.com.


Last updated: 2026 09

About

Audit any public instagram.com account's audience without a login. Samples the accounts that actually engage with its posts and returns one scored row per engager: bot score, engagement rate and the reason tags behind every verdict.

Resources

Stars

0 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors