Instagram Scraper - Posts, Profiles & Reels avatar

Instagram Scraper - Posts, Profiles & Reels

Deprecated

Pricing

from $0.26 / 1,000 item collecteds

Go to Apify Store
Instagram Scraper - Posts, Profiles & Reels

Instagram Scraper - Posts, Profiles & Reels

Deprecated

Scrape Instagram profiles, posts, reels, comments, followers, hashtags, locations, search, and more through 36 API endpoints with automatic pagination. Powered by Titan Network.

Pricing

from $0.26 / 1,000 item collecteds

Rating

0.0

(0)

Developer

Titan Network

Titan Network

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

17 days ago

Last modified

Categories

Share

Dataclawd Instagram Data Collector

A reliable Actor that collects Instagram data through the Dataclawd API — a managed social-data API with API-key authentication, usage reporting and prepaid credits. This Actor is a thin client: it calls the endpoint you choose (or maps a friendly task to the right endpoint automatically), follows pagination (endCursor / maxId) automatically, extracts each item, and writes one row per item into the default Dataset (each successful row is also the pay-per-event billing event).

It covers 36 endpoints: user info/profile/posts/reels/followers/following, post detail/comments/likers/download & video links, hashtag info/recent/top, search (users/hashtags/places/combined), explore sections, location feeds & city directories, music info, and media/user ID conversion.

Powered by Dataclawd API · Not affiliated with Instagram / Meta, Inc.

Quick start

  1. Open the Actor's Input tab.
  2. Pick a task (collection task), e.g. userPosts — the endpoint is selected automatically.
  3. Fill userId (numeric Instagram user ID — use the idUserId endpoint to convert a username to an ID, e.g. instagram → 25025320).
  4. Optionally set count (items per page), maxItems / maxPages to control volume, or add multiple objects to targets to batch-collect several users.
  5. Click Run and open the Dataset tab when it finishes — every collected item is one row with the raw JSON in the data field.

Input

FieldTypeDefaultDescription
taskenum—Recommended. Collection intent, auto-maps to an endpoint (see Tasks below). Leave empty only if you set endpoint explicitly.
endpointenumautoAdvanced. auto = resolve from task; or set an explicit endpoint to override.
userIdstring—Instagram numeric user ID. Needed by user* list endpoints (posts/reels/followers/...).
usernamestring—Instagram username without @ (e.g. instagram). Used by profile / ID-conversion endpoints.
shortcodestring—Post/reel shortcode (the part after /p/ or /reel/ in a post URL), e.g. C4PYiJQubeR. Needed by post* endpoints.
urlstring—Full post/reel URL (alternative to shortcode).
qstring—Search keyword. Needed by hashtag* / search* endpoints.
typeenumblendedSearch type: blended / users / hashtags / places.
countinteger12Items per page (IG commonly uses 12 / 30).
locationPkstring—Instagram location ID. Needed by location* endpoints.
lat / lngstring—Coordinates for directory* endpoints.
tabenumrankedLocation posts tab: ranked (top) / recent.
countryCode / cityId / pagestring/int—Directory endpoints (directoryCities / directoryLocations).
musicIdstring—Music ID. Needed by musicInfo.
sectionIdstring—Explore section ID (optional for explore*).
maxIdstring(empty)Start cursor for search/hashtag/explore/music endpoints.
targetsarray of object(empty)Batch mode: each element is a param set, collected one after another.
proxyUrlstring(empty)Proxy used by Dataclawd's server when fetching Instagram.
httpMethodenumGETDataclawd Instagram endpoints are GET-only; POST usually returns endpoint_not_found.
maxItemsinteger100Stop after this many items in total (max 100 000).
maxPagesinteger10Max pages per target (max 100).
startCursorstring(empty)Start from a previous cursor (endCursor / maxId).
itemsPathstring(empty)Dot-path to the item list. Leave empty for auto-detection (largest array).
nextCursorPathstring(empty)Dot-path to the next-page cursor. Leave empty for auto-detection (end_cursor / max_id / profile_grid_items_cursor ...).
paginationParamenumautoQuery param used for paging: auto (per endpoint), endCursor, maxId, or none.

Endpoints

Tasks

Pick a task and the Actor selects the endpoint for you. Fill the params below (username / userId / shortcode / q …).

TaskEndpointRequired params
userInfo User infouserInfousername or userId
userPosts User postsuserPostsuserId
userReels User ReelsuserReelsuserId
userFollowers User followersuserFollowersuserId
userFollowing User followinguserFollowinguserId
postDetail Post detailspostDetailshortcode or url
postComments Post commentspostCommentsshortcode or url
postLikers Post likerspostLikersshortcode or url
hashtagTop Hashtag tophashtagTopq
hashtagRecent Hashtag recenthashtagRecentq
search General searchsearchq
searchUsers Search userssearchUsersq
searchPlaces Search placessearchPlacesq
locationFeeds Location postslocationFeedslocationPk
explore Explore pageexploreSection—

ID conversion — idUserId, idUsername, idMediaId, idMediaShortcode User — userInfo, userProfile, userProfile2, userWebProfile, userPosts, userPosts2, userReels, userTagged, userReposts, userRelatedProfiles, userFollowers, userFollowing Post — postDetail, postComments, postTopComments, postLikers, postDownloadLink, postVideoUrl Hashtag — hashtagInfo, hashtagRecent, hashtagTop Search — search, searchUsers, searchHashtags, searchPlaces Explore — exploreSection, exploreSections Location / directory — locationFeeds, locationInfo2, directoryCities, directoryLocations Music — musicInfo

Output (Dataset — one row per collected item)

FieldDescription
platformAlways instagram.
endpointThe endpoint that produced this row.
statuscompleted (data collected) or failed (request-level failure).
targetHuman-readable target summary, e.g. userId=25025320.
pagePage number this item came from (1-based).
itemIndexSequential index within the run (1-based).
cursorCursor used to fetch that page (-1 = first page).
dataThe collected item — the raw JSON returned by Dataclawd (an IG media / user / comment object).
errorStageWhere a failure happened: connect_failed / server_error / auth_error / rate_limited / request_error / parse_error / empty_result.
errorCodeHTTP status code (if any), e.g. 429.
errorError message on failed rows.

Note: some Instagram endpoints require Dataclawd's own Instagram session to be logged in. If Dataclawd's backend is logged out / rate-limited, the API returns login_required / "Please wait a few minutes" style payloads; the Actor writes them as failed rows with the original message in error (best-effort collection still works for endpoints that are available).

Run status & status message

While running, the Actor's status message is updated periodically (Collecting … → Collecting …: wrote n items → Done: …). The run succeeds (SUCCEEDED) as soon as at least one item is collected. If every target fails or returns nothing, the run ends FAILED with an explanatory message. Billing happens only for completed rows.

How to verify the results

  1. Run status: the run is SUCCEEDED when at least one item was written; the status message shows live progress.
  2. Per-item state: open the run's Dataset — every collected item has status=completed with the raw JSON in data; failed rows explain why (errorStage / errorCode / error).
  3. Spot-check against the API: re-run one target in the Dataclawd API Console and compare with the row's data.
  4. Billing: only completed rows carry the item-collected pay-per-event charge.

Developer note: for automated checks (run status, row schema, optional Dataclawd API replay), use python3 tools/verify_output.py --run <RUN_ID> [--api].

Pricing

Pay-per-event (PPE): one item-collected event per successfully collected item. Failed rows do not carry this event.

Publishing note: define an item-collected event with your price in the Actor's Monetization tab. The platform's built-in synthetic event apify-default-dataset-item is charged automatically per default-dataset item when PPE is enabled; set its price to 0 if you don't want a per-item base fee.

Configuration (publisher)

The following environment variables are configured on the Actor's version settings and are not visible to end users:

  • DATACLAWD_API_HOST — Dataclawd API base URL (e.g. https://api.dataclawd.example).
  • DATACLAWD_API_KEY — Dataclawd API key (sent as X-API-Key).

Disclaimer

This Actor is not affiliated with, endorsed by, or sponsored by Instagram or Meta, Inc. Collect only data you are authorized to collect, and comply with Instagram's Terms of Service, the Dataclawd service terms, and applicable law.