awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com

Instagram profile scraper

Ranking updated Jul 4, 2026

For a tool for scraping public Instagram profiles, the first results are instaloader/instaloader, datalux/osintgram (Osintgram is a self-hostable CLI tool that uses headless browser automation to extract public Instagram profile data for OSINT purposes, fitting the category of an Instagram scraper with likely support for posts, comments, followers, and session handling) and justanotherarchivist/snscrape. ping/instagram_private_api and subzeroid/instagrapi round out the shortlist. Compare the match explanations and check the project documentation against your requirements.

Find the best Instagram profile scrapers. We compare top open-source tools ranked by activity and reliability to help you pick the right one.

Instagram profile scraper

Find the best repos with AI.We'll search the best matching repositories with AI.
  • instaloader/instaloaderinstaloader avatar

    instaloader/instaloader

    11,619View on GitHub↗

    Instaloader is a Python library and command-line utility designed for the automated retrieval, archiving, and analysis of Instagram content. It provides a programmatic interface to fetch media, captions, and metadata from public or private profiles, hashtags, and stories, while maintaining persistent user sessions for authorized access. The tool distinguishes itself through robust archive management and traffic control mechanisms. It supports incremental synchronization, allowing users to resume interrupted downloads and update local collections without redundant requests. To ensure reliable

    Instaloader is a dedicated command-line tool and library for scraping Instagram data, supporting extraction of posts, media, captions, comments, follower and following lists, with built-in pagination handling, rate-limit management, session-based login, and export to CSV/JSON — fitting your requirements for a self-hostable Instagram scraper.

    PythonSocial Media Extraction ToolsSocial Media Scrapers
    View on GitHub↗11,619
  • datalux/osintgramDatalux avatar

    Datalux/Osintgram

    13,179View on GitHub↗

    Osintgram is a command-line utility designed for open-source intelligence gathering and the extraction of public data from social media profiles. It functions as a framework for collecting and processing user information to assist in digital investigations and the mapping of digital footprints. The tool distinguishes itself through a modular architecture that organizes intelligence-gathering tasks into independent scripts, all sharing a unified session state and data processing pipeline. It utilizes headless browser automation and session-based interactions to mimic legitimate user behavior,

    Osintgram is a self-hostable CLI tool that uses headless browser automation to extract public Instagram profile data for OSINT purposes, fitting the category of an Instagram scraper with likely support for posts, comments, followers, and session handling.

    PythonSocial Media Extraction ToolsSocial Media Scrapers
    View on GitHub↗13,179
  • justanotherarchivist/snscrapeJustAnotherArchivist avatar

    JustAnotherArchivist/snscrape

    5,398View on GitHub↗

    snscrape is a Python-based social media web scraper and crawler designed to extract public posts, profiles, and hashtags from social networks without the use of official APIs. It functions as an archival tool and a utility for open-source intelligence data collection, allowing for the gathering of publicly available information to investigate trends and people. The tool facilitates social media data extraction for research and archival purposes, enabling the creation of historical records of conversations and user activity. It supports workflows for academic social analysis and the export of

    snscrape is a Python-based social media scraper that can extract posts, media, comments, and profile data from Instagram without using the official API, supports cursor-based pagination, exports to JSON/CSV, and runs as a self-hostable command-line tool—matching the requested Instagram scraping features.

    PythonSocial Media Extraction Tools
    View on GitHub↗5,398
  • ping/instagram_private_apiping avatar

    ping/instagram_private_api

    3,234View on GitHub↗

    This Python library is a private API wrapper that provides programmatic access to Instagram features by communicating with internal mobile endpoints. It functions as a social media automation toolkit for managing profiles, publishing media, and interacting with the social graph. The library uses a reverse-engineered API to mimic the communication patterns and request headers of mobile applications. It incorporates a session manager that persists authentication cookies and client metadata to maintain active logins and reduce the frequency of authentication handshakes. Its capabilities cover m

    This Python library wraps Instagram's private API to programmatically access posts, comments, and profile data, making it a solid fit for scraping publicly available Instagram content, though it is a general automation toolkit rather than a dedicated scraper with built-in CLI and export utilities.

    PythonSocial Media Extraction Tools
    View on GitHub↗3,234
  • subzeroid/instagrapisubzeroid avatar

    subzeroid/instagrapi

    6,366View on GitHub↗

    Instagrapi is a Python library that provides a complete interface to Instagram's private API for scraping posts, comments, followers, and stories, with login support and pagination handling, making it a solid fit for building an Instagram data extraction tool.

    PythonSocial Media Extraction Tools
    View on GitHub↗6,366
  • th3unkn0n/osi.igth3unkn0n avatar

    th3unkn0n/osi.ig

    1,472View on GitHub↗

    th3unkn0n/osi.ig is an open-source Python Instagram scraper built for OSINT, directly addressing the need to extract posts, comments, and follower data from public pages, though you should verify handling of pagination and export formats from its code or docs.

    PythonSocial Media Intelligence
    View on GitHub↗1,472
  • huaying/instagram-crawlerhuaying avatar

    huaying/instagram-crawler

    1,335View on GitHub↗

    This project is a web scraping and automation tool designed to collect public data from Instagram and perform automated social media interactions. It functions by gathering profile details, captions, media files, and engagement metrics directly from web pages, bypassing the need for official developer interfaces or platform-specific credentials. The tool distinguishes itself by combining data extraction with automated engagement capabilities. It allows users to programmatically interact with content by liking posts that match specific search criteria or hashtags, aiming to increase account vi

    This Python-based Instagram scraper extracts posts, profile data, and hashtag results without using the official API, making it a self-hostable command-line tool that fits the core scraping purpose, though explicit comment and follower data extraction is not confirmed.

    PythonSocial Media Scrapers
    View on GitHub↗1,335
  • evil0ctal/douyin_tiktok_download_apiEvil0ctal avatar

    Evil0ctal/Douyin_TikTok_Download_API

    16,336View on GitHub↗

    This project is a RESTful media extraction service that provides a programmatic interface for downloading video and image content from social media platforms. It functions as a scraper that parses shared URLs and user profile identifiers to isolate direct media streams and associated metadata from platform-specific data structures. The service distinguishes itself through its ability to emulate cryptographic signatures and security tokens required to authenticate requests against protected backend services. By simulating headless browser behavior and managing cookies and headers, the system b

    This repository is a media extraction service for Douyin and TikTok, not Instagram, so it scrapes the wrong platform for the visitor's stated need.

    PythonSocial Media Extraction ToolsSocial Media Scrapers
    View on GitHub↗16,336
  • twintproject/twinttwintproject avatar

    twintproject/twint

    16,319View on GitHub↗

    Twint is an open-source intelligence and data extraction framework designed to gather public social media information. It functions as a command-line utility that retrieves posts, user profiles, and follower lists directly from web interfaces, bypassing the need for official platform developer credentials or authentication keys. The tool distinguishes itself by enabling automated, large-scale data collection through terminal-based orchestration. It supports granular filtering by keywords, geographic locations, time ranges, and account status, allowing researchers to build targeted datasets fo

    Twint is a Twitter-specific scraping tool, not an Instagram scraper—it extracts tweets and Twitter profiles but lacks any Instagram features like extracting posts, comments, or follower data from Instagram pages.

    PythonSocial Media Extraction ToolsSocial Media Scrapers
    View on GitHub↗16,319
  • qeeqbox/social-analyzerqeeqbox avatar

    qeeqbox/social-analyzer

    21,134View on GitHub↗

    Social-analyzer is an open-source intelligence framework designed for the automated discovery, correlation, and verification of digital identities across online platforms. It functions as a comprehensive engine for gathering social media intelligence, utilizing distributed browser automation to extract metadata and profile information from hundreds of websites simultaneously. The platform distinguishes itself through its ability to perform cross-platform identity correlation using heuristic-based pattern matching and name permutation generation. It processes these findings through a confidenc

    Social Analyzer is an OSINT framework for cross-platform identity discovery and verification, not a dedicated Instagram scraper for extracting posts, comments, and follower data from specific pages.

    JavaScriptSocial Media Extraction ToolsSocial Media Scrapers
    View on GitHub↗21,134
  • nanmicoder/mediacrawlerNanmiCoder avatar

    NanmiCoder/MediaCrawler

    51,294View on GitHub↗

    MediaCrawler is an automated web scraping framework designed to extract public posts, comments, and creator metadata from various social media platforms. It functions as a headless browser automator, utilizing real browser instances to render dynamic content and execute the client-side scripts necessary for interacting with modern web interfaces. The system distinguishes itself through a focus on session persistence and network flexibility. It supports remote debugging to reuse active browser sessions and cookies, which helps minimize the risk of triggering platform security challenges. To ma

    MediaCrawler is a general automated framework for extracting public data from multiple social media platforms, but it is not specifically built for Instagram and does not mention Instagram in its description or tags, so it is a broader tool rather than a dedicated Instagram scraper.

    PythonSocial Media Extraction ToolsSocial Media Scrapers
    View on GitHub↗51,294
  • hu17889/go_spiderhu17889 avatar

    hu17889/go_spider

    1,821View on GitHub↗

    Go Spider is a modular framework designed for building concurrent web scrapers and data extraction workflows. It provides a structured engine for orchestrating automated crawling tasks, managing request scheduling, and processing web content through a unified pipeline. The framework distinguishes itself through a highly configurable architecture that allows developers to inject custom logic for downloaders, schedulers, and storage components via interface-driven contracts. It manages network interactions using middleware-based request throttling and URL deduplication, ensuring that crawling o

    Go Spider is a general-purpose concurrent web scraping framework, not a dedicated Instagram scraper — it provides the building blocks for custom crawlers but does not directly extract Instagram posts, followers, or comments.

    GoRequest Throttling
    View on GitHub↗1,821
Compare the top 10 at a glance
RepositoryStarsLanguageLicenseLast push
instaloader/instaloader11.6KPythonmitJan 18, 2026
datalux/osintgram13.2KPythonGPL-3.0Aug 25, 2025
justanotherarchivist/snscrape5.4KPythonGPL-3.0Nov 15, 2023
ping/instagram_private_api3.2KPythonmitMay 6, 2024
subzeroid/instagrapi6.4KPythonNOASSERTIONJun 20, 2026
th3unkn0n/osi.ig1.5KPython—Feb 1, 2024
huaying/instagram-crawler1.3KPythonMITMay 3, 2024
evil0ctal/douyin_tiktok_download_api16.3KPythonapache-2.0Oct 12, 2025
twintproject/twint16.3KPythonmitFeb 23, 2023
qeeqbox/social-analyzer21.1KJavaScriptagpl-3.0Jan 12, 2026

Related searches

  • Instagram client
  • a federated Instagram
  • a library for scraping web data
  • a residential proxy service for web scraping
  • a web scraping tool for data extraction
  • a web scraping framework for Python
  • a tool for managing multiple social accounts
  • a username enumeration OSINT tool