awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
twintproject avatar

twintproject/twintArchived

0
View on GitHub↗
16,319 stars·2,786 forks·Python·mit·36 views

Twint

Twint is an open-source intelligence and data extraction framework designed to gather public social media information. It functions as a command-line utility that retrieves posts, user profiles, and follower lists directly from web interfaces, bypassing the need for official platform developer credentials or authentication keys.

The tool distinguishes itself by enabling automated, large-scale data collection through terminal-based orchestration. It supports granular filtering by keywords, geographic locations, time ranges, and account status, allowing researchers to build targeted datasets for sentiment analysis or network relationship mapping. The architecture includes state tracking to resume interrupted sessions and an integrated pipeline for real-time text translation during the collection process.

Beyond raw extraction, the project provides a modular output system that streams data into local files, databases, or external search engines. This design facilitates integration with visualization tools for generating network graphs and interactive dashboards, supporting long-term research workflows and investigative analysis.

Features

  • Command Line Interfaces - Provides a command-line interface for orchestrating large-scale social media data collection tasks.
  • OSINT Automation Frameworks - Automates the retrieval of public web content and mapping of account connections for investigative OSINT analysis.
  • Social Media Extraction Tools - Collects public social media posts, profiles, and follower lists without requiring official developer credentials.
  • Social Media Scrapers - Extracts public social media posts and user profiles via command-line tools without needing official platform API access.
  • Social Media Intelligence Gatherers - Collects and filters public social media activity to support intelligence gathering and long-term research.
  • Social Media Research Tools - Enables researchers to filter and retrieve social media content for investigative and academic analysis.
  • HTTP Request Builders - Fetches public web content by mimicking browser requests to bypass platform authentication and rate limits.
  • Public Data Gathering Frameworks - Gathers public social data from platforms without needing official access keys.
  • Web Data Pipelines - Provides automated workflows to extract and load public social media data into local storage for research.
  • Data Query Filters - Enables granular filtering of social media activity by keywords, location, and time ranges.
  • Search Result Filtering - Filters and searches social media content using criteria like keywords, location, and verification status.
  • Web Data Extraction - Extracts public profile data directly from web interfaces for research and analysis.
  • OSINT Intelligence - Scrape and analyze Twitter data without using the official API.
  • Structured Data Extraction - Extracts structured data from raw web responses by parsing HTML document elements.
  • CLI Task Managers - Executes data extraction tasks directly from the terminal using specific search parameters.
  • Social Network Analysis Tools - Maps and visualizes social network connections and influence patterns using collected follower and interaction data.
  • Cursor-Based Pagination - Implements cursor-based pagination to maintain state and resume interrupted data collection sessions.
  • Data Exporters - Exports scraped data into multiple formats and streams it directly into external databases.
  • Scraping Resumption Handlers - Tracks progress to resume interrupted data collection sessions from the last processed item.
  • Automated Extraction Schedulers - Automates recurring data collection tasks to maintain up-to-date datasets without manual intervention.
  • Local Data Stores - Saves gathered information into local files and databases for long-term research.
  • Relationship Graph Visualizers - Generates network graphs from stored user and follower data to visualize social relationships.
  • Modular Architectures - Features a modular output architecture that streams collected data into various local files and external databases.

Star history

Star history chart for twintproject/twintStar history chart for twintproject/twint

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with Twint

These projects share indexed features with Twint. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • datalux/osintgramDatalux avatar

    Datalux/Osintgram

    13,179View on GitHub↗

    Osintgram is a command-line utility designed for open-source intelligence gathering and the extraction of public data from social media profiles. It functions as a framework for collecting and processing user information to assist in digital investigations and the mapping of digital footprints. The tool distinguishes itself through a modular architecture that organizes intelligence-gathering tasks into independent scripts, all sharing a unified session state and data processing pipeline. It utilizes headless browser automation and session-based interactions to mimic legitimate user behavior,

    Pythonanalysishackinginformation-gathering
    View on GitHub↗13,179
  • instaloader/instaloaderinstaloader avatar

    instaloader/instaloader

    11,619View on GitHub↗

    Instaloader is a Python library and command-line utility designed for the automated retrieval, archiving, and analysis of Instagram content. It provides a programmatic interface to fetch media, captions, and metadata from public or private profiles, hashtags, and stories, while maintaining persistent user sessions for authorized access. The tool distinguishes itself through robust archive management and traffic control mechanisms. It supports incremental synchronization, allowing users to resume interrupted downloads and update local collections without redundant requests. To ensure reliable

    Pythondownloaderinstagraminstagram-client
    View on GitHub↗11,619
  • qeeqbox/social-analyzerqeeqbox avatar

    qeeqbox/social-analyzer

    21,134View on GitHub↗

    Social-analyzer is an open-source intelligence framework designed for the automated discovery, correlation, and verification of digital identities across online platforms. It functions as a comprehensive engine for gathering social media intelligence, utilizing distributed browser automation to extract metadata and profile information from hundreds of websites simultaneously. The platform distinguishes itself through its ability to perform cross-platform identity correlation using heuristic-based pattern matching and name permutation generation. It processes these findings through a confidenc

    JavaScriptanalysisanalyzercli
    View on GitHub↗21,134
  • mxrch/ghuntmxrch avatar

    mxrch/GHunt

    19,089View on GitHub↗

    GHunt is a Google account investigator and open-source intelligence framework designed to retrieve publicly available information and metadata associated with Google accounts. It functions as an OSINT data extractor and offensive security framework used to identify user identities and uncover hidden metadata. The tool extracts public profile data from various Google services and exports the findings into structured JSON formats. This allows for the collection and analysis of digital footprints to support security research and reconnaissance.

    Python
    View on GitHub↗19,089
Compare all 30 related projects→

Frequently asked questions

What does twintproject/twint do?

Twint is an open-source intelligence and data extraction framework designed to gather public social media information. It functions as a command-line utility that retrieves posts, user profiles, and follower lists directly from web interfaces, bypassing the need for official platform developer credentials or authentication keys.

What are the main features of twintproject/twint?

The main features of twintproject/twint are: Command Line Interfaces, OSINT Automation Frameworks, Social Media Extraction Tools, Social Media Scrapers, Social Media Intelligence Gatherers, Social Media Research Tools, HTTP Request Builders, Public Data Gathering Frameworks.

Which projects share features with twintproject/twint?

Projects with overlapping indexed features include: datalux/osintgram — Osintgram is a command-line utility designed for open-source intelligence gathering and the extraction of public data… instaloader/instaloader — Instaloader is a Python library and command-line utility designed for the automated retrieval, archiving, and analysis… qeeqbox/social-analyzer — Social-analyzer is an open-source intelligence framework designed for the automated discovery, correlation, and… mxrch/ghunt — GHunt is a Google account investigator and open-source intelligence framework designed to retrieve publicly available… apify/crawlee — Crawlee is a web scraping framework designed for building scalable, reliable, and distributed data extraction… plausible/analytics — This project is an open-source, privacy-focused web analytics platform designed for high-throughput data ingestion and…