awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetÀ proposNotre méthodologiePresseServeur MCP
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to jsvine/waybackpack

Open-source alternatives to Waybackpack

23 open-source projects similar to jsvine/waybackpack, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Waybackpack alternative.

  • akamhy/waybackpyAvatar de akamhy

    akamhy/waybackpy

    589Voir sur GitHub↗

    Wayback Machine API interface & a command-line tool

    Python
    Voir sur GitHub↗589
  • hoothin/userscriptsAvatar de hoothin

    hoothin/UserScripts

    4,065Voir sur GitHub↗

    UserScripts is a collection of JavaScript browser userscripts designed to modify website behavior and add custom functionality to web browsers. It serves as a multi-purpose toolset for web page content automation, web interface enhancement, and specialized web scraping and downloading. The project distinguishes itself through a wide range of specialized utilities, including a browser-based text transformer for character encoding and terminology mapping, and tools for bypassing content censorship. It provides advanced web scraping capabilities such as deciphering obfuscated download links, agg

    JavaScriptadd-onauto-scrollautopager
    Voir sur GitHub↗4,065
  • hartator/wayback-machine-downloaderAvatar de hartator

    hartator/wayback-machine-downloader

    5,897Voir sur GitHub↗

    This project is a command line web archiver and retrieval tool designed to download entire websites and recreate their directory structures from the Internet Archive Wayback Machine. It functions as an archive crawler that automates the bulk retrieval of archived site content for local backup, analysis, or restoration. The tool distinguishes itself through the use of concurrent download optimization to increase the speed of large site retrievals. It enables the reconstruction of historical website versions by mapping URL paths to local file system folders and resolving specific snapshots usin

    Ruby
    Voir sur GitHub↗5,897

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Find more with AI search
  • drawrowfly/tiktok-scraperAvatar de drawrowfly

    drawrowfly/tiktok-scraper

    5,120Voir sur GitHub↗

    This project is a specialized TikTok API scraper and data extractor. It functions as a proxy-based web scraper designed to collect user metadata, video posts, and trend feeds, while providing a webhook data pipeline to route scraped information to external URLs via HTTP requests. The tool includes a watermark-free video downloader that saves high-definition content to local storage. It employs cryptographic request signing for server authentication and utilizes session cookie authentication combined with proxy rotation to manage network traffic and avoid rate limits. Capabilities cover bulk

    TypeScript
    Voir sur GitHub↗5,120
  • egbertbouman/youtube-comment-downloaderAvatar de egbertbouman

    egbertbouman/youtube-comment-downloader

    1,189Voir sur GitHub↗
    Pythondata-sciencedata-scraperpython
    Voir sur GitHub↗1,189
  • footsiefat/zspotifyF

    Footsiefat/zspotify

    0Voir sur GitHub↗
    Voir sur GitHub↗0
  • geiserx/wayback-archiveAvatar de GeiserX

    GeiserX/Wayback-Archive

    16Voir sur GitHub↗

    Download complete websites from the Wayback Machine with full asset preservation for offline viewing

    Python
    Voir sur GitHub↗16
  • haax9/waybackpdfAvatar de Haax9

    Haax9/WaybackPDF

    58Voir sur GitHub↗

    OSINT tool to download archived PDF files from archive.org for a given website.

    Python
    Voir sur GitHub↗58
  • husseinphp/web-archiveAvatar de husseinphp

    husseinphp/web-archive

    69Voir sur GitHub↗
    JavaScript
    Voir sur GitHub↗69
  • lc/gauAvatar de lc

    lc/gau

    4,831Voir sur GitHub↗

    Gau is a command-line tool and passive URL enumerator designed to discover and aggregate known and historical web addresses for specific target domains. It functions as a collection framework that retrieves domain-specific data from public web archives and threat intelligence providers. The tool focuses on passive reconnaissance and open-source intelligence research to map attack surfaces without sending requests directly to target infrastructure. It aggregates data from multiple external sources to identify accessible web endpoints and forgotten pages. The system includes capabilities for r

    Goalienvaultgauhacktoberfest
    Voir sur GitHub↗4,831
  • lorenzoromani1983/wayback-keyword-searchAvatar de lorenzoromani1983

    lorenzoromani1983/wayback-keyword-search

    180Voir sur GitHub↗

    This tool downloads each page from the Wayback Machine for a specific domain and enables further keyword search on each saved page.

    Python
    Voir sur GitHub↗180
  • lunnlew/stream-downloaderAvatar de lunnlew

    lunnlew/stream-downloader

    25Voir sur GitHub↗

    Stream Downloader是一个用于从网络下载媒体资源(视频,音频,图片等)的工具。

    JavaScript
    Voir sur GitHub↗25
  • miniglome/archive.org-downloaderAvatar de MiniGlome

    MiniGlome/Archive.org-Downloader

    1,300Voir sur GitHub↗

    Python3 script to download archive.org books in PDF format

    Python
    Voir sur GitHub↗1,300
  • miserlou/soundscrapeAvatar de Miserlou

    Miserlou/SoundScrape

    1,444Voir sur GitHub↗

    SoundCloud (and Bandcamp and Mixcloud) downloader in Python.

    Python
    Voir sur GitHub↗1,444
  • mohan3d/slideshare-goAvatar de mohan3d

    mohan3d/slideshare-go

    12Voir sur GitHub↗

    API-less slideshare downloader in golang.

    Go
    Voir sur GitHub↗12
  • soimort/you-getAvatar de soimort

    soimort/you-get

    56,839Voir sur GitHub↗

    This project is a command-line utility designed to fetch video, audio, and image content from a wide range of web platforms. It functions by parsing page metadata and utilizing modular, site-specific scripts to extract direct media stream URLs from complex web structures, enabling the local archiving of digital media for offline use. The tool distinguishes itself through its ability to handle authenticated content, allowing users to inject browser-stored session cookies to access restricted or private media. It also supports real-time media streaming by piping remote content directly into ext

    Python
    Voir sur GitHub↗56,839
  • xnl-h4ck3r/waymoreAvatar de xnl-h4ck3r

    xnl-h4ck3r/waymore

    2,672Voir sur GitHub↗

    Find way more from the Wayback Machine, Common Crawl, Alien Vault OTX, URLScan, VirusTotal, GhostArchive & Intelligence X!

    Python
    Voir sur GitHub↗2,672
  • anmolksachan/thetimemachineAvatar de anmolksachan

    anmolksachan/TheTimeMachine

    542Voir sur GitHub↗

    Weaponizing WaybackUrls for Recon, BugBounties , OSINT, Sensitive Endpoints and what not

    Python
    Voir sur GitHub↗542
  • bellingcat/wayback-google-analyticsAvatar de bellingcat

    bellingcat/wayback-google-analytics

    238Voir sur GitHub↗

    A lightweight tool for scraping current and historic Google Analytics data

    Python
    Voir sur GitHub↗238
  • spotdl/spotify-downloaderAvatar de spotDL

    spotDL/spotify-downloader

    23,996Voir sur GitHub↗

    Spotify-downloader is a command-line utility designed to archive music from Spotify by matching track URLs to external video sources. It functions as a high-fidelity downloader that retrieves audio content and saves it as local files, ensuring optimal sound quality by selecting the highest available bitrate from the source media. The tool distinguishes itself through its ability to maintain local music collections by mirroring remote playlist states. It performs local-remote synchronization to determine which tracks require downloading or removal, while utilizing a modular architecture to dec

    Pythondownload-musichacktoberfestmp3
    Voir sur GitHub↗23,996
  • wkentaro/gdownAvatar de wkentaro

    wkentaro/gdown

    5,116Voir sur GitHub↗

    gdown is a command-line tool that downloads public files and folders from Google Drive without requiring authentication. It bypasses the mandatory virus-scan warning page to retrieve large files that conventional download tools block, and can resume interrupted transfers using HTTP range requests. Beyond simple file downloads, gdown can recursively download entire folder hierarchies while preserving the local directory structure. It lists the contents of a public folder as structured JSON without downloading the files themselves, and resolves a file's real name and extension without retrievin

    Pythoncurldownloaddownloader
    Voir sur GitHub↗5,116
  • sparklemotion/mechanizeAvatar de sparklemotion

    sparklemotion/mechanize

    4,443Voir sur GitHub↗

    Mechanize is a Ruby library for web browser automation and headless browser emulation. It allows for programmatically navigating websites and simulating human behavior without a graphical user interface. The library provides an automated interface for populating and submitting web forms, including text fields, checkboxes, and file uploads. It manages stateful sessions by automatically storing and sending cookies across multiple requests to maintain user authentication and identity. Additional capabilities include web data scraping, the ability to download remote web content, and the maintena

    Ruby
    Voir sur GitHub↗4,443
  • adbar/trafilaturaAvatar de adbar

    adbar/trafilatura

    5,319Voir sur GitHub↗

    Trafilatura is a Python library and command-line tool for extracting clean, structured text and metadata from web pages. It downloads HTML content, identifies the main body of text, and strips away navigation, ads, and other boilerplate, returning the core article content along with fields like title, author, date, and URL. The tool can also extract user comments and test whether a page contains extractable text, making it a general-purpose web text extraction library. What distinguishes Trafilatura from simpler extractors is its configurable extraction pipeline, which offers high-speed, high

    Pythonarticle-extractorcorpus-buildercorpus-tools
    Voir sur GitHub↗5,319