awesome-repositories.com
Blog
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectDespreCum realizăm clasamentulPresăServer MCP
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
yangyangwithgnu avatar

yangyangwithgnu/hardseed

0
View on GitHub↗
9,207 stele·1,782 fork-uri·C++·GPL-2.0·2 vizualizări

Hardseed

Hardseed is a command-line forum media scraper and content archiver. It extracts images and torrent seeds from internet forums in bulk, saving the media and associated metadata to local directories.

The tool utilizes a proxy-enabled web scraping engine that rotates network traffic through multiple proxy servers and protocols to bypass rate limits and network restrictions. It includes a keyword-based media filter that matches user-defined strings within entry titles to include or exclude specific topics.

The system manages data extraction through batch-processed thread iteration and range-limited topic processing. Downloaded assets are organized into a folder hierarchy using template-based directory mapping derived from category and timestamp metadata. All operations are managed via a terminal-based command interface.

Features

  • Forum Media Scrapers - Extracts images and torrent seeds from internet forums in bulk to create local archives.
  • Automated Media Archivers - Automates the collection and local storage of forum media and metadata.
  • Command Line Interfaces - Provides a terminal-based interface for managing scraping and filtering routines.
  • Title-Based Content Filtering - Filters forum topics by matching user-defined keywords within entry titles.
  • Torrent Seed Scrapers - Scrapes and saves torrent files from targeted forum threads into organized directories.
  • CLI Archivers - Saves forum media and metadata using a terminal application with structured naming conventions.
  • Multi-Protocol Proxy Clients - Implements a client-side engine capable of routing traffic through multiple proxy protocols to bypass network restrictions.
  • Proxy Rotation Services - Distributes network traffic across a rotating pool of proxy servers to circumvent rate limits.
  • Proxy Routing - Routes automated download requests through multiple proxies to bypass network restrictions.
  • Proxy-Enabled Media Fetchers - Routes data extraction traffic through proxy servers to bypass rate limits and restrictions.
  • Torrent Seed Collectors - Automatically gathers torrent metadata and magnet links from forum topics.
  • Index Range Selection - Constrains the scraping engine to a specific numerical range of forum IDs.
  • Batch Processing Utilities - Processes forum entries in sequenced groups to extract media without overloading the source server.
  • Batch Media Scraping - Extracts images and torrent seeds by iterating through forum threads in sequenced groups.
  • Range-Limited Indexing - Provides the ability to target specific numerical ranges of forum IDs during the scraping process.
  • Metadata-Driven Directory Mapping - Organizes downloaded assets into a folder hierarchy based on category and timestamp metadata.

Istoric stele

Graficul istoricului de stele pentru yangyangwithgnu/hardseedGraficul istoricului de stele pentru yangyangwithgnu/hardseed

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Întrebări frecvente

Ce face yangyangwithgnu/hardseed?

Hardseed is a command-line forum media scraper and content archiver. It extracts images and torrent seeds from internet forums in bulk, saving the media and associated metadata to local directories.

Care sunt principalele funcționalități ale yangyangwithgnu/hardseed?

Principalele funcționalități ale yangyangwithgnu/hardseed sunt: Forum Media Scrapers, Automated Media Archivers, Command Line Interfaces, Title-Based Content Filtering, Torrent Seed Scrapers, CLI Archivers, Multi-Protocol Proxy Clients, Proxy Rotation Services.

Care sunt câteva alternative open-source pentru yangyangwithgnu/hardseed?

Alternativele open-source pentru yangyangwithgnu/hardseed includ: apify/crawlee-python — Crawlee-python is a web crawling framework for building scalable scrapers using Python. It serves as a comprehensive… rg3/youtube-dl — This project is a command-line video downloader and web media extractor written in Python. It is designed to retrieve… ripmeapp/ripme — Ripme is a batch media downloader and web media scraper designed for extracting images and videos from image-hosting… drawrowfly/tiktok-scraper — This project is a specialized TikTok API scraper and data extractor. It functions as a proxy-based web scraper… jarun/googler — Googler is a command line search client and headless search wrapper that allows users to execute Google searches… calcprogrammer1/openrgb — OpenRGB is a centralized software suite for controlling colors and lighting effects across various brands of RGB…

Alternative open-source pentru Hardseed

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu Hardseed.
  • rg3/youtube-dlAvatar rg3

    rg3/youtube-dl

    140,520Vezi pe GitHub↗

    This project is a command-line video downloader and web media extractor written in Python. It is designed to retrieve video and audio streams from various hosting platforms for local storage or real-time streaming via standard output. The system utilizes a framework of custom extractor classes to handle different websites and allows for the development of new extractors to extend compatibility. It supports accessing restricted, private, or region-locked content through the use of session cookies, user-agent headers, and proxy server routing. Capabilities include media format selection based

    Python
    Vezi pe GitHub↗140,520
  • apify/crawlee-pythonAvatar apify

    apify/crawlee-python

    8,097Vezi pe GitHub↗

    Crawlee-python is a web crawling framework for building scalable scrapers using Python. It serves as a comprehensive tool for web scraping automation, providing a system to extract structured data from websites using both lightweight HTTP requests and headless browser automation. The framework is distinguished by its anti-bot evasion capabilities, which include browser fingerprint impersonation and tiered proxy rotation to bypass detection systems and solve challenges such as Cloudflare. It also incorporates artificial intelligence for autonomous website navigation and schema-based data extra

    Pythonapifyautomationbeautifulsoup
    Vezi pe GitHub↗8,097
  • drawrowfly/tiktok-scraperAvatar drawrowfly

    drawrowfly/tiktok-scraper

    5,120Vezi pe GitHub↗

    This project is a specialized TikTok API scraper and data extractor. It functions as a proxy-based web scraper designed to collect user metadata, video posts, and trend feeds, while providing a webhook data pipeline to route scraped information to external URLs via HTTP requests. The tool includes a watermark-free video downloader that saves high-definition content to local storage. It employs cryptographic request signing for server authentication and utilizes session cookie authentication combined with proxy rotation to manage network traffic and avoid rate limits. Capabilities cover bulk

    TypeScript
    Vezi pe GitHub↗5,120
  • ripmeapp/ripmeAvatar RipMeApp

    RipMeApp/ripme

    4,041Vezi pe GitHub↗

    Ripme is a batch media downloader and web media scraper designed for extracting images and videos from image-hosting platforms and social media sites. It functions as an image gallery downloader and a network client capable of retrieving full albums and paginated content. The project includes a custom media ripper framework that allows for the definition of new extraction rules to support websites lacking native support. It features a proxy-enabled network layer for routing requests through HTTP or SOCKS servers and supports session-based content retrieval using authentication cookies and cus

    Javaalbumarchivalarchive
    Vezi pe GitHub↗4,041
Vezi toate cele 30 alternative pentru Hardseed→