awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetÀ proposNotre méthodologiePresseServeur MCP
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
n0tan3rd avatar

n0tan3rd/squidwarc

0
View on GitHub↗
176 stars·25 forks·JavaScript·Apache-2.0·2 vuesn0tan3rd.github.io/Squidwarc↗

Squidwarc

Squidwarc is a high fidelity, user scriptable, archival crawler that uses Chrome or Chromium with or without a head.

Features

  • JavaScript Crawling Frameworks - High-fidelity archival crawler using Chrome.
  • Web Scraping - High-fidelity archival crawler using headless Chrome.

Historique des stars

Graphique de l'historique des stars pour n0tan3rd/squidwarcGraphique de l'historique des stars pour n0tan3rd/squidwarc

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Start searching with AI

Alternatives open source à Squidwarc

Projets open source similaires, classés selon le nombre de fonctionnalités partagées avec Squidwarc.
  • bda-research/node-crawlerAvatar de bda-research

    bda-research/node-crawler

    6,785Voir sur GitHub↗

    node-crawler is a programmable web crawler for Node.js that manages request queues and automates data extraction. It functions as a rate-limited HTTP client and a headless HTML parser, providing the infrastructure to visit large sets of URLs asynchronously while preventing duplicate processing through task deduplication. The project distinguishes itself through a proxy rotation manager that cycles user agents and proxy servers to bypass access restrictions. It utilizes the HTTP/2 protocol to improve request performance and server compatibility during large-scale scraping operations. The syst

    TypeScriptcheeriocrawlerextract-data
    Voir sur GitHub↗6,785
  • brendonboshell/supercrawlerAvatar de brendonboshell

    brendonboshell/supercrawler

    381Voir sur GitHub↗

    Supercrawler is a Node.js web crawler. It is designed to be highly configurable and easy to use.

    JavaScript
    Voir sur GitHub↗381
  • antivanov/js-crawlerAvatar de antivanov

    antivanov/js-crawler

    257Voir sur GitHub↗

    js-crawler

    TypeScript
    Voir sur GitHub↗257
  • cgiffard/node-simplecrawlerAvatar de cgiffard

    cgiffard/node-simplecrawler

    2,133Voir sur GitHub↗

    simplecrawler is designed to provide a basic, flexible and robust API for crawling websites. It was written to archive, analyse, and search some very large websites and has happily chewed through hundreds of thousands of pages and written tens of gigabytes to disk without issue.

    JavaScript
    Voir sur GitHub↗2,133
Voir les 30 alternatives à Squidwarc→

Questions fréquentes

Que fait n0tan3rd/squidwarc ?

Squidwarc is a high fidelity, user scriptable, archival crawler that uses Chrome or Chromium with or without a head.

Quelles sont les fonctionnalités principales de n0tan3rd/squidwarc ?

Les fonctionnalités principales de n0tan3rd/squidwarc sont : JavaScript Crawling Frameworks, Web Scraping.

Quelles sont les alternatives open-source à n0tan3rd/squidwarc ?

Les alternatives open-source à n0tan3rd/squidwarc incluent : cgiffard/node-simplecrawler — simplecrawler is designed to provide a basic, flexible and robust API for crawling websites. It was written to… ionicabizau/scrape-it — scrape-it is a Node.js web scraper and HTML parser designed to extract structured data from websites and HTML files.… antivanov/js-crawler — js-crawler. brendonboshell/supercrawler — Supercrawler is a Node.js web crawler. It is designed to be highly configurable and easy to use. bda-research/node-crawler — node-crawler is a programmable web crawler for Node.js that manages request queues and automates data extraction. It… lapwinglabs/x-ray — X-Ray is a web scraping framework and asynchronous web crawler designed to extract structured data from websites. It…