awesome-repositories.com
Blog
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektÜber unsRanking-MethodikPresseMCP-Server
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
n0tan3rd avatar

n0tan3rd/squidwarc

0
View on GitHub↗
176 Stars·25 Forks·JavaScript·Apache-2.0·2 Aufrufen0tan3rd.github.io/Squidwarc↗

Squidwarc

Squidwarc is a high fidelity, user scriptable, archival crawler that uses Chrome or Chromium with or without a head.

Features

  • JavaScript Crawling Frameworks - High-fidelity archival crawler using Chrome.
  • Web Scraping - High-fidelity archival crawler using headless Chrome.

Star-Verlauf

Star-Verlauf für n0tan3rd/squidwarcStar-Verlauf für n0tan3rd/squidwarc

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Häufig gestellte Fragen

Was macht n0tan3rd/squidwarc?

Squidwarc is a high fidelity, user scriptable, archival crawler that uses Chrome or Chromium with or without a head.

Was sind die Hauptfunktionen von n0tan3rd/squidwarc?

Die Hauptfunktionen von n0tan3rd/squidwarc sind: JavaScript Crawling Frameworks, Web Scraping.

Welche Open-Source-Alternativen gibt es zu n0tan3rd/squidwarc?

Open-Source-Alternativen zu n0tan3rd/squidwarc sind unter anderem: cgiffard/node-simplecrawler — simplecrawler is designed to provide a basic, flexible and robust API for crawling websites. It was written to… ionicabizau/scrape-it — scrape-it is a Node.js web scraper and HTML parser designed to extract structured data from websites and HTML files.… antivanov/js-crawler — js-crawler. brendonboshell/supercrawler — Supercrawler is a Node.js web crawler. It is designed to be highly configurable and easy to use. bda-research/node-crawler — node-crawler is a programmable web crawler for Node.js that manages request queues and automates data extraction. It… lapwinglabs/x-ray — X-Ray is a web scraping framework and asynchronous web crawler designed to extract structured data from websites. It…

Open-Source-Alternativen zu Squidwarc

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit Squidwarc.
  • bda-research/node-crawlerAvatar von bda-research

    bda-research/node-crawler

    6,785Auf GitHub ansehen↗

    node-crawler is a programmable web crawler for Node.js that manages request queues and automates data extraction. It functions as a rate-limited HTTP client and a headless HTML parser, providing the infrastructure to visit large sets of URLs asynchronously while preventing duplicate processing through task deduplication. The project distinguishes itself through a proxy rotation manager that cycles user agents and proxy servers to bypass access restrictions. It utilizes the HTTP/2 protocol to improve request performance and server compatibility during large-scale scraping operations. The syst

    TypeScriptcheeriocrawlerextract-data
    Auf GitHub ansehen↗6,785
  • brendonboshell/supercrawlerAvatar von brendonboshell

    brendonboshell/supercrawler

    381Auf GitHub ansehen↗

    Supercrawler is a Node.js web crawler. It is designed to be highly configurable and easy to use.

    JavaScript
    Auf GitHub ansehen↗381
  • antivanov/js-crawlerAvatar von antivanov

    antivanov/js-crawler

    257Auf GitHub ansehen↗

    js-crawler

    TypeScript
    Auf GitHub ansehen↗257
  • cgiffard/node-simplecrawlerAvatar von cgiffard

    cgiffard/node-simplecrawler

    2,133Auf GitHub ansehen↗

    simplecrawler is designed to provide a basic, flexible and robust API for crawling websites. It was written to archive, analyse, and search some very large websites and has happily chewed through hundreds of thousands of pages and written tens of gigabytes to disk without issue.

    JavaScript
    Auf GitHub ansehen↗2,133
Alle 30 Alternativen zu Squidwarc anzeigen→