awesome-repositories.com
Blog
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektÜber unsRanking-MethodikPresseMCP-Server
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
scrapy avatar

scrapy/scrapely

0
View on GitHub↗
1,887 Stars·273 Forks·HTML·4 Aufrufe

Scrapely

Scrapely

Features

  • Python Crawling Frameworks - Pure-python library for HTML screen-scraping.

Star-Verlauf

Star-Verlauf für scrapy/scrapelyStar-Verlauf für scrapy/scrapely

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu Scrapely

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit Scrapely.
  • chineking/colaAvatar von chineking

    chineking/cola

    1,501Auf GitHub ansehen↗

    A high-level distributed crawling framework.

    Python
    Auf GitHub ansehen↗1,501
  • cocrawler/cocrawlerAvatar von cocrawler

    cocrawler/cocrawler

    194Auf GitHub ansehen↗

    CoCrawler is a versatile web crawler built using modern tools and concurrency.

    Python
    Auf GitHub ansehen↗194
  • codelucas/newspaperAvatar von codelucas

    codelucas/newspaper

    14,982Auf GitHub ansehen↗

    Newspaper is a Python library designed for scraping, parsing, and analyzing web-based information. It functions as a framework for automated news aggregation and large-scale web content extraction, providing tools to download, clean, and structure text, metadata, and media from diverse online sources. The project distinguishes itself through a pipeline-oriented architecture that combines heuristic-based content extraction with natural language processing. It automatically identifies and isolates article bodies from web page boilerplate while simultaneously performing language detection, keywo

    HTMLcrawlercrawlingnews
    Auf GitHub ansehen↗14,982
  • binux/pyspiderAvatar von binux

    binux/pyspider

    16,809Auf GitHub ansehen↗

    PySpider is a Python web crawling framework designed for automated data extraction. It provides a pipeline for periodically fetching web content, processing HTML, and persisting scraped information into database backends. The system features a web-based management interface for editing scraping scripts, monitoring task progress, and reviewing collected data. It includes a headless browser JavaScript renderer to capture rendered HTML from dynamic web pages and a distributed architecture that uses message queues to scale crawling workloads across multiple nodes. The framework also covers task

    Python
    Auf GitHub ansehen↗16,809
Alle 21 Alternativen zu Scrapely anzeigen→

Häufig gestellte Fragen

Was macht scrapy/scrapely?

Scrapely

Was sind die Hauptfunktionen von scrapy/scrapely?

Die Hauptfunktionen von scrapy/scrapely sind: Python Crawling Frameworks.

Welche Open-Source-Alternativen gibt es zu scrapy/scrapely?

Open-Source-Alternativen zu scrapy/scrapely sind unter anderem: chineking/cola — A high-level distributed crawling framework. cocrawler/cocrawler — CoCrawler is a versatile web crawler built using modern tools and concurrency. codelucas/newspaper — Newspaper is a Python library designed for scraping, parsing, and analyzing web-based information. It functions as a… douban/brownant — |Build Status| |Coverage Status| |PyPI Version| |PyPI Downloads| |Wheel Status|. gaojiuli/gain — Taken Over By Shad0w For Responsible Disclosure [Kiwi BBP]. binux/pyspider — PySpider is a Python web crawling framework designed for automated data extraction. It provides a pipeline for…