awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectDespreCum realizăm clasamentulPresăServer MCP
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
scrapy avatar

scrapy/scrapely

0
View on GitHub↗
1,887 stele·273 fork-uri·HTML·6 vizualizări

Scrapely

Scrapely

Features

  • Python Crawling Frameworks - Pure-python library for HTML screen-scraping.

Istoric stele

Graficul istoricului de stele pentru scrapy/scrapelyGraficul istoricului de stele pentru scrapy/scrapely

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru Scrapely

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu Scrapely.
  • chineking/colaAvatar chineking

    chineking/cola

    1,501Vezi pe GitHub↗

    A high-level distributed crawling framework.

    Python
    Vezi pe GitHub↗1,501
  • cocrawler/cocrawlerAvatar cocrawler

    cocrawler/cocrawler

    194Vezi pe GitHub↗

    CoCrawler is a versatile web crawler built using modern tools and concurrency.

    Python
    Vezi pe GitHub↗194
  • codelucas/newspaperAvatar codelucas

    codelucas/newspaper

    14,982Vezi pe GitHub↗

    Newspaper is a Python library designed for scraping, parsing, and analyzing web-based information. It functions as a framework for automated news aggregation and large-scale web content extraction, providing tools to download, clean, and structure text, metadata, and media from diverse online sources. The project distinguishes itself through a pipeline-oriented architecture that combines heuristic-based content extraction with natural language processing. It automatically identifies and isolates article bodies from web page boilerplate while simultaneously performing language detection, keywo

    HTMLcrawlercrawlingnews
    Vezi pe GitHub↗14,982
  • binux/pyspiderAvatar binux

    binux/pyspider

    16,809Vezi pe GitHub↗

    PySpider is a Python web crawling framework designed for automated data extraction. It provides a pipeline for periodically fetching web content, processing HTML, and persisting scraped information into database backends. The system features a web-based management interface for editing scraping scripts, monitoring task progress, and reviewing collected data. It includes a headless browser JavaScript renderer to capture rendered HTML from dynamic web pages and a distributed architecture that uses message queues to scale crawling workloads across multiple nodes. The framework also covers task

    Python
    Vezi pe GitHub↗16,809
Vezi toate cele 21 alternative pentru Scrapely→

Întrebări frecvente

Ce face scrapy/scrapely?

Scrapely

Care sunt principalele funcționalități ale scrapy/scrapely?

Principalele funcționalități ale scrapy/scrapely sunt: Python Crawling Frameworks.

Care sunt câteva alternative open-source pentru scrapy/scrapely?

Alternativele open-source pentru scrapy/scrapely includ: chineking/cola — A high-level distributed crawling framework. cocrawler/cocrawler — CoCrawler is a versatile web crawler built using modern tools and concurrency. codelucas/newspaper — Newspaper is a Python library designed for scraping, parsing, and analyzing web-based information. It functions as a… douban/brownant — |Build Status| |Coverage Status| |PyPI Version| |PyPI Downloads| |Wheel Status|. gaojiuli/gain — Taken Over By Shad0w For Responsible Disclosure [Kiwi BBP]. binux/pyspider — PySpider is a Python web crawling framework designed for automated data extraction. It provides a pipeline for…