awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेसMCP सर्वर
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
scrapy avatar

scrapy/scrapely

0
View on GitHub↗
1,887 स्टार्स·273 फोर्क्स·HTML·6 व्यूज़

Scrapely

Scrapely

Features

  • Python Crawling Frameworks - Pure-python library for HTML screen-scraping.

स्टार हिस्ट्री

scrapy/scrapely के लिए स्टार हिस्ट्री चार्टscrapy/scrapely के लिए स्टार हिस्ट्री चार्ट

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI

Scrapely के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो Scrapely के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • chineking/colachineking का अवतार

    chineking/cola

    1,501GitHub पर देखें↗

    A high-level distributed crawling framework.

    Python
    GitHub पर देखें↗1,501
  • cocrawler/cocrawlercocrawler का अवतार

    cocrawler/cocrawler

    194GitHub पर देखें↗

    CoCrawler is a versatile web crawler built using modern tools and concurrency.

    Python
    GitHub पर देखें↗194
  • codelucas/newspapercodelucas का अवतार

    codelucas/newspaper

    14,982GitHub पर देखें↗

    Newspaper is a Python library designed for scraping, parsing, and analyzing web-based information. It functions as a framework for automated news aggregation and large-scale web content extraction, providing tools to download, clean, and structure text, metadata, and media from diverse online sources. The project distinguishes itself through a pipeline-oriented architecture that combines heuristic-based content extraction with natural language processing. It automatically identifies and isolates article bodies from web page boilerplate while simultaneously performing language detection, keywo

    HTMLcrawlercrawlingnews
    GitHub पर देखें↗14,982
  • binux/pyspiderbinux का अवतार

    binux/pyspider

    16,809GitHub पर देखें↗

    PySpider is a Python web crawling framework designed for automated data extraction. It provides a pipeline for periodically fetching web content, processing HTML, and persisting scraped information into database backends. The system features a web-based management interface for editing scraping scripts, monitoring task progress, and reviewing collected data. It includes a headless browser JavaScript renderer to capture rendered HTML from dynamic web pages and a distributed architecture that uses message queues to scale crawling workloads across multiple nodes. The framework also covers task

    Python
    GitHub पर देखें↗16,809
Scrapely के सभी 21 विकल्प देखें→

अक्सर पूछे जाने वाले प्रश्न

scrapy/scrapely क्या करता है?

Scrapely

scrapy/scrapely की मुख्य विशेषताएं क्या हैं?

scrapy/scrapely की मुख्य विशेषताएं हैं: Python Crawling Frameworks।

scrapy/scrapely के कुछ ओपन-सोर्स विकल्प क्या हैं?

scrapy/scrapely के ओपन-सोर्स विकल्पों में शामिल हैं: chineking/cola — A high-level distributed crawling framework. cocrawler/cocrawler — CoCrawler is a versatile web crawler built using modern tools and concurrency. codelucas/newspaper — Newspaper is a Python library designed for scraping, parsing, and analyzing web-based information. It functions as a… douban/brownant — |Build Status| |Coverage Status| |PyPI Version| |PyPI Downloads| |Wheel Status|. gaojiuli/gain — Taken Over By Shad0w For Responsible Disclosure [Kiwi BBP]. binux/pyspider — PySpider is a Python web crawling framework designed for automated data extraction. It provides a pipeline for…