awesome-repositories.com
博客
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目关于排名机制媒体报道MCP 服务器
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
propublica avatar

propublica/upton

0
View on GitHub↗
1,599 星标·109 分支·HTML·MIT·2 次浏览

Upton

A batteries-included framework for easy web-scraping. Just add CSS! (Or do more.)

Features

  • Web Crawling - Batteries-included framework for simplified web scraping.
  • Ruby Crawling Frameworks - Batteries-included framework for easy scraping.

Star 历史

propublica/upton 的 Star 历史图表propublica/upton 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

Upton 的开源替代方案

相似的开源项目,按与 Upton 的功能重合度排序。
  • postmodern/spidrpostmodern 的头像

    postmodern/spidr

    837在 GitHub 上查看↗

    A versatile Ruby web spidering library that can spider a site, multiple domains, certain links or infinitely. Spidr is designed to be fast and easy to use.

    Ruby
    在 GitHub 上查看↗837
  • sparklemotion/mechanizesparklemotion 的头像

    sparklemotion/mechanize

    4,443在 GitHub 上查看↗

    Mechanize is a Ruby library for web browser automation and headless browser emulation. It allows for programmatically navigating websites and simulating human behavior without a graphical user interface. The library provides an automated interface for populating and submitting web forms, including text fields, checkboxes, and file uploads. It manages stateful sessions by automatically storing and sending cookies across multiple requests to maintain user authentication and identity. Additional capabilities include web data scraping, the ability to download remote web content, and the maintena

    Ruby
    在 GitHub 上查看↗4,443
  • felipecsl/wombatfelipecsl 的头像

    felipecsl/wombat

    1,362在 GitHub 上查看↗

    Lightweight Ruby web crawler/scraper with an elegant DSL which extracts structured data from pages.

    Rubycrawlerdslruby
    在 GitHub 上查看↗1,362
  • mendableai/firecrawl-mcp-servermendableai 的头像

    mendableai/firecrawl-mcp-server

    6,602在 GitHub 上查看↗

    This project is a Model Context Protocol server that connects large language models to web scraping and crawling tools. It functions as a bridge, allowing LLM clients to utilize a web crawling engine and scraping utilities to extract and process web data. The server integrates a markdown web converter that transforms dynamic web pages and PDF documents into clean markdown to optimize consumption by AI models. It also provides a browser automation interface for controlling headless sessions and bypassing access restrictions. The system covers broad capabilities including large-scale website d

    JavaScript
    在 GitHub 上查看↗6,602
查看 Upton 的所有 12 个替代方案→

常见问题解答

propublica/upton 是做什么的?

A batteries-included framework for easy web-scraping. Just add CSS! (Or do more.)

propublica/upton 的主要功能有哪些?

propublica/upton 的主要功能包括:Web Crawling, Ruby Crawling Frameworks。

propublica/upton 有哪些开源替代品?

propublica/upton 的开源替代品包括: felipecsl/wombat — Lightweight Ruby web crawler/scraper with an elegant DSL which extracts structured data from pages. sparklemotion/mechanize — Mechanize is a Ruby library for web browser automation and headless browser emulation. It allows for programmatically… postmodern/spidr — A versatile Ruby web spidering library that can spider a site, multiple domains, certain links or infinitely. Spidr is… lorien/web-scraping — This project is a comprehensive resource directory for web data extraction, providing a curated collection of tools… mendableai/firecrawl-mcp-server — This project is a Model Context Protocol server that connects large language models to web scraping and crawling… adithya-s-k/omniparse — Omniparse is a multimodal content parser and generative AI ingestion engine designed to convert documents, images, and…