awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
postmodern avatar

postmodern/spidr

0
View on GitHub↗
837 stars·107 forks·Ruby·MIT·9 views

Spidr

A versatile Ruby web spidering library that can spider a site, multiple domains, certain links or infinitely. Spidr is designed to be fast and easy to use.

Features

  • Web Crawling - Versatile library for spidering domains and links.
  • Ruby Crawling Frameworks - Flexible site spidering and link discovery.
  • Test Automation Frameworks - Web spidering library for Ruby.

Star history

Star history chart for postmodern/spidrStar history chart for postmodern/spidr

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Spidr

Similar open-source projects, ranked by how many features they share with Spidr.
  • propublica/uptonpropublica avatar

    propublica/upton

    1,599View on GitHub↗

    A batteries-included framework for easy web-scraping. Just add CSS! (Or do more.)

    HTML
    View on GitHub↗1,599
  • sparklemotion/mechanizesparklemotion avatar

    sparklemotion/mechanize

    4,443View on GitHub↗

    Mechanize is a Ruby library for web browser automation and headless browser emulation. It allows for programmatically navigating websites and simulating human behavior without a graphical user interface. The library provides an automated interface for populating and submitting web forms, including text fields, checkboxes, and file uploads. It manages stateful sessions by automatically storing and sending cookies across multiple requests to maintain user authentication and identity. Additional capabilities include web data scraping, the ability to download remote web content, and the maintena

    Ruby
    View on GitHub↗4,443
  • felipecsl/wombatfelipecsl avatar

    felipecsl/wombat

    1,362View on GitHub↗

    Lightweight Ruby web crawler/scraper with an elegant DSL which extracts structured data from pages.

    Rubycrawlerdslruby
    View on GitHub↗1,362
  • mendableai/firecrawl-mcp-servermendableai avatar

    mendableai/firecrawl-mcp-server

    6,602View on GitHub↗

    This project is a Model Context Protocol server that connects large language models to web scraping and crawling tools. It functions as a bridge, allowing LLM clients to utilize a web crawling engine and scraping utilities to extract and process web data. The server integrates a markdown web converter that transforms dynamic web pages and PDF documents into clean markdown to optimize consumption by AI models. It also provides a browser automation interface for controlling headless sessions and bypassing access restrictions. The system covers broad capabilities including large-scale website d

    JavaScript
    View on GitHub↗6,602
See all 30 alternatives to Spidr→

Frequently asked questions

What does postmodern/spidr do?

A versatile Ruby web spidering library that can spider a site, multiple domains, certain links or infinitely. Spidr is designed to be fast and easy to use.

What are the main features of postmodern/spidr?

The main features of postmodern/spidr are: Web Crawling, Ruby Crawling Frameworks, Test Automation Frameworks.

What are some open-source alternatives to postmodern/spidr?

Open-source alternatives to postmodern/spidr include: sparklemotion/mechanize — Mechanize is a Ruby library for web browser automation and headless browser emulation. It allows for programmatically… felipecsl/wombat — Lightweight Ruby web crawler/scraper with an elegant DSL which extracts structured data from pages. propublica/upton — A batteries-included framework for easy web-scraping. Just add CSS! (Or do more.). mendableai/firecrawl-mcp-server — This project is a Model Context Protocol server that connects large language models to web scraping and crawling… lorien/web-scraping — This project is a comprehensive resource directory for web data extraction, providing a curated collection of tools… adithya-s-k/omniparse — Omniparse is a multimodal content parser and generative AI ingestion engine designed to convert documents, images, and…