awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Y2Z avatar

Y2Z/monolith

0
View on GitHub↗
15,283 stars·459 forks·Rust·CC0-1.0·18 viewscrates.io/crates/monolith↗

Monolith

Monolith is a single-file HTML web archiver and asset bundler. It functions as a command-line interface and programmatic library designed to save complete web pages and their associated assets into a single HTML file for offline viewing.

The tool crawls URLs to discover and fetch linked stylesheets, scripts, and images, which are then embedded into the document as data URLs. It includes capabilities for session injection via external cookie files and authentication handling to backup protected or member-only content.

The project covers broader functional areas including automated web scraping, asset filtering by domain or resource type, and network request customization. It also supports proxy routing through environment variables to manage network restrictions.

Features

  • Single-File Web Archivers - Provides the core capability of saving complete web pages and their assets into a single HTML file for offline viewing.
  • Web Page Archiving - Saving complete websites as single HTML files to preserve content and styling for permanent offline access.
  • Offline Single-File Applications - Bundles complete web pages and assets into a single HTML or MHTML file for offline viewing.
  • Web Content Extraction Utilities - Implements a programmatic interface for collecting web content and media into single files for offline use.
  • Base64 Asset Embedding - Converts external page assets into base64 encoded data URIs for embedding directly within the HTML file.
  • DOM Resource Crawlers - Traverses the page DOM recursively to find and fetch all associated stylesheets, scripts, and images.
  • HTML Asset Bundlers - Embeds images, styles, and scripts into HTML files using data URLs to preserve page layout offline.
  • Offline Web Page Archivers - Saves complete web pages as local files by crawling URLs and combining all content into one document.
  • Single-File Distributions - Bundles page source and assets into a single standalone HTML file for offline distribution.
  • Asset Type Filters - Allows users to exclude specific resource types, such as video or scripts, from the final bundle.
  • Dynamic to Static Conversion - Transforms dynamic web pages into portable HTML files by embedding all assets as data URLs.
  • Resource Domain Filters - Allows restricting asset retrieval to specific domains to control which external resources are bundled.
  • Authenticated Content Backups - Preserves private or member-only web data by handling session cookies during the archiving process.
  • Domain-Restricted Crawling - Limits the crawling process to a specific root domain or block-list to control asset inclusion.
  • Page Bundling Automation - Provides programmatic and command-line interfaces to automate the bundling of multiple URLs.
  • Authentication Handling - Includes mechanisms for managing user identity and session cookies to access protected or member-only web content.
  • Cookie-Based Authentication Bridges - Enables the use of external cookie files to inject authenticated sessions for capturing protected content.
  • Web Scraping and Automation - Programmatically captures full page snapshots and assets from multiple URLs at scale.

Star history

Star history chart for y2z/monolithStar history chart for y2z/monolith

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Monolith

Similar open-source projects, ranked by how many features they share with Monolith.
  • obsidianmd/obsidian-clipperobsidianmd avatar

    obsidianmd/obsidian-clipper

    3,032View on GitHub↗

    This project is a markdown web clipper and local-first web archiver. It functions as a browser extension that extracts web page content and highlights, saving them as structured markdown files for personal knowledge management and long-term preservation. The utility acts as a template-based content extractor, transforming raw website data into formatted notes. It uses custom variables and processing filters to organize how captured information is structured before it is sent to a local directory.

    TypeScriptbravebrowser-extensionchrome
    View on GitHub↗3,032
  • danburzo/percollatedanburzo avatar

    danburzo/percollate

    4,647View on GitHub↗

    Percollate is a command-line tool for converting web pages and RSS feeds into structured files. It functions as a web content converter, static document generator, and page bundler that transforms online content into PDF, EPUB, HTML, or Markdown formats. The tool creates self-contained documents by embedding external images as encoded data URLs and applying custom HTML templates and CSS stylesheets. It can combine multiple web URLs or feed entries into a single digital book featuring a generated table of contents and hyperlinked index. Additional capabilities include the decomposition of Ato

    JavaScript
    View on GitHub↗4,647
  • dosyago/dndosyago avatar

    dosyago/dn

    3,905View on GitHub↗

    dn is a local browser data archive and web archiver designed to save and render web pages from Chromium browsers for offline viewing and permanent storage. It functions as a self-hosted repository for browsing history and page content, operating as an offline web content server that hosts saved data as if the original sites were still online. The system includes a full-text search engine that indexes all saved web pages, enabling the instant recovery of specific information across the local collection. It utilizes a domain-based filtering system to block specific website addresses from being

    JavaScript
    View on GitHub↗3,905
  • do-say-go/dnDO-SAY-GO avatar

    DO-SAY-GO/dn

    3,905View on GitHub↗

    dn is a self-hosted personal web archiving system that automatically intercepts and stores web pages on a local device. It uses a proxy-based request interception model to capture browser traffic and save content for offline access without an internet connection. The system features a local full-text search engine that indexes all saved page content for information retrieval across the collection. It includes a dedicated browser interface that simulates online connectivity to serve archived files, mimicking the original live web environment. Administrative control is provided through a web-b

    JavaScriptarchivearchiverdisk
    View on GitHub↗3,905
See all 30 alternatives to Monolith→

Frequently asked questions

What does y2z/monolith do?

Monolith is a single-file HTML web archiver and asset bundler. It functions as a command-line interface and programmatic library designed to save complete web pages and their associated assets into a single HTML file for offline viewing.

What are the main features of y2z/monolith?

The main features of y2z/monolith are: Single-File Web Archivers, Web Page Archiving, Offline Single-File Applications, Web Content Extraction Utilities, Base64 Asset Embedding, DOM Resource Crawlers, HTML Asset Bundlers, Offline Web Page Archivers.

What are some open-source alternatives to y2z/monolith?

Open-source alternatives to y2z/monolith include: obsidianmd/obsidian-clipper — This project is a markdown web clipper and local-first web archiver. It functions as a browser extension that extracts… danburzo/percollate — Percollate is a command-line tool for converting web pages and RSS feeds into structured files. It functions as a web… dosyago/dn — dn is a local browser data archive and web archiver designed to save and render web pages from Chromium browsers for… do-say-go/dn — dn is a self-hosted personal web archiving system that automatically intercepts and stores web pages on a local… tagspaces/tagspaces — TagSpaces is an offline-first file tagging and organization platform that lets you manage local files with portable… spyglass-search/spyglass — Spyglass is a local-first information retrieval system designed to aggregate, index, and search personal data and web…