awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
dosyago avatar

dosyago/dn

0
View on GitHub↗
3,905 stars·148 forks·JavaScript·18 viewsdosaygo.com/downloadnet↗

Dn

dn is a local browser data archive and web archiver designed to save and render web pages from Chromium browsers for offline viewing and permanent storage. It functions as a self-hosted repository for browsing history and page content, operating as an offline web content server that hosts saved data as if the original sites were still online.

The system includes a full-text search engine that indexes all saved web pages, enabling the instant recovery of specific information across the local collection. It utilizes a domain-based filtering system to block specific website addresses from being archived via a blacklist.

The project covers local content serving through Chromium-based page rendering and file-system web mirroring to maintain original visual layouts. It also provides tools for managing system resources, such as allocating storage and memory limits for the archive.

Features

  • Web Page Archiving - Captures and saves complete versions of web pages to create a permanent local archive.
  • Data Extraction - Captures and stores browsing history and page content specifically from Chromium-based browsers.
  • Browser Data Archives - Provides a self-hosted repository for storing browsing history and page content without relying on cloud services.
  • Local Content Viewers - Serves downloaded digital content via a local web server for seamless browser-based viewing.
  • Full-Text Search Engines - Indexes large volumes of local web content to provide rapid full-text retrieval.
  • Full-Text Search Indexes - Creates full-text search indexes across all archived pages for instant information recovery.
  • Local-First Storage - Prioritizes local device storage for all archived web data to ensure permanent offline availability.
  • Full Text Indexing - Processes and stores textual content from archived pages to enable fast keyword retrieval.
  • Web Page Rendering - Implements the process of interpreting HTML and CSS to visually render saved web pages locally.
  • Offline Content Servers - Hosts and renders saved website data on a local server to preserve the original online browsing experience.
  • Offline Page Serving - Renders saved pages through a local server so they can be browsed as if the original sites were online.
  • Offline Web Page Archivers - Saves complete web pages as local files to facilitate reading and research without an internet connection.
  • Directory Structure Mirroring - Organizes archived web assets in a local directory structure that mimics the original website hierarchy.
  • Domain Filtering - Provides a blacklist system to block specific website domains from being captured in the archive.
  • Browser Applications - Archiving and indexing tool for offline browsing.

Star history

Star history chart for dosyago/dnStar history chart for dosyago/dn

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Dn

Similar open-source projects, ranked by how many features they share with Dn.
  • do-say-go/dnDO-SAY-GO avatar

    DO-SAY-GO/dn

    3,905View on GitHub↗

    dn is a self-hosted personal web archiving system that automatically intercepts and stores web pages on a local device. It uses a proxy-based request interception model to capture browser traffic and save content for offline access without an internet connection. The system features a local full-text search engine that indexes all saved page content for information retrieval across the collection. It includes a dedicated browser interface that simulates online connectivity to serve archived files, mimicking the original live web environment. Administrative control is provided through a web-b

    JavaScriptarchivearchiverdisk
    View on GitHub↗3,905
  • tagspaces/tagspacestagspaces avatar

    tagspaces/tagspaces

    4,935View on GitHub↗

    TagSpaces is an offline-first file tagging and organization platform that lets you manage local files with portable metadata stored directly in filenames or sidecar JSON files, eliminating the need for a central database. It functions as a full-text file search engine, a Kanban board file organizer, a local AI file assistant, an S3-compatible cloud file manager, and a web clipper and bookmark manager, all within a single application. The project distinguishes itself through a local-first architecture where all file operations, indexing, and AI processing run entirely on the device, with cloud

    TypeScriptelectronjavascriptnote-taking
    View on GitHub↗4,935
  • y2z/monolithY2Z avatar

    Y2Z/monolith

    15,283View on GitHub↗

    Monolith is a single-file HTML web archiver and asset bundler. It functions as a command-line interface and programmatic library designed to save complete web pages and their associated assets into a single HTML file for offline viewing. The tool crawls URLs to discover and fetch linked stylesheets, scripts, and images, which are then embedded into the document as data URLs. It includes capabilities for session injection via external cookie files and authentication handling to backup protected or member-only content. The project covers broader functional areas including automated web scrapin

    Rustcome-and-take-ite-hoardingits-mine
    View on GitHub↗15,283
  • go-ego/riotgo-ego avatar

    go-ego/riot

    6,059View on GitHub↗

    Riot is a Go-based distributed search engine and indexing server designed for full-text indexing and retrieval. It functions as a retrieval system that sorts documents by relevance using BM25 ranking algorithms, term frequency, and inverse document frequency. The engine provides specialized support for the Chinese language, featuring concurrent text segmentation and phonetic Pinyin mapping to match romanized input with characters. It utilizes a distributed architecture that employs hash-based index sharding to balance data load and throughput across multiple server nodes. The system covers a

    Gogogolanggwk
    View on GitHub↗6,059
See all 30 alternatives to Dn→

Frequently asked questions

What does dosyago/dn do?

dn is a local browser data archive and web archiver designed to save and render web pages from Chromium browsers for offline viewing and permanent storage. It functions as a self-hosted repository for browsing history and page content, operating as an offline web content server that hosts saved data as if the original sites were still online.

What are the main features of dosyago/dn?

The main features of dosyago/dn are: Web Page Archiving, Data Extraction, Browser Data Archives, Local Content Viewers, Full-Text Search Engines, Full-Text Search Indexes, Local-First Storage, Full Text Indexing.

What are some open-source alternatives to dosyago/dn?

Open-source alternatives to dosyago/dn include: do-say-go/dn — dn is a self-hosted personal web archiving system that automatically intercepts and stores web pages on a local… tagspaces/tagspaces — TagSpaces is an offline-first file tagging and organization platform that lets you manage local files with portable… y2z/monolith — Monolith is a single-file HTML web archiver and asset bundler. It functions as a command-line interface and… huichen/wukong — Wukong is a distributed full-text search engine designed for indexing and retrieving text documents. It functions as a… go-ego/riot — Riot is a Go-based distributed search engine and indexing server designed for full-text indexing and retrieval. It… obsidianmd/obsidian-clipper — This project is a markdown web clipper and local-first web archiver. It functions as a browser extension that extracts…