awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
omnivore-app avatar

omnivore-app/omnivore

0
View on GitHub↗
15,882 stars·1,254 forks·TypeScript·agpl-3.0·29 viewsomnivore.work↗

Omnivore

Omnivore is an open-source, self-hostable read-it-later application designed to centralize web articles, newsletters, and digital documents into a personal library. It functions as a comprehensive content archiver that captures web pages and stores them locally, ensuring permanent access and readability regardless of internet connectivity.

The platform distinguishes itself through an event-sourced synchronization engine that maintains a consistent state across multiple devices by replaying user actions. It utilizes a headless web scraping service to extract clean text and metadata from raw web pages, providing a uniform reading experience. Users can manage their collections through a research-oriented workflow that supports highlighting passages and attaching personal notes to saved content.

The application provides a full suite of content management capabilities, including offline reading, cross-device progress synchronization, and structured data persistence. It is distributed as an open-source project, allowing users to maintain full control over their personal data and reading history.

Features

  • Read-It-Later Applications - Functions as a centralized platform for capturing web articles to read offline with support for annotations and cross-device synchronization.
  • Web Content Archivers - Provides a self-hosted platform for capturing, organizing, and archiving web articles and newsletters into a personal library.
  • Read-It-Later Platforms - Centralizes web articles and newsletters into a personal library for reading at your convenience.
  • Self-Hosted Applications - Allows users to host their own content library, maintaining full control over personal data.
  • Data Synchronization Engines - Uses an event-sourced engine to synchronize library state across devices.
  • Cross-Device Synchronization Engines - Synchronizes reading progress and library state across multiple devices.
  • Document Archiving Systems - Archives web pages and documents to local storage for permanent offline access.
  • Offline Caching - Enables offline access to saved articles and documents by persisting content locally.
  • Self-Hosted - Enables users to maintain a personal, self-hosted reading list that synchronizes progress and highlights across devices.
  • Event Sourcing - Maintains library state across devices by replaying an immutable log of user actions.
  • Feed Generators - Open-source read-it-later app that supports RSS feeds.
  • Media & Communication - Read-it-later service with offline support.
  • Web Scraping - Provides a headless scraping service to extract clean, readable content from raw web pages.
  • Document Annotators - Supports highlighting and attaching personal notes to saved content for research and review.
  • Data Annotation Workflows - Facilitates research workflows by allowing users to mark passages and attach notes to saved content.
  • Note Taking Tools - Provides tools for capturing and organizing personal insights and notes within saved documents.
  • Optimistic UI Updates - Provides immediate UI feedback for user actions like highlighting or archiving before server confirmation.
  • Content-Addressable Storage - Uses cryptographic hashing to store and deduplicate web content, ensuring data integrity within the local library.
  • Relational Database Persistence - Uses relational database storage to manage complex relationships between articles, annotations, and metadata.

Star history

Star history chart for omnivore-app/omnivoreStar history chart for omnivore-app/omnivore

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does omnivore-app/omnivore do?

Omnivore is an open-source, self-hostable read-it-later application designed to centralize web articles, newsletters, and digital documents into a personal library. It functions as a comprehensive content archiver that captures web pages and stores them locally, ensuring permanent access and readability regardless of internet connectivity.

What are the main features of omnivore-app/omnivore?

The main features of omnivore-app/omnivore are: Read-It-Later Applications, Web Content Archivers, Read-It-Later Platforms, Self-Hosted Applications, Data Synchronization Engines, Cross-Device Synchronization Engines, Document Archiving Systems, Offline Caching.

What are some open-source alternatives to omnivore-app/omnivore?

Open-source alternatives to omnivore-app/omnivore include: awesome-selfhosted/awesome-selfhosted — This project is a community-curated directory of open-source software designed for deployment in private server… wallabag/wallabag — Wallabag is a self-hosted, open-source bookmark manager designed to archive web content for later reading. It… kovidgoyal/calibre — Calibre is a comprehensive suite for digital library management, serving as a centralized hub for organizing,… keygen-sh/keygen-api — Keygen is a software licensing and distribution API platform designed to validate product keys, manage user… kenshin/simpread — Simpread is a browser extension that transforms any web page into a clean, distraction-free reading layout optimized… karakeep-app/karakeep — Karakeep is a self-hosted, open-source platform designed for personal knowledge management and web content archiving.…

Open-source alternatives to Omnivore

Similar open-source projects, ranked by how many features they share with Omnivore.
  • awesome-selfhosted/awesome-selfhostedawesome-selfhosted avatar

    awesome-selfhosted/awesome-selfhosted

    299,516View on GitHub↗

    This project is a community-curated directory of open-source software designed for deployment in private server environments and home labs. It serves as a comprehensive resource for discovering independent, self-hosted alternatives to mainstream cloud services, enabling users to maintain full data ownership and control over their digital infrastructure. The directory is structured through a hierarchical taxonomy that organizes a vast collection of applications into logical categories, ranging from media management and data analytics to private communication and team productivity tools. It dis

    awesomeawesome-listcloud
    View on GitHub↗299,516
  • wallabag/wallabagwallabag avatar

    wallabag/wallabag

    12,777View on GitHub↗

    Wallabag is a self-hosted, open-source bookmark manager designed to archive web content for later reading. It functions as a personal knowledge management tool, allowing users to collect, store, and organize web pages into a centralized, searchable library. The platform provides a distraction-free reading experience by extracting the primary text and images from web pages while removing advertisements and navigation menus. This process ensures that saved articles remain accessible for offline reading, preserving the content even if the original source is removed from the internet. The system

    PHPhacktoberfestphpread-it-later
    View on GitHub↗12,777
  • kovidgoyal/calibrekovidgoyal avatar

    kovidgoyal/calibre

    24,146View on GitHub↗

    Calibre is a comprehensive suite for digital library management, serving as a centralized hub for organizing, converting, and editing e-book collections. It functions as a multi-purpose platform that combines a relational database for metadata tracking with a powerful processing engine capable of transforming document formats and restructuring internal markup. Beyond local management, the software acts as a content server, enabling users to host their libraries over a network for remote access and reading via standard web browsers. The project distinguishes itself through its deep extensibili

    Pythoncalibreebookebook-formats
    View on GitHub↗24,146
  • keygen-sh/keygen-apikeygen-sh avatar

    keygen-sh/keygen-api

    1,507View on GitHub↗

    Keygen is a software licensing and distribution API platform designed to validate product keys, manage user entitlements, and enforce device activation policies. It provides a comprehensive set of identity and licensing services that enable software vendors to control access to desktop and server applications while tracking usage limits and seat counts across hardware machines. The platform includes enterprise identity provider integration using standard SAML protocols, single sign-on authentication, and role-based access control with granular permission restrictions to secure system operatio

    Gherkinapifair-sourcelicense-keys
    View on GitHub↗1,507
See all 30 alternatives to Omnivore→