awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoAcerca deCómo clasificamosPrensaServidor MCP
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
groupultra avatar

groupultra/telegram-search

0
View on GitHub↗
3,949 estrellas·259 forks·TypeScript·AGPL-3.0·4 vistassearch.lingogram.app↗

Telegram Search

Telegram Search is a self-hosted platform designed to export, index, and archive personal or group message history. It functions as a private search engine that transforms scattered communication logs and media assets into a searchable knowledge library, allowing users to maintain full control over their data through containerized infrastructure.

The platform distinguishes itself by utilizing vector-based semantic indexing to enable fuzzy retrieval across historical datasets. It incorporates an optical character recognition pipeline to extract text from images and media files, ensuring that visual content is as discoverable as text-based messages. Users can navigate directly from search results back to the original source within the messaging application using protocol-based deep links.

Beyond basic search, the system integrates generative artificial intelligence to provide context-aware summaries and answers based on stored chat logs. This retrieval-augmented generation capability allows for intelligent analysis of historical threads, while automated media archiving offloads heavy assets to external storage to maintain a lightweight local database. The entire stack is deployed via containerized configurations to simplify the management of the interface, database, and storage services.

Features

  • Telegram Chat Importers - Exports and preserves personal or group message history into a private database for long-term storage and offline access.
  • Retrieval Augmented Generation - Combines stored chat archives with large language models to provide context-aware answers and summaries based on private user data.
  • Semantic Vector Search - Performs fuzzy and vector-based searches across text and images to quickly locate specific information within archived communication history.
  • Vector Indexing - Converts message content into high-dimensional numerical embeddings to enable fast similarity searches across large historical datasets.
  • Chat Analysis - Analyzes historical and unread message threads to provide intelligent summaries and context-aware answers to user queries.
  • Local Chat Archiving - Provides a containerized infrastructure for backing up and organizing chat media and message history into local or remote storage.
  • Retrieval Augmented Generation Platforms - Analyzes stored chat logs to provide automated summaries and context-aware answers to user queries.
  • Vector Databases and Search - Utilizes embeddings to enable fuzzy and semantic retrieval of historical text and image content.
  • Personal Knowledge Management - Indexes chat content and media assets to turn scattered conversations into a searchable library of information and shared resources.
  • Chat History Synchronization - Synchronizes and stores message data into local or remote databases, processing content for vector-based search to ensure conversations remain accessible.
  • Chat Query Engines - Enables retrieval-augmented generation to answer questions using chat history as context, helping users manage communication more effectively.
  • Optical Character Recognition Indexers - Extracts embedded text from images and media files during the ingestion process to make visual content searchable by keyword.
  • Media Library Indexers - Automatically indexes shared media and links through text recognition and image description, while backing up files to secure storage.
  • Object Storage Persistence - Offloads heavy media assets to external storage services to maintain a lightweight local database while ensuring long-term data preservation.
  • Deep Linking Utilities - Provides direct protocol-based links from search results to the original message source for context verification.
  • Fuzzy File Content Searches - Performs multi-language fuzzy and semantic searches across text and images using vector embeddings to locate specific historical conversations.
  • Containerized Service Deployment - Supports containerized deployment of the interface, database, and storage services using a single configuration file to simplify private infrastructure hosting.
  • Multi-Service Container Orchestration - Packages the application, database, and storage components into isolated environments to simplify self-hosting and infrastructure management.
  • Self-Hosted Infrastructure - Deploys and manages private search and storage services using containerized configurations to maintain full control over data.

Historial de estrellas

Gráfico del historial de estrellas de groupultra/telegram-searchGráfico del historial de estrellas de groupultra/telegram-search

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Colecciones destacadas con Telegram Search

Colecciones seleccionadas manualmente donde aparece Telegram Search.
  • Bases de datos para historial de chat

Alternativas open-source a Telegram Search

Proyectos open-source similares, clasificados según cuántas características comparten con Telegram Search.
  • yichuan-w/leannAvatar de yichuan-w

    yichuan-w/LEANN

    11,985Ver en GitHub↗

    LEANN is a framework for local retrieval augmented generation and vector indexing. It functions as a system for building local knowledge bases and source code search engines that combine large language models with retrieved private data to generate context-aware responses. The project distinguishes itself through a vision-model based document layout extractor for parsing complex PDF figures and diagrams, and a source code search engine that employs structure-aware chunking to preserve function and class boundaries. It also implements the Model Context Protocol to integrate real-time data sour

    Pythonaifaissgpt-oss
    Ver en GitHub↗11,985
  • hellodigua/chatlabAvatar de hellodigua

    hellodigua/ChatLab

    4,522Ver en GitHub↗

    ChatLab is a self-hosted chat database and data pipeline designed to normalize, store, and analyze large-scale social conversation histories. It functions as an analytics platform that uses large language models to extract patterns and insights from messaging data imported from multiple platforms. The system distinguishes itself through an AI-powered analysis engine that utilizes vector-based history analysis and agent-based function calling to summarize conversation trends. It further identifies behavioral patterns by generating visual analytics, including heatmaps, word clouds, and activity

    TypeScriptaichat-analysischat-history
    Ver en GitHub↗4,522
  • awesome-selfhosted/awesome-selfhostedAvatar de awesome-selfhosted

    awesome-selfhosted/awesome-selfhosted

    299,516Ver en GitHub↗

    This project is a community-curated directory of open-source software designed for deployment in private server environments and home labs. It serves as a comprehensive resource for discovering independent, self-hosted alternatives to mainstream cloud services, enabling users to maintain full data ownership and control over their digital infrastructure. The directory is structured through a hierarchical taxonomy that organizes a vast collection of applications into logical categories, ranging from media management and data analytics to private communication and team productivity tools. It dis

    awesomeawesome-listcloud
    Ver en GitHub↗299,516
  • superduper-io/superduperAvatar de superduper-io

    superduper-io/superduper

    5,298Ver en GitHub↗

    Superduper is an AI agent development kit and LLM application framework designed to build autonomous agents and data-driven applications. It functions as a RAG orchestration platform and vector search infrastructure, coordinating AI models with database storage to perform multi-step computations and actions using persisted data states. The project distinguishes itself by providing a database-integrated machine learning pipeline that executes training and inference tasks directly on data hosted within SQL and NoSQL databases. It allows for the deployment of self-hosted AI infrastructure on pri

    Pythonaichatbotdata
    Ver en GitHub↗5,298
Ver las 30 alternativas a Telegram Search→

Preguntas frecuentes

¿Qué hace groupultra/telegram-search?

Telegram Search is a self-hosted platform designed to export, index, and archive personal or group message history. It functions as a private search engine that transforms scattered communication logs and media assets into a searchable knowledge library, allowing users to maintain full control over their data through containerized infrastructure.

¿Cuáles son las características principales de groupultra/telegram-search?

Las características principales de groupultra/telegram-search son: Telegram Chat Importers, Retrieval Augmented Generation, Semantic Vector Search, Vector Indexing, Chat Analysis, Local Chat Archiving, Retrieval Augmented Generation Platforms, Vector Databases and Search.

¿Qué alternativas de código abierto existen para groupultra/telegram-search?

Las alternativas de código abierto para groupultra/telegram-search incluyen: yichuan-w/leann — LEANN is a framework for local retrieval augmented generation and vector indexing. It functions as a system for… hellodigua/chatlab — ChatLab is a self-hosted chat database and data pipeline designed to normalize, store, and analyze large-scale social… awesome-selfhosted/awesome-selfhosted — This project is a community-curated directory of open-source software designed for deployment in private server… superduper-io/superduper — Superduper is an AI agent development kit and LLM application framework designed to build autonomous agents and… thinkany-ai/rag-search — This project provides a search service designed to retrieve and rerank web content for use in large language model… weaviate/weaviate — Weaviate is an AI-native vector database designed to store and index high-dimensional vector embeddings alongside…