awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
searxng avatar

searxng/searxng-docker

0
View on GitHub↗
3,157 stars·665 forks·agpl-3.0·9 vues

Searxng Docker

This project is a containerized search infrastructure designed to deploy a privacy-focused metasearch engine. It acts as a self-hosted search proxy that aggregates results from multiple external web, image, and academic search providers while anonymizing requests and stripping trackers to protect user identity.

The system utilizes Docker to orchestrate the search instance, integrating caching mechanisms and reverse proxy support to ensure a private and efficient search environment. It employs a modular adapter-based integration to standardize diverse external API responses and a processing pipeline to clean and format data.

The software covers a wide range of capability areas, including advanced search aggregation across niche and general domains, bot defense with sliding-window rate limiting, and comprehensive result sanitization. It also provides tools for localization, custom plugin development, and the integration of key-value databases for search performance optimization.

Deployment is facilitated through automated installation scripts and support for various application gateways and SSL reverse proxy configurations.

Features

  • Metasearch Engines - Operates as a privacy-focused metasearch engine that aggregates results from multiple external providers into one interface.
  • Privacy Proxies - Acts as a privacy-preserving intermediary between users and search engines to block tracking and profiling.
  • Data Caching - Implements a persistent local cache using SQLite to store key-value pairs and binary blobs.
  • Search Result Formatters - Processes raw search engine responses through a pipeline of formatters to clean and standardize the data.
  • Aggregated Result Templates - Uses specialized HTML templates to render different media types consistently in the aggregated results list.
  • Standardized Result Schemas - Structures diverse data from multiple external engines into a consistent internal format for titles and snippets.
  • Search Configuration - Provides centralized configuration for adjusting engine sources and UI preferences to customize result retrieval.
  • Category-Based Filters - Groups search engines into functional tabs to filter results by content type.
  • Privacy-Based Cleaning - Removes tracking URLs and filters domains from search results to increase user privacy and result relevance.
  • Key-Value - Utilizes a Valkey database for high-performance persistent storage of key-value pairs.
  • Search Engine Configuration - Defines categories, timeouts, API keys, and proxy settings for each integrated external search provider.
  • Response Field Mapping - Uses JSON configurations to map external API response fields into standardized internal search result formats.
  • Search Query Utilities - Transforms raw user input into formatted request parameters required by various external search engines.
  • Search Result Aggregators - Simulates browser requests to fetch results from multiple general external engines including pagination and language headers.
  • Docker Container Deployments - Packages the metasearch engine and its dependencies into Docker images for consistent self-hosted deployment.
  • Containerized Deployment Orchestration - Provides a containerized deployment orchestration setup using Docker to manage the search engine and its dependencies.
  • Response Parsing Utilities - Extracts data from diverse external HTTP responses and converts it into structured lists of search results.
  • Anonymizing Proxies - Acts as an anonymizing proxy that strips tracking and masks user identity when querying search providers.
  • Identity Masking Proxies - Routes outgoing requests through SOCKS or HTTP proxies to hide the server IP from external search providers.
  • Tracking Parameter Removers - Strips identifying query arguments and tracking parameters from search result URLs to protect user privacy.
  • Bot Blocking - Limits the number of requests to protect the search instance from automated bot traffic.
  • Bot Detection - Analyzes User-Agent and Connection headers to identify and block automated bot traffic.
  • Private Search Engines - Implements a self-hosted, containerized search platform that prioritizes user privacy and avoids data collection.
  • Privacy-Focused Search Engines - Provides a search aggregator that anonymizes requests and blocks trackers to ensure user privacy.
  • Privacy Image Proxies - Masks the user's IP address by routing external image requests through a privacy proxy.
  • Inbound - Identifies automated requests and manages rate limits to protect the server from abuse.
  • Privacy Log Suppression - Prevents the recording of request history and user IPs by redirecting server logs to a null device.
  • Search Aggregators - Queries multiple external search services simultaneously and normalizes their outputs into a single unified format.
  • Answer Engines - Provides specialized modules that deliver immediate, concise answers to specific queries.
  • Instant Answer Generators - Provides plugins that generate instant, direct answers to specific queries at the top of search results.
  • Biomedical - Retrieves factual responses and citations for medical queries from MEDLINE and life science journals.
  • Image and Gallery - Displays image search results in a thumbnail gallery with detailed views for resolution and format.
  • Search Result Layouts - Customizes the visual layout, theme, and navigation behavior of the search results page.
  • Bot Protection - Mitigates automated abuse and spam by capping the total number of requests the instance accepts.
  • Distributed Caching - Integrates with remote key-value stores to implement distributed data caching for improved retrieval speed.
  • Local Data Engines - Executes search queries against locally hosted databases to enable offline information retrieval.
  • Search Engine Plugins - Supports the development of custom search plugins to modify requests and process results within the search pipeline.
  • Search Domains - Provides a framework for building specialized search plugins and API integrations for niche or local data sources.
  • Search Engine Selectors - Allows users to restrict searches to a specific engine or category using a specialized prefix syntax.
  • Utility Computation Extensions - Integrates specialized tools such as calculators and unit converters to enhance search output quality.
  • Search Result Aggregators - Retrieves and normalizes search data from Bing to unify it within a privacy-focused interface.
  • Scholarly Search Aggregators - Fetches bibliographic data and research literature specifically from the arXiv API.
  • Search Result Exporters - Converts aggregated search results into structured formats like JSON or RSS for use in external applications.
  • Search Result Filtering - Limits search results to a specific language by applying a language filter prefix.
  • Engine Constraints - Restricts search results to specific engines or categories using a prefix modifier.
  • Search Suggestions - Suggests alternative search terms to help users refine or redirect their queries.
  • Multilingual Search Engines - Creates multiple engine instances with different language codes to retrieve results in several languages.
  • Search - Stores aggregated search results in a key-value database with expiration times to improve performance and reduce provider load.
  • Search Result Caches - Stores retrieved search results in a key-value store to reduce latency and upstream requests.
  • String Localization - Handles translation strings and locale mappings to ensure the interface is served in the user's preferred language.
  • Research Paper Indexes - Aggregates and retrieves scholarly articles and research papers from global academic indexes.
  • Reverse Image Search Tools - Allows users to find original sources or similar content by analyzing uploaded or linked images via external services.
  • Search Language Configuration - Defines the default language for queries and restricts available languages for the user.
  • Targeted Search Parameters - Defines custom search parameters and sorting methods to create targeted shortcuts for specific content types.
  • Search Utility Keywords - Provides instant utility triggers for tasks like generating UUIDs or performing quick calculations.
  • Plugin Extensibility - Features a plugin architecture that allows adding specialized functional extensions to the search process.
  • Search Infrastructure Management - Manages the deployment and maintenance of search instances using Docker and reverse proxy configurations.
  • SSL Termination Proxies - Integrates with NGINX or Apache to handle SSL termination and secure incoming HTTP requests.
  • Instant Fact Retrieval - Fetches immediate answers to common queries through internal data sources to provide direct results.
  • Video Search Aggregators - Aggregates and displays video content from decentralized hosting platforms.
  • Outgoing Identity Control - Specifies source IP addresses and user-agent suffixes to manage network interfaces and prevent blocking by search engines.
  • IP Address Filters - Implements mechanisms to block or allow search requests based on predefined IPv4 and IPv6 address lists.
  • Network Traffic Proxying - Routes outgoing requests through SOCKS or HTTP proxies using round-robin distribution to mask the server identity.
  • Bot Detection Bypass - Implements techniques to mimic human browser behavior and session tokens to avoid CAPTCHAs when retrieving data from search engines.
  • Academic Publisher API Integration - Retrieves academic papers and research data from a global publisher using a dedicated API.
  • Mathematical Problem Solving Toolkits - Parses and calculates mathematical expressions entered by the user directly within the search interface.
  • Measurement Unit Conversions - Transforms measured values between different systems of measurement based on query symbols.
  • Academic Search Engines - Provides specialized search capabilities for locating scholarly articles, academic works, and research literature.
  • Bot Challenge Verifications - Verifies legitimate browser sessions using token exchanges via CSS resources to filter suspicious requests.
  • Bot Detection Systems - Detects CAPTCHAs and block pages to identify when search engines are rate-limiting the instance.
  • Tor Routing - Routes outgoing search requests through the Tor network to hide the user's identity.
  • Request Rate Limiting - SearXNG implements a limiter to control the volume of incoming requests and protect against abuse.
  • Environment Variable-Based Configuration - Allows overriding application settings through a hierarchy of YAML files and environment variables for flexible deployment.
  • Integration Adapters - Implements an architectural abstraction layer using modular adapters to normalize diverse external search API responses.
  • Processing Pipelines - Processes raw search engine data through a sequential pipeline of plugins and filters to clean and format the output.
  • Search Result Processing Pipelines - Implements a modular pipeline of plugins to process and modify aggregated search results.
  • Provider Suspension Management - Automatically stops requests to search engines that return CAPTCHAs or rate-limit errors.
  • IP-Based Rate Limiting - Implements sliding-window rate limiting per IP address to prevent bot abuse and avoid provider blacklisting.
  • Typo Correction Suggestions - Identifies potential typos and suggests corrected search queries to the user.
  • Service Engine Health Monitors - Suspends access to external search engines for a specified duration after detecting errors or rate limits.
  • Search Keyword Suggesters - Provides real-time search query suggestions by routing requests to backend autocomplete providers.
  • Branding Customization - Provides administrative settings to customize the visual identity, colors, and branding of the search interface.
  • Direct Answer Displays - Renders concise information such as weather forecasts and translations directly in the results interface.
  • Infinite Scrolling - Automatically loads the next page of search results as the user scrolls to the bottom of the page.
  • Interface Text Localization - Manages the translation of user interface strings via an external platform to support multiple languages.
  • Sliding Window Counters - Tracks request frequency over rolling time intervals using sliding windows to identify and block abusive IPs.
  • API Query Interfaces - Provides a programmatic interface to execute search queries and retrieve results in JSON, CSV, or RSS formats.
  • User Agent Generators - Produces randomized or specific identification strings to mimic various clients and devices during network requests.
  • Performance Optimizations - Uses key-value caching with Valkey or SQLite to reduce upstream requests and improve response times.
  • Template-Driven Rendering - Uses specialized server-side HTML templates to render different media types such as images and academic papers.
  • Web Scraping - Uses XPath expressions to scrape and parse structured data from websites that do not provide native APIs.

Historique des stars

Graphique de l'historique des stars pour searxng/searxng-dockerGraphique de l'historique des stars pour searxng/searxng-docker

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Start searching with AI

Alternatives open source à Searxng Docker

Projets open source similaires, classés selon le nombre de fonctionnalités partagées avec Searxng Docker.
  • searxng/searxngAvatar de searxng

    searxng/searxng

    32,180Voir sur GitHub↗

    This project is a privacy-focused, self-hosted metasearch engine that aggregates results from a wide array of web, academic, and media sources into a single, unified interface. By acting as a proxy between the user and external search providers, it strips identifying headers and tracking parameters from requests, ensuring that search activity remains anonymous and protected from third-party profiling. The platform distinguishes itself through a modular, plugin-based architecture that allows for extensive customization of search behavior, result filtering, and interface branding. It supports a

    Pythonbingbravedegoogle
    Voir sur GitHub↗32,180
  • searx/searxAvatar de searx

    searx/searx

    13,513Voir sur GitHub↗

    Searx is a privacy-respecting metasearch engine and search result aggregator. It functions as a self-hosted search proxy that queries diverse web services, databases, and local indices to present a single unified list of results. The project prevents user tracking and profiling by acting as an intermediary between the client and search services. It strips identifying information from queries, removes tracker URLs and HTTP referrers from outgoing links, and can route traffic through proxies or the Tor network to mask user identity. The system supports multilingual search and result filtering

    Python
    Voir sur GitHub↗13,513
  • deedy5/ddgsAvatar de deedy5

    deedy5/ddgs

    2,754Voir sur GitHub↗

    ddgs is a metasearch engine and web content extractor that provides a toolkit for programmatically retrieving search results from DuckDuckGo. It functions as a search API server and a Model Context Protocol server to integrate web search capabilities directly into large language model environments. The project distinguishes itself by aggregating text, image, news, and video results from multiple providers into a single interface. It includes a utility for fetching URLs and converting HTML content into markdown, plain text, or structured data. The system covers a broad range of search capabil

    Pythonapiddgsmcp
    Voir sur GitHub↗2,754
  • facebook/rocksdbAvatar de facebook

    facebook/rocksdb

    31,767Voir sur GitHub↗

    RocksDB is a high-performance, embeddable persistent key-value library and storage engine based on Log-Structured Merge-trees. It is designed to provide durable storage for large-scale datasets, integrating directly into applications to manage data on flash and RAM-based hardware. The engine is distinguished by its focus on minimizing read and write amplification through multi-threaded compaction and custom memory allocators. It features specialized optimizations for flash storage, including support for zoned block devices, and provides the ability to extend store behavior via external plugin

    C++databasestorage-engine
    Voir sur GitHub↗31,767
Voir les 30 alternatives à Searxng Docker→

Questions fréquentes

Que fait searxng/searxng-docker ?

This project is a containerized search infrastructure designed to deploy a privacy-focused metasearch engine. It acts as a self-hosted search proxy that aggregates results from multiple external web, image, and academic search providers while anonymizing requests and stripping trackers to protect user identity.

Quelles sont les fonctionnalités principales de searxng/searxng-docker ?

Les fonctionnalités principales de searxng/searxng-docker sont : Metasearch Engines, Privacy Proxies, Data Caching, Search Result Formatters, Aggregated Result Templates, Standardized Result Schemas, Search Configuration, Category-Based Filters.

Quelles sont les alternatives open-source à searxng/searxng-docker ?

Les alternatives open-source à searxng/searxng-docker incluent : searxng/searxng — This project is a privacy-focused, self-hosted metasearch engine that aggregates results from a wide array of web,… searx/searx — Searx is a privacy-respecting metasearch engine and search result aggregator. It functions as a self-hosted search… deedy5/ddgs — ddgs is a metasearch engine and web content extractor that provides a toolkit for programmatically retrieving search… facebook/rocksdb — RocksDB is a high-performance, embeddable persistent key-value library and storage engine based on Log-Structured… chaitin/safeline — SafeLine is a containerized web application firewall and reverse proxy designed to secure web services by inspecting… pagefind/pagefind — Pagefind is a static site search engine that indexes HTML files to provide a browser-based search experience without…