2 个仓库
Caching mechanisms that use vector-based semantic matching to retrieve AI model responses.
Distinct from Response Caching: Specializes response caching by using semantic similarity rather than exact key matches.
Explore 2 awesome GitHub repositories matching data & databases · Semantic Caching. Refine with filters or upvote what's useful.
Higress is an AI API gateway and cloud-native traffic manager that functions as a Kubernetes ingress controller. It provides a centralized system for routing, securing, and optimizing traffic directed toward large language models, AI agents, and microservice architectures. The project distinguishes itself through deep AI orchestration, including the ability to host and manage Model Context Protocol servers that transform REST APIs into tools for AI agents. It features specialized AI infrastructure for model request proxying, protocol translation across multiple providers, and semantic-based c
Stores model responses and dialogue context using semantic matching to reduce token usage and latency.
Caches responses for semantically similar queries using vector-based matching to reduce latency and cost.