1 Repo
Specialized execution backends for managing and running fine-tuning adapters.
Distinct from Performance Optimizations: Focuses on adapter-specific execution backends, distinct from general performance optimization.
Explore 1 awesome GitHub repository matching artificial intelligence & ml · Adapter Execution Backends. Refine with filters or upvote what's useful.
Sglang is a high-performance inference engine and serving system designed for large language and multimodal models. It provides a programmable interface for orchestrating complex generation workflows, enabling developers to coordinate multi-turn dialogues, tool invocations, and reasoning chains through a domain-specific language. The platform is built to support production-scale deployments, offering an OpenAI-compatible API that allows for integration with existing application ecosystems. The system distinguishes itself through a disaggregated architecture that separates compute-intensive pr
Balances compatibility and high-concurrency performance for adapter-heavy workloads by selecting specialized backends.