awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
intelligent-machine-learning avatar

intelligent-machine-learning/glakeArchived

0
View on GitHub↗
502 stars·44 forks·Python·Apache-2.0·8 views

Glake

GLake: optimizing GPU memory management and IO transmission.

Features

  • Inference Serving Engines - Flexible virtual tensor management for efficient serving.

Star history

Star history chart for intelligent-machine-learning/glakeStar history chart for intelligent-machine-learning/glake

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Glake

Similar open-source projects, ranked by how many features they share with Glake.
  • hsword/spotserveHsword avatar

    Hsword/SpotServe

    134View on GitHub↗

    SpotServe: Serving Generative Large Language Models on Preemptible Instances

    View on GitHub↗134
  • microsoft/vattentionmicrosoft avatar

    microsoft/vattention

    495View on GitHub↗

    Dynamic Memory Management for Serving LLMs without PagedAttention

    C
    View on GitHub↗495
  • nvidia/tensorrt-llmNVIDIA avatar

    NVIDIA/TensorRT-LLM

    12,913View on GitHub↗

    TensorRT-LLM is a platform and toolkit designed for compiling, optimizing, and serving transformer-based models on accelerated hardware. It functions as a framework that transforms machine learning models into efficient execution graphs, providing an engine to refine these models for specific hardware to maximize throughput and minimize latency during text generation. The project distinguishes itself through advanced execution strategies that manage the entire inference pipeline. It utilizes kernel-level fusion and static graph execution to optimize mathematical operations and computational f

    Pythonblackwellcudallm-serving
    View on GitHub↗12,913
  • rulinshao/lightseqRulinShao avatar

    RulinShao/LightSeq

    223View on GitHub↗

    Official repository for DistFlashAttn: Distributed Memory-efficient Attention for Long-context LLMs Training

    Python
    View on GitHub↗223

Frequently asked questions

What does intelligent-machine-learning/glake do?

GLake: optimizing GPU memory management and IO transmission.

What are the main features of intelligent-machine-learning/glake?

The main features of intelligent-machine-learning/glake are: Inference Serving Engines.

What are some open-source alternatives to intelligent-machine-learning/glake?

Open-source alternatives to intelligent-machine-learning/glake include: hsword/spotserve — SpotServe: Serving Generative Large Language Models on Preemptible Instances. microsoft/vattention — Dynamic Memory Management for Serving LLMs without PagedAttention. nvidia/tensorrt-llm — TensorRT-LLM is a platform and toolkit designed for compiling, optimizing, and serving transformer-based models on… rulinshao/lightseq — Official repository for DistFlashAttn: Distributed Memory-efficient Attention for Long-context LLMs Training.