awesome-repositories.com
博客
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目关于排名机制媒体报道MCP 服务器
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
cortexlabs avatar

cortexlabs/cortex

0
View on GitHub↗
8,013 星标·595 分支·Go·Apache-2.0·5 次浏览cortexlabs.com↗

Cortex

Cortex is a Kubernetes-based machine learning infrastructure platform designed for deploying, scaling, and managing models and workloads. It functions as a serverless inference engine and GPU cluster orchestrator, providing the tools necessary to execute real-time, asynchronous, and batch model predictions.

The platform utilizes declarative infrastructure-as-code for provisioning model clusters and environments. It optimizes operational costs by elastically scaling CPU and GPU resources through the use of spot instances.

The system covers a broad set of operational capabilities, including workload orchestration, private cloud network isolation with integrated identity management, and observability pipelines that stream logs and performance metrics to external monitoring tools.

Features

  • Production Serving Infrastructure - Deploys and serves machine learning models in production environments with scalable infrastructure and automated settings.
  • Serverless Inference Engines - Provides a serverless inference engine that automatically scales real-time, asynchronous, and batch model predictions.
  • GPU Resource Scaling - Dynamically adjusts GPU compute capacity using spot instances to balance performance and operational costs.
  • GPU Resource Orchestrators - Elastically provisions and optimizes CPU and GPU resources using spot instances for AI workloads.
  • Kubernetes ML Platforms - Provides a production platform for deploying, scaling, and managing machine learning models and workloads on Kubernetes.
  • ML Infrastructure Managers - Automates the provisioning and scaling of CPU and GPU compute clusters for large-scale ML workloads.
  • ML Orchestration Deployments - Orchestrates the deployment and scaling of machine learning models across production infrastructure to handle traffic loads.
  • Model Inference Clusters - Provisions specialized infrastructure and environment settings specifically for serving machine learning models.
  • Serverless Inference Engines - Executes real-time or batch model predictions that scale automatically based on request volume or queue length.
  • Workload Orchestration - Orchestrates real-time and batch processes that scale automatically based on request volume or queue length.
  • Private AI Deployments - Deploys machine learning workloads on private infrastructure to ensure data security and access control.
  • Asynchronous Task Processing - Provides a queued system for executing non-real-time machine learning workloads through background workers.
  • Cloud Infrastructure Cost Optimization - Reduces operational expenses through the use of spot instances and elastic compute scaling.
  • Infrastructure Provisioning Tools - Automates the creation of model clusters using declarative infrastructure-as-code configurations.
  • Virtual Private Clouds - Runs ML workloads within isolated virtual private clouds with integrated identity management for secure access.
  • Compute Instance Scaling - Elastically scales CPU and GPU compute instances using spot instances to reduce operational expenses.
  • Declarative Infrastructure Tools - Uses infrastructure-as-code and configuration templates to provision machine learning environments and clusters.
  • Private Network Security - Runs workloads within isolated virtual private clouds with integrated identity management for secure access control.
  • Observability Pipelines - The project tracks system behavior and errors by streaming metrics and logs to external monitoring tools or dashboards.
  • General Machine Learning - Platform for deploying ML models in production.
  • MLOps and Lifecycle - Deploy machine learning models.
  • Deep Learning Implementations - Platform for deploying machine learning models as web services.

Star 历史

cortexlabs/cortex 的 Star 历史图表cortexlabs/cortex 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

Cortex 的开源替代方案

相似的开源项目,按与 Cortex 的功能重合度排序。
  • clearml/clearmlclearml 的头像

    clearml/clearml

    6,740在 GitHub 上查看↗

    ClearML is a comprehensive MLOps platform designed to manage the end-to-end machine learning lifecycle, from initial experimentation to production deployment. It provides a suite of integrated tools including a pipeline orchestrator for automating workflows, an experiment tracking tool for logging hyperparameters and metrics, and a metadata-driven data versioning system for managing large-scale datasets and model artifacts. The platform is distinguished by its advanced compute management and serving capabilities. It features a GPU compute manager that supports fractional resource slicing and

    Python
    在 GitHub 上查看↗6,740
  • bentoml/bentomlbentoml 的头像

    bentoml/BentoML

    8,456在 GitHub 上查看↗

    BentoML is a machine learning model serving framework and GPU-accelerated inference server designed to package, deploy, and scale AI models as production-ready REST APIs. It functions as an AI model lifecycle manager and an inference graph orchestrator, enabling the chaining of multiple models and custom logic into complex pipelines for advanced task sequences. The framework distinguishes itself through a dynamic batching engine that optimizes GPU throughput and an artifact-based packaging system that bundles model weights and dependencies into immutable archives for consistent deployment. It

    Pythonai-inferencedeep-learninggenerative-ai
    在 GitHub 上查看↗8,456
  • nebuly-ai/nebullvmnebuly-ai 的头像

    nebuly-ai/nebullvm

    8,338在 GitHub 上查看↗

    Nebullvm is an AI inference accelerator, GPU resource orchestrator, and performance optimization library for large language models. It functions as an optimization layer designed to lower operational costs by aligning model execution with underlying hardware architectures. The system maximizes cluster efficiency through real-time dynamic partitioning and elastic quotas for shared hardware resources. It employs alignment methods and techniques to reduce the hardware and data requirements necessary for tuning large language models. The project covers broad capability areas including AI infrast

    Python
    在 GitHub 上查看↗8,338
  • boto/boto3boto 的头像

    boto/boto3

    9,834在 GitHub 上查看↗

    Boto3 is the AWS SDK for Python, providing a programmatic interface for managing and automating AWS cloud infrastructure and services. It serves as a cloud management API client and resource manager for provisioning, configuring, and scaling virtual servers, databases, and storage. The library enables the implementation of infrastructure-as-code through declarative templates and scripts, allowing for the deployment of identical resource stacks across multiple accounts and geographic regions. It also provides a framework for coordinating distributed workflows, serverless functions, and contain

    Pythonawsaws-sdkcloud
    在 GitHub 上查看↗9,834
查看 Cortex 的所有 30 个替代方案→

常见问题解答

cortexlabs/cortex 是做什么的?

Cortex is a Kubernetes-based machine learning infrastructure platform designed for deploying, scaling, and managing models and workloads. It functions as a serverless inference engine and GPU cluster orchestrator, providing the tools necessary to execute real-time, asynchronous, and batch model predictions.

cortexlabs/cortex 的主要功能有哪些?

cortexlabs/cortex 的主要功能包括:Production Serving Infrastructure, Serverless Inference Engines, GPU Resource Scaling, GPU Resource Orchestrators, Kubernetes ML Platforms, ML Infrastructure Managers, ML Orchestration Deployments, Model Inference Clusters。

cortexlabs/cortex 有哪些开源替代品?

cortexlabs/cortex 的开源替代品包括: clearml/clearml — ClearML is a comprehensive MLOps platform designed to manage the end-to-end machine learning lifecycle, from initial… bentoml/bentoml — BentoML is a machine learning model serving framework and GPU-accelerated inference server designed to package,… nebuly-ai/nebullvm — Nebullvm is an AI inference accelerator, GPU resource orchestrator, and performance optimization library for large… boto/boto3 — Boto3 is the AWS SDK for Python, providing a programmatic interface for managing and automating AWS cloud… pycaret/pycaret — PyCaret is a Python AutoML platform and MLOps lifecycle manager designed to automate machine learning workflows. It… h2oai/h2o-3 — h2o-3 is a distributed machine learning platform and automated machine learning framework designed for training and…