awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to dstackai/dstack

Open-source alternatives to Dstack

30 open-source projects similar to dstackai/dstack, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Dstack alternative.

  • apple/corenetapple 的头像

    apple/corenet

    6,999在 GitHub 上查看↗

    Corenet is a deep learning training framework and computer vision model library designed for developing neural networks across vision, text, and audio modalities. It functions as a distributed training orchestrator for scaling workloads across multiple compute nodes and provides a multimodal data pipeline for processing image, text, and video data. The project includes a model conversion toolkit for transforming weights and architectures between different machine learning frameworks. It also provides tools for optimizing model performance on Apple Silicon and reducing response latency in gene

    Jupyter Notebook
    在 GitHub 上查看↗6,999
  • aqueducthq/aqueductaqueducthq 的头像

    aqueducthq/aqueduct

    519在 GitHub 上查看↗

    Aqueduct is no longer being maintained. Aqueduct allows you to run LLM and ML workloads on any cloud infrastructure.

    Go
    在 GitHub 上查看↗519
  • argoproj/argoargoproj 的头像

    argoproj/argo

    16,770在 GitHub 上查看↗

    Argo is a cloud native CI/CD platform and Kubernetes workflow engine. It functions as a container pipeline orchestrator and job scheduler, managing multi-step sequences of containers as jobs using directed acyclic graphs within a cluster. The system acts as a progressive delivery controller, reducing release risk through automated Canary and Blue-Green deployment strategies. It provides declarative GitOps synchronization to mirror the state of a git repository directly into the cluster environment for continuous delivery automation. The platform covers a broad range of capabilities including

    Go
    在 GitHub 上查看↗16,770
  • argoproj/argo-workflowsargoproj 的头像

    argoproj/argo-workflows

    16,466在 GitHub 上查看↗

    Argo Workflows is a container-native workflow engine that functions as a Kubernetes custom resource controller. It orchestrates complex sequences of containerized tasks by executing them as directed acyclic graphs, allowing for dependency management and parallel processing within a cluster. The system extends the native Kubernetes control plane to manage the full lifecycle of automated processes, from initial triggering to final resource cleanup. The platform distinguishes itself through its controller-pattern reconciliation, which continuously monitors workflow states to align them with desi

    Goairflowargoargo-workflows
    在 GitHub 上查看↗16,466

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Find more with AI search
  • axolotl-ai-cloud/axolotlaxolotl-ai-cloud 的头像

    axolotl-ai-cloud/axolotl

    12,059在 GitHub 上查看↗

    Axolotl is a configuration-driven framework designed for the fine-tuning, evaluation, and quantization of large language models. It functions as a comprehensive orchestrator for distributed training, enabling users to manage complex workflows across multi-node and multi-GPU environments. By utilizing structured configuration files, the platform streamlines the setup of training parameters, dataset paths, and hardware distribution strategies. The project distinguishes itself through its support for diverse training methodologies, including full-parameter tuning, parameter-efficient adaptation,

    Pythonfine-tuningllm
    在 GitHub 上查看↗12,059
  • bindsnet/bindsnetBindsNET 的头像

    BindsNET/bindsnet

    1,654在 GitHub 上查看↗
    Pythondynamicgpu-computingmachine-learning
    在 GitHub 上查看↗1,654
  • codefuse-ai/mftcodercodefuse-ai 的头像

    codefuse-ai/MFTCoder

    714在 GitHub 上查看↗

    High Accuracy and efficiency multi-task fine-tuning framework for Code LLMs. This work has been accepted by KDD 2024.

    Python
    在 GitHub 上查看↗714
  • combust/mleapcombust 的头像

    combust/mleap

    1,537在 GitHub 上查看↗

    MLeap: Deploy ML Pipelines to Production

    Scala
    在 GitHub 上查看↗1,537
  • continualai/avalancheContinualAI 的头像

    ContinualAI/avalanche

    2,061在 GitHub 上查看↗

    Avalanche: an End-to-End Library for Continual Learning based on PyTorch.

    Python
    在 GitHub 上查看↗2,061
  • cordum-io/cordumcordum-io 的头像

    cordum-io/cordum

    483在 GitHub 上查看↗

    The open agent control plane. Govern autonomous AI agents with pre-execution policy enforcement, approval gates, and audit trails. Works with LangChain, CrewAI, MCP, and any framework.

    Go
    在 GitHub 上查看↗483
  • couler-proj/coulercouler-proj 的头像

    couler-proj/couler

    944在 GitHub 上查看↗

    Unified Interface for Constructing and Managing Workflows on different workflow engines, such as Argo Workflows, Tekton Pipelines, and Apache Airflow.

    Python
    在 GitHub 上查看↗944
  • dagworks-inc/hamiltondagworks-inc 的头像

    dagworks-inc/hamilton

    2,528在 GitHub 上查看↗

    Apache Hamilton helps data scientists and engineers define testable, modular, self-documenting dataflows, that encode lineage/tracing and metadata. Runs and scales everywhere python does.

    Jupyter Notebook
    在 GitHub 上查看↗2,528
  • deepseek-ai/3fsdeepseek-ai 的头像

    deepseek-ai/3FS

    9,970在 GitHub 上查看↗

    3FS is a distributed file system and RDMA storage cluster designed for high-performance AI training and inference workloads. It functions as a strongly consistent storage layer that utilizes a disaggregated architecture to pool SSDs and memory resources across multiple nodes. The system provides specialized storage implementations including an AI training checkpoint store for parallel state preservation and a distributed key-value cache store for decoder layer vectors to optimize inference processing. It ensures data integrity through chain replication and apportioned query distribution. The

    C++
    在 GitHub 上查看↗9,970
  • determined-ai/determineddetermined-ai 的头像

    determined-ai/determined

    3,224在 GitHub 上查看↗

    Determined is an open-source machine learning platform that simplifies distributed training, hyperparameter tuning, experiment tracking, and resource management. Works with PyTorch and TensorFlow.

    Go
    在 GitHub 上查看↗3,224
  • dotflow-io/dotflowdotflow-io 的头像

    dotflow-io/dotflow

    7在 GitHub 上查看↗

    🎲 Dotflow turns an idea into flow! — Lightweight Python library for execution pipelines

    Python
    在 GitHub 上查看↗7
  • facebookresearch/fairseqfacebookresearch 的头像

    facebookresearch/fairseq

    32,228在 GitHub 上查看↗

    Fairseq is a PyTorch toolkit for sequence-to-sequence modeling, specializing in neural machine translation, automatic speech recognition, and large-scale language model training. It provides a framework for processing and aligning diverse data sources, including text, audio, and video, to support tasks such as speech-to-text conversion and multimodal sequence learning. The project is distinguished by its distributed training capabilities, which utilize parameter sharding, mixed-precision training, and CPU offloading to handle models that exceed single-device memory. It also includes specializ

    Python
    在 GitHub 上查看↗32,228
  • flyteorg/flyteflyteorg 的头像

    flyteorg/flyte

    7,095在 GitHub 上查看↗

    Flyte is a Kubernetes-based machine learning orchestrator and containerized pipeline manager designed for coordinating AI workflows and data pipelines. It functions as an engine for defining and executing resilient pipelines, utilizing a data lineage tracker to maintain immutable execution states and ensure reproducible outputs. The platform distinguishes itself by packaging individual tasks into separate containers to ensure dependency isolation and environment consistency. It provides specialized capabilities for machine learning, including the transformation of trained models into scalable

    Go
    在 GitHub 上查看↗7,095
  • future-agi/simulate-sdkfuture-agi 的头像

    future-agi/simulate-sdk

    59在 GitHub 上查看↗

    Enterprise Grade, Voice AI simulation SDK for testing your AI Agents

    Python
    在 GitHub 上查看↗59
  • googlecontainertools/skaffoldGoogleContainerTools 的头像

    GoogleContainerTools/skaffold

    15,856在 GitHub 上查看↗

    Skaffold is a command-line tool that automates the build, push, and deployment lifecycle for containerized applications on Kubernetes. It functions as a continuous development engine, monitoring source code for changes to trigger incremental updates, manifest hydration, and automated deployments to a cluster. By abstracting the underlying build and deployment tools, it provides a unified interface for managing the inner development loop. The platform distinguishes itself through its environment-aware configuration and flexible build orchestration. It supports diverse build strategies, includi

    Gocontainersdeveloper-toolsdocker
    在 GitHub 上查看↗15,856
  • h2oai/h2o-3h2oai 的头像

    h2oai/h2o-3

    7,493在 GitHub 上查看↗

    h2o-3 is a distributed machine learning platform and automated machine learning framework designed for training and deploying predictive models using distributed in-memory computing. It functions as a deep learning framework and a distributed model scoring engine, capable of operating as a Kubernetes ML cluster to process large datasets in parallel. The platform distinguishes itself through automated machine learning capabilities that automatically select the best algorithms and hyperparameters to optimize model performance. It provides specialized deep learning toolkits for tasks including i

    Jupyter Notebookautomlbig-datadata-science
    在 GitHub 上查看↗7,493
  • huggingface/autotrain-advancedhuggingface 的头像

    huggingface/autotrain-advanced

    4,580在 GitHub 上查看↗

    This project is a multimodal model trainer and machine learning fine-tuning tool that provides a containerized workflow for adapting pre-trained models to specific tasks. It features a no-code web interface and a dashboard for training large language models and other machine learning datasets without writing code. The system distinguishes itself by integrating a no-code interface with remote GPU orchestration, allowing users to deploy containerized training environments on cloud infrastructure or local hardware. It includes a dedicated integrator for uploading trained model weights and config

    Python
    在 GitHub 上查看↗4,580
  • huggingface/nanotronhuggingface 的头像

    huggingface/nanotron

    2,718在 GitHub 上查看↗

    Minimalistic large language model 3D-parallelism training

    Python
    在 GitHub 上查看↗2,718
  • instill-ai/vdpinstill-ai 的头像

    instill-ai/vdp

    2,316在 GitHub 上查看↗

    🔮 Instill Core is a full-stack AI infrastructure tool for data, model and pipeline orchestration, designed to streamline every aspect of building versatile AI-first applications

    Python
    在 GitHub 上查看↗2,316
  • iterative/cmliterative 的头像

    iterative/cml

    4,178在 GitHub 上查看↗

    CML is a pipeline automation tool for training and evaluating machine learning models, functioning as a CI/CD system for machine learning. It serves as a cloud compute orchestrator and Git-based workflow manager that automates model training cycles through branch management, automated commits, and integrated reporting. The project distinguishes itself by provisioning ephemeral cloud instances or Kubernetes nodes to provide specialized hardware for compute-heavy tasks. It also manages remote compute runners, allowing the connection of self-hosted GPU clusters or on-premise machines to execute

    JavaScript
    在 GitHub 上查看↗4,178
  • kubeflow-kale/kalekubeflow-kale 的头像

    kubeflow-kale/kale

    694在 GitHub 上查看↗

    Kubeflow’s superfood for Data Scientists

    Python
    在 GitHub 上查看↗694
  • kubeflow/kubeflowkubeflow 的头像

    kubeflow/kubeflow

    15,739在 GitHub 上查看↗

    Kubeflow is a Kubernetes machine learning platform and containerized toolkit designed to orchestrate the entire machine learning lifecycle. It functions as an MLOps workflow orchestrator and infrastructure layer for building, training, and deploying models within containerized environments. The project provides specialized infrastructure for scaling compute resources and managing GPU workloads for large-scale distributed training. It automates the transition of models from experimental development to production through workflow orchestration and model deployment services. The platform covers

    在 GitHub 上查看↗15,739
  • kubeflow/pipelineskubeflow 的头像

    kubeflow/pipelines

    4,154在 GitHub 上查看↗

    This project is a containerized machine learning workflow engine and orchestrator designed to automate the end-to-end lifecycle of machine learning models on Kubernetes clusters. It functions as an MLOps pipeline compiler that transforms a domain-specific language into structured specifications for portable and scalable deployment. The platform provides a multi-tenant environment with isolated namespaces and identity provider authentication. It distinguishes itself through a combination of container-based task isolation, strongly typed artifact management for data passing, and content-address

    Python
    在 GitHub 上查看↗4,154
  • logicalclocks/hopsworkslogicalclocks 的头像

    logicalclocks/hopsworks

    1,302在 GitHub 上查看↗

    Hopsworks - Data-Intensive AI platform with a Feature Store

    Java
    在 GitHub 上查看↗1,302
  • logspace-ai/langflowlogspace-ai 的头像

    logspace-ai/langflow

    149,776在 GitHub 上查看↗

    Langflow is a low-code platform for designing and deploying multi-step AI agent pipelines and large language model sequences. It provides a visual environment to map logic and data flow between components, serving as an orchestrator for managing conversations and data retrieval across multiple autonomous agents. The platform distinguishes itself through a drag-and-drop interface that allows for the construction of complex AI pipelines without extensive boilerplate code. It enables the conversion of these internal workflows into standardized tools for external connectivity via the Model Contex

    Python
    在 GitHub 上查看↗149,776
  • ludwig-ai/ludwigludwig-ai 的头像

    ludwig-ai/ludwig

    11,717在 GitHub 上查看↗

    Ludwig is a multimodal machine learning platform and low-code framework designed for building, training, and deploying neural networks. It enables the construction of models that process text, images, audio, and tabular data through a unified interface using declarative configuration files rather than custom code. The system features a specialized low-code framework for large language models, supporting supervised fine-tuning, preference alignment, and a constrained decoding tool to force structured data output via logit extraction. It also includes an automated model architecture search to i

    Pythoncomputer-visiondata-centricdata-science
    在 GitHub 上查看↗11,717