awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Netflix avatar

Netflix/chaosmonkey

0
View on GitHub↗
16,597 星标·1,271 分支·Go·apache-2.0·20 次浏览

Chaosmonkey

Chaos Monkey is a chaos engineering tool designed to verify the resilience of distributed systems by intentionally terminating production instances. It functions as a fault injection service that identifies weaknesses in cloud-based architectures by simulating real-world hardware and software outages.

The platform operates through a centralized orchestration engine that executes periodic disruption cycles based on predefined configuration rules. It employs a rule-based selection process that evaluates instance metadata against safety constraints to ensure that only eligible targets are disrupted, while a persistent data store tracks execution history to prevent excessive system instability.

The system integrates with cloud environments through a plugin-based abstraction layer that translates generic termination commands into provider-specific API calls. It monitors infrastructure lifecycle events to ensure that disruption actions remain aligned with current service health and deployment status, supporting automated site reliability engineering workflows.

Features

  • Fault Injection Testing - Tests the reliability of microservices by intentionally terminating instances to verify that the overall architecture remains operational.
  • Failure Simulation Tools - Simulates random failures in production environments to ensure that distributed systems can automatically recover from unexpected outages.
  • Instance Termination Tools - Stops production infrastructure components randomly to verify that services remain operational during unexpected failures.
  • Resilient Infrastructure - Identifies weaknesses in distributed systems by simulating real-world outages and hardware disruptions in production environments.
  • System Reliability - Identifies weaknesses in cloud-based architectures by simulating real-world outages and hardware disruptions in production.
  • Chaos Engineering - Resiliency tool for testing random instance failures.
  • Reliability And Chaos - Tool for randomly terminating instances to verify application resiliency.
  • Chaos Engineering Tools - Randomly terminates instances to test system resiliency.
  • Automated Service Reliability - Implements automated chaos experiments to validate system stability and improve incident response readiness.
  • Cloud Infrastructure Management - Monitors infrastructure state changes to ensure that automated disruption actions align with current service health and deployment status.
  • Infrastructure Abstraction Layers - Translates generic termination commands into provider-specific API calls through a modular interface layer.
  • Target Selection Rules - Evaluates instance metadata against safety constraints to identify eligible infrastructure components for disruption.
  • Orchestration Engines - Coordinates periodic execution cycles to trigger failure events based on predefined schedules and configuration rules.
  • Task Schedulers - Triggers periodic execution cycles to select and terminate infrastructure targets based on predefined configuration rules.
  • Distributed Coordination Systems - Tracks execution history and active disruption windows to prevent overlapping or excessive infrastructure instability.
  • Plugin-Based Architectures - Translates generic termination commands into provider-specific API calls through a modular interface layer.

Star 历史

netflix/chaosmonkey 的 Star 历史图表netflix/chaosmonkey 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

Chaosmonkey 的开源替代方案

相似的开源项目,按与 Chaosmonkey 的功能重合度排序。
  • netflix/simianarmyNetflix 的头像

    Netflix/SimianArmy

    7,984在 GitHub 上查看↗

    SimianArmy is a chaos engineering framework and resilience testing tool designed to induce random infrastructure failures in cloud environments. It functions as a cloud instance termination tool that simulates unplanned outages to verify that distributed architectures maintain high availability and fault tolerance. The system identifies and terminates cloud server instances to ensure applications can tolerate unexpected hardware failures without interrupting service. This process allows for the verification of automated failover mechanisms and the identification of weaknesses in system reliab

    Java
    在 GitHub 上查看↗7,984
  • shopify/toxiproxyShopify 的头像

    Shopify/toxiproxy

    12,088在 GitHub 上查看↗

    Toxiproxy is a framework designed for chaos engineering and network resilience testing. It functions as a programmable TCP proxy that intercepts and routes data streams between clients and servers, allowing developers to simulate unstable network conditions such as latency, bandwidth throttling, and connection failures. The tool provides a control plane that enables the dynamic manipulation of network conditions on active connections in real time. By integrating into automated test suites, it allows for the programmatic injection of faults to validate how distributed systems and microservices

    Gochaosdowngo
    在 GitHub 上查看↗12,088
  • voltagent/awesome-claude-code-subagentsVoltAgent 的头像

    VoltAgent/awesome-claude-code-subagents

    21,906在 GitHub 上查看↗

    This project provides a framework for managing multi-agent systems, designed to automate complex software development, infrastructure, and business workflows. It functions as a multi-agent workflow orchestrator that routes tasks to domain-specific workers while maintaining state persistence and infrastructure automation. By leveraging large language models, the system decomposes high-level objectives into actionable plans, ensuring that complex operations are executed with consistency and reliability. The framework distinguishes itself through its hierarchical agent registry and policy-driven

    Shellai-agent-frameworkai-agent-toolsai-agents
    在 GitHub 上查看↗21,906
  • chaos-mesh/chaos-meshchaos-mesh 的头像

    chaos-mesh/chaos-mesh

    7,761在 GitHub 上查看↗

    Chaos Mesh is a cloud-native fault injection tool and Kubernetes chaos engineering platform designed to verify system resilience. It functions as a testing framework for designing and executing automated failure scenarios to evaluate how containerized workloads recover from disruptions. The project acts as a multi-cluster chaos orchestrator, providing a centralized control plane to manage and monitor experiments across multiple remote Kubernetes clusters from a single interface. It includes a dashboard for the visual scheduling of experiments and the coordination of complex failure scenarios.

    Go
    在 GitHub 上查看↗7,761
查看 Chaosmonkey 的所有 30 个替代方案→

常见问题解答

netflix/chaosmonkey 是做什么的?

Chaos Monkey is a chaos engineering tool designed to verify the resilience of distributed systems by intentionally terminating production instances. It functions as a fault injection service that identifies weaknesses in cloud-based architectures by simulating real-world hardware and software outages.

netflix/chaosmonkey 的主要功能有哪些?

netflix/chaosmonkey 的主要功能包括:Fault Injection Testing, Failure Simulation Tools, Instance Termination Tools, Resilient Infrastructure, System Reliability, Chaos Engineering, Reliability And Chaos, Chaos Engineering Tools。

netflix/chaosmonkey 有哪些开源替代品?

netflix/chaosmonkey 的开源替代品包括: netflix/simianarmy — SimianArmy is a chaos engineering framework and resilience testing tool designed to induce random infrastructure… shopify/toxiproxy — Toxiproxy is a framework designed for chaos engineering and network resilience testing. It functions as a programmable… voltagent/awesome-claude-code-subagents — This project provides a framework for managing multi-agent systems, designed to automate complex software development,… chaos-mesh/chaos-mesh — Chaos Mesh is a cloud-native fault injection tool and Kubernetes chaos engineering platform designed to verify system… litmuschaos/litmus — Litmus is a cloud native chaos engineering platform and fault injection tool used to design and execute controlled… getmoto/moto — Moto is a cloud service mockery framework and API mock server that simulates AWS infrastructure locally. It allows…