Recommended Scenario

Container redundancy

Test resilience to container failures by shutting down a randomly selected container. Verify that your container runtime automatically restarts or replaces it.

Experiment types

Shutdown

Targets

Containers

Length

5 minutes

How it works

How this Scenario works

This Scenario shuts down a randomly selected container, simulating an unexpected container failure. This tests whether your container runtime's restart policies and health check mechanisms detect the failure and recover automatically.

Use cases

Why run this Scenario?

This Scenario uses the same principle as Chaos Monkey: if a host or container shuts down unexpectedly, the underlying platform should detect this and automatically restart or replace it.

  • Validate that container restart policies are configured correctly and recover failed containers within acceptable timeframes.
  • Verify that load balancers or service meshes route traffic away from the failed container during recovery.
  • Test that container health checks detect the failure and trigger restarts appropriately.
  • Confirm that container-level failures don't cascade into broader application outages.
Result

What to expect when you run it

When a container fails, the container runtime or orchestrator automatically restarts or replaces it, and traffic is routed to healthy replicas.

Avoid downtime. Use Gremlin to turn failure into resilience.

Gremlin empowers you to proactively root out failure before it causes downtime. See how you can harness chaos to build resilient systems by requesting a demo of Gremlin.

Product Hero ImageShape