Test EC2 reliability with Gremlin. Run CPU, Memory, Disk, and Shutdown experiments plus built-in reliability tests to validate autoscaling and failover.
EC2 instances carry most AWS-hosted workloads, from web tiers to batch processors. Teams plan capacity around the assumption that an instance will be there when needed, until a noisy neighbor, an exhausted disk, or an unexpected reboot proves otherwise.
Because the Gremlin Agent installs directly on the instance, you get full host coverage. Run a CPU or Memory experiment to see how your application handles resource pressure, a Disk or IO experiment to test a slow or full volume, or a Shutdown experiment to simulate the instance disappearing outright.
Gremlin's built-in reliability tests turn these one-off checks into a repeatable practice. The CPU, Memory, and Disk I/O tests validate your autoscaling thresholds, and the Host and Zone tests confirm your service survives losing an instance or an entire availability zone.
Recommended Scenarios
CPU scalability - Linux
Test that your Linux-hosted service scales as expected when CPU capacity is limited. Gremlin consumes CPU in three stages—50%, 75%, and 90%—to validate scaling thresholds.
CPU scalability - Containers
Test that your service scales as expected when CPU capacity is limited. Gremlin will consume CPU in 3 stages: 50%, 75%, and 90%.
Region Evacuation - Linux
Test your Linux-hosted service's availability when an entire cloud region becomes unavailable. Verify that traffic automatically fails over to backup regions without impacting the user experience.
Zone redundancy - Linux
Test your Linux-hosted service's availability when a randomly selected availability zone becomes unreachable. Verify that traffic fails over to secondary zones.
Intelligent Health Checks - AWS
Automatically monitor your AWS services during testing in one-click with Intelligent Health Checks.
AWS integration
Integrate Gremlin with your AWS account for automatic Health Check creation, service discovery, and more.
Using Failure Flags by proxy
Use Failure Flags, Gremlin's application-level fault injection feature, without changing a single line of code.
Installing Gremlin on Amazon ECS
Learn how to install Gremlin on EC2-backed Amazon Elastic Container Service (ECS) deployments.
Installing Gremlin on AWS - Configuring your VPC
Amazon Web Services (AWS) has unique networking requirements that must be implemented for Gremlin to run successfully…
Avoid downtime. Use Gremlin to turn failure into resilience.
Gremlin empowers you to proactively root out failure before it causes downtime. See how you can harness chaos to build resilient systems by requesting a demo of Gremlin.
