Test how your application handles DynamoDB throttling or outages. Run Gremlin network experiments and reliability tests to validate retry logic.
Amazon DynamoDB almost never goes offline, and that reliability is precisely what makes it dangerous. Teams stop planning for it to fail. So when a problem arises, such as throttling during a traffic spike, the application meets a condition nobody designed for.
AWS documents sixteen distinct throttling reasons in its throttling resolution guide, and each one reaches your customers the same way: as a request that failed. The damage usually follows a predictable sequence:
- An uneven access pattern pushes one partition past its throughput limit while the table's overall metrics still look healthy.
- DynamoDB returns throttling exceptions, which your SDK retries automatically.
- Without proper backoff and jitter, those retries keep the partition saturated and extend the event well past its natural end.
- Requests pile up behind SDK timeouts that were never tuned, and the failure spreads to services that never even touched DynamoDB.
Gremlin lets you rehearse that sequence safely, from your side of the connection, without touching the table. Latency experiments recreate the delay a throttled request experiences. Blackhole experiments remove the endpoint entirely so you can confirm your application degrades instead of collapsing. Gremlin helps you turn these results into actionable insights that you can use to build services that can handle even the most unexpected DynamoDB failure modes.
* Testing a dependency doesn't require installing anything on it. Gremlin runs the experiment from the service that calls it, so compatibility depends on that host rather than on the dependency itself. Check our compatibility documentation for supported operating systems and platforms, or get in touch if you don't see yours.
Intelligent Health Checks - AWS
Automatically monitor your AWS services during testing in one-click with Intelligent Health Checks.
AWS integration
Integrate Gremlin with your AWS account for automatic Health Check creation, service discovery, and more.
Using Failure Flags by proxy
Use Failure Flags, Gremlin's application-level fault injection feature, without changing a single line of code.
Installing Gremlin on Amazon ECS
Learn how to install Gremlin on EC2-backed Amazon Elastic Container Service (ECS) deployments.
Installing Gremlin on AWS - Configuring your VPC
Amazon Web Services (AWS) has unique networking requirements that must be implemented for Gremlin to run successfully…
Avoid downtime. Use Gremlin to turn failure into resilience.
Gremlin empowers you to proactively root out failure before it causes downtime. See how you can harness chaos to build resilient systems by requesting a demo of Gremlin.
