Test GCP Distributed Cloud Edge reliability with Gremlin. Simulate network latency and blackholes to validate edge connectivity resilience.
Google Distributed Cloud Edge runs workloads close to where data is produced, often in locations with limited connectivity and nobody technical on site. Both characteristics make failure handling more important and testing it more awkward.
The connection back to the core is the dependency most likely to break and least likely to have been tested. Applications developed against reliable connectivity treat the central control plane as always available. At an edge site that assumption can fail for hours, and what the workload does during that window is usually unknown until it happens.
Install the Gremlin Agent on your edge nodes and you can answer it deliberately. Blackhole experiments sever the path back to the core so you can confirm whether local operation continues or blocks on a call that will never return. Latency experiments simulate the degraded link that's more common than a clean disconnection.
For sites where a truck roll is the recovery plan, the difference between autonomous operation and aspirational autonomy is worth confirming in advance.
Unreliable Networks - Linux
Simulate unreliable network conditions on Linux hosts by adding latency to API calls. Test whether your users are affected when network response times degrade to hundreds or thousands of milliseconds.
DNS Redundancy - Linux
Test your Linux-hosted service's availability when its primary DNS server is unreachable. Verify that DNS failover routes traffic correctly through secondary providers.
Zone redundancy - Linux
Test your Linux-hosted service's availability when a randomly selected availability zone becomes unreachable. Verify that traffic fails over to secondary zones.
* Gremlin is designed to work on any cloud platform that provides Linux or Windows hosts. We haven't individually tested every service we cover, and not all are officially supported. Check our compatibility documentation for tested operating systems and known caveats, or get in touch if you don't see yours.
Intelligent Health Checks - GCP
Automatically monitor your Google Cloud services during testing in one-click with Intelligent Health Checks.
Avoid downtime. Use Gremlin to turn failure into resilience.
Gremlin empowers you to proactively root out failure before it causes downtime. See how you can harness chaos to build resilient systems by requesting a demo of Gremlin.
