Loading SaaS profile...
Loading SaaS profile...
Proactive reliability and chaos engineering platform for cloud infrastructure.
- Provides fault injection capabilities to safely test system resilience against real-world failures. - Automates reliability testing and dependency discovery across complex cloud environments. - Validates disaster recovery processes, including zone, region, and datacenter failovers. - Monitors detected risks and supplies actionable reliability insights to prevent outages.
Proactively identifies system vulnerabilities before customer-facing outages occur. Automates complex chaos engineering experiments to reduce manual toil and testing overhead. Provides safe, controlled fault injection that isolates testing from production data loss. Integrates seamlessly into modern CI/CD pipelines and cloud architectures.
Category: AI & Automation
Team Size: 26-100
Visit WebsiteGremlin is a reliability platform designed to help engineering teams proactively manage system stability at scale. The platform enables users to uncover infrastructure risks, run automated fault injection tests, and validate disaster recovery processes. By simulating outages safely before they impact customers, teams can prevent downtime and improve application resilience. It is trusted by leading enterprise cloud engineering teams worldwide.
Gremlin was founded in 2016 by Kolton Andrus and Matthew Fornaciari in San Jose, United States. Drawing from their backgrounds with large-scale distributed systems at companies like Amazon and Netflix, the founders sought to make chaos engineering accessible to all engineering teams. The company has raised $26.8M in venture funding to build a safer, more reliable internet.