Steadybit is a reliability platform that lets SRE and platform teams discover assets across their tech stack and run no‑code chaos engineering experiments to uncover and fix resilience weaknesses before they cause outages. It offers an open‑source extension framework for 20+ integrations, a drag‑and‑drop timeline editor with templates, and automated reliability advice to improve observability alert coverage and incident response.
Funding
$6M raised to dateRaised to date based on public sources. This may differ from the amount the company actually raised and is based only on what is publicly available on the internet.

2OFounders
Product
Problem
Organizations often lack systematic ways to test how their services behave under failure conditions, leading to undetected reliability weaknesses that can cause outages and prolonged incident response times.
Solution
Steadybit provides a reliability platform that enables teams to discover targets across their entire tech stack and run controlled chaos engineering experiments without writing code. An open‑source extension framework connects to cloud services, Kubernetes, message brokers, observability tools, and more, allowing fault injection and health checks at network, resource, and application layers. Users create experiments with a drag‑and‑drop, timeline‑based editor or use ready‑made templates, then visualize results and receive automated reliability advice. The platform supports both SaaS and on‑premises deployments, integrates with existing CI/CD pipelines, and offers role‑based access to foster a culture of proactive reliability engineering.
Target Audience
Primary customers are Site Reliability Engineering (SRE) and platform teams in medium to large enterprises that need to proactively assess and improve the resilience of cloud‑native, containerized, and hybrid infrastructures.
Features
- Open‑source extensions for 20+ technologies (e.g., AWS, Azure, Kubernetes, Kafka, Datadog, Istio) enabling target discovery and fault injection across cloud, on‑prem, and hybrid environments
- No‑code experiment editor with drag‑and‑drop actions, timeline control, and reusable templates for common failure scenarios
- Single Steadybit agent per network boundary simplifies installation and secure communication with all integrated services
- Automated reliability advice that assesses target compliance with best‑practice resilience patterns
- Visualization of discovered assets and grouping by metadata to map system topology and impact scopes
- Role‑based team and permission management to coordinate reliability initiatives across SRE and platform teams