Steadybit, generally seen as one of the biggest players in chaos engineering and reliability testing, has launched what it calls the industry’s first Model Context Protocol (MCP) Server, which is able to integrate chaos engineering data directly into large language model workflows.
For QA and SRE teams, particularly in the highly regulated banking and financial services sector, this represents a significant advance in proactive reliability testing, according to Benjamin Wilms, the CEO and co-founder of Bonn, Germany-based Steadybit.
“This means a new paradigm for how teams analyse and act on reliability data,” stressed Wilms, elaborating that “teams can now query the outcomes of chaos experiments using natural language in tools like Claude, Gemini, or ChatGPT, and that means faster insights, more targeted testing, and smarter system resilience strategies.
He explained that the MCP Server provides a standardised interface that connects Steadybit’s chaos engineering platform with multiple AI-powered testing systems, allowing Site Reliability Engineering (SRE) and QA teams to extract experiment insights in real time.
For financial firms seeking to reduce downtime, avoid regulatory penalties, and stay ahead of operational risk mandates such as the EU’s DORA framework, this innovation is especially timely, Wilms argued.
“With high-profile outages still making headlines, the cost of untested systems is clear,” he continued. “This is about empowering QA and platform teams to stress-test systems and understand the limits of resilience, before something breaks.”
The ability to integrate experiment data with observability tools like Datadog or incident response platforms such as PagerDuty means teams can now query not only what happened during a chaos test, but why, and what to do next, he pointed out.
Chaos engineering, a practice that ensures systems are robust and resilient against unpredictable online environments, is increasingly a priority for QA teams within the financial services space and other industries.
In the financial sector, where testing environments must reflect highly complex, interdependent systems, this new capability could streamline how QA professionals plan, run, and refine reliability experiments.
“Teams can now type a prompt like ‘what gaps exist in our current experiment coverage for our payment service?’ and get answers instantly,” said Wilms.
From summarising chaos test results by team to recommending new test types based on historical gaps, the MCP Server opens new workflows that blend AI ease with rigorous testing discipline.
When combined with other system data, the results can influence incident response strategies and improve mean time to resolution, he added.
“We are making chaos engineering not only easier to adopt, but smarter to use, especially at scale.”
– Benjamin Wilms
Salesforce’s Director of Software Engineering, Krishna Palati, was prepared to endorse the solution: “This MCP will enable us to just type a prompt to pull custom reports, analyse reliability testing gaps, and get insights on what experiments to run next.”
Steadybit’s latest innovation reflects a broader effort to bring chaos engineering out of the specialist silo and into mainstream QA and DevOps pipelines.
The company’s platform allows teams to create experiments through a no-code editor and run them across distributed systems. The new MCP Server adds a layer of intelligence that helps prioritize future tests based on impact, coverage, and business context.
“Every team and tech stack is different,” Wilms emphasised. “What we’re doing with the MCP is making chaos engineering not only easier to adopt, but smarter to use, especially at scale.”
As QA leaders in banking, insurance, and fintech look to future-proof their operational resilience strategies, tools like the Steadybit MCP may well become essential for continuous, AI-powered verification of system reliability.
Capital raise
Steadybit is a Germany-based chaos engineering platform focused on proactive reliability testing. Its tooling helps SRE, QA, and platform teams identify weaknesses before outages occur, using a combination of automated experiments, strong observability integrations, and AI-driven insights.
The server launch comes roughly a year after the firm raised around $6 million in fresh funding.
The May 2024 funding round, which was led by Paladin Capital Group, allowed the company to make new and additional investments in its hugely popular chaos engineering platform. Fresh allocations came from existing investors Boldstart Ventures, Angular Ventures, and NewForge.
The new funding have mostly been used to help Steadybit expand its teams in product development, engineering, sales, and marketing.

Ken Pentimonti, managing director of Washington-based Paladin Capital Group, explained to QA Financial why his firm decided to pump more capital into Steadybit.
“As digital systems become more complex and interdependent, organisations are at greater risk of failures and system outages,” he said.
“Many products fail to meet customer expectations as software engineering teams are focused on releasing new features quickly rather than product reliability,” Pentimonti continued.
“Organisations that interface with customers digitally are realising that service reliability is a key priority for growth and customer retention,” he noted, pointing out that “this is fuelling demand for Steadybit’s platform, designed to reduce outages and provide visibility into distributed systems to detect issues.”
A range of large banks, financial services firms and other large companies have integrated the solution to pre-empt and mitigate system vulnerabilities, enhance their overall performance and user experience.
Alongside the funding announcement, Steadybit introduced a reliability advice feature which monitored all information collected about the system under observation to discover potential reliability gaps.
Users are given instructions on how to fix errors and offered experiments that validate the effectiveness of those interventions, creating their own libraries of good engineering practices, explained Wilms.
He concluded by stressing that “as the demand for accessibility tools grows, tools like this one become essential for businesses to ensure their digital content is accessible to all.”
NEXT WEEK

NEW EVENT

Why not become a QA Financial subscriber?
It’s entirely FREE
* Receive our weekly newsletter every Wednesday * Get priority invitations to our Forum events *

REGULATION & COMPLIANCE
Looking for more news on regulations and compliance requirements driving developments in software quality engineering at financial firms? Visit our dedicated Regulation & Compliance page here.
READ MORE
- Banks confront rising agentic AI testing challenge
- Testing turns ‘reactive’ as documentation lags
- Inside UBS’s landmark testing challenge
- Bank of England raises the testing bar for frontier AI
- Tricentis, Tabnine and Snyk: the latest vendor and product news
WATCH NOW



