This analysis demonstrates improved system reliability and reduced mean time to recovery through chaos engineering in CI/CD pipelines.
As organizations adopt DevOps practices to accelerate software delivery, integrating chaos engineering into Continuous Integration and Continuous Deployment (CI/CD) pipelines offers a proactive approach to testing system robustness. This paper explores the systematic integration of chaos engineering with CI/CD to enable resilience testing at every stage of the software lifecycle. Through an extensive literature review, architectural analysis, theoretical framing, and implementation strategies, this study identifies effective fault injection models, observability techniques, and resilience metrics suitable for automated pipelines. The results show that chaos-infused CI/CD pipelines significantly enhance system reliability, reduce Mean Time to Recovery (MTTR), and align with Site Reliability Engineering (SRE) objectives. The paper concludes by outlining best practices, limitations, and future opportunities in building self-healing, fault-tolerant systems.
No takes yet. Share an insight, caveat, or question.
Pavan Kumar Adapala (2022) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: