Building resilient systems requires more than knowing individual tools—it demands the ability to design architectures that anticipate failure and recover effectively. In this intermediate course, you will learn how to apply resilience engineering principles to modern distributed systems, focusing on high availability, fault tolerance, and disaster recovery planning.

Building Resilient Systems
Grow your skills with Coursera Plus for $239/year (usually $399). Save now.

Recommended experience
What you'll learn
Explain core resilience engineering principles and differentiate between failure types in modern distributed systems.
Analyze system architectures to identify single points of failure and resilience gaps that could impact availability.
Develop disaster recovery strategies aligned with defined business requirements such as RTO and RPO.
Evaluate monitoring, observability, and incident response practices to improve system reliability and operational resilience.
Details to know

Add to your LinkedIn profile
April 2026
4 assignments
See how employees at top companies are mastering in-demand skills

There are 4 modules in this course
Offered by
Why people choose Coursera for their career

Felipe M.

Jennifer J.

Larry W.

Chaitanya A.

Open new doors with Coursera Plus
Unlimited access to 10,000+ world-class courses, hands-on projects, and job-ready certificate programs - all included in your subscription
Advance your career with an online degree
Earn a degree from world-class universities - 100% online
Join over 3,400 global companies that choose Coursera for Business
Upskill your employees to excel in the digital economy



