CertClue
Courses · Cloud
Platform / Reliability Engineer

Resilient Hosting: One App, Many Providers

Keep three systems up across providers for eight days: map single points of failure, survive a results-day spike, fail a hospital database over and back without losing prescriptions, recover from a bad migration, keep government data in-country, and drill the switch.
1.5 hrs taught · 6 to 10 hrs applied 7 modules 20 lessons 4 portfolio artifacts Completion certificate Updated September 2026
Created by the CertClue team
What you'll build

Real portfolio pieces built during the course, not a certificate for its own sake. Each one is work you can show.

  • Resilience baseline: failure map and targetsEverything the inherited school portal depends on, what would take it down, and the recovery targets agreed with the school, with the cost of each option.
  • Failover and failback runbookThe steps that switched the hospital system to its second-provider copy, and the steps that bring it back without losing a prescription.
  • Data residency architecture decisionAn architecture decision record for the state agency's permit portal: resilience across two providers, with every copy of citizens' data kept in Nigeria.
  • Game day and cost review reportThe failover drill's honest result against its target, the gaps it found and their fixes, the redundancy worth its cost, and the next drill.

What you'll learn

Resilience, kept short
The job, decoded
A day in the seat
The rhythm of the job
Into the simulation: Ikoyo Digital
Building the portfolio
What comes next

Course content · 7 modules, 20 lessons

Sign up to unlock every lesson - the titles below show exactly what is inside.

What the job is, why the data is the hard part, and targets that come from the business.

What a platform or reliability engineer does
The app is easy to copy; the data is not
RPO and RTO from business impact

Requirements

  • Comfortable with the fundamentals this course's own Module 1 covers, or equivalent experience.
  • No prior experience in this field is required to start.
  • A computer with a reliable internet connection.
  • Comfortable using a web browser - no software to install.

Description

Every CertClue course follows the same seven-part shape: fundamentals, the role translated out of job-posting language, a real working day, the job's recurring rhythms, a multi-day simulation, the portfolio you build along the way, and a handoff into your next move. Here is what that looks like for platform / reliability engineer.

Who this course is for

Anyone aiming to become a platform / reliability engineer, including career changers with no background in it yet. This is the entry rung of a realistic ladder:

entry
Cloud / DevOps Engineer

Deploys and runs an app on one provider and restores it when something breaks.

Deploying to one providerManaged databasesBackupsBasic monitoring
mid
Platform / Reliability Engineer

Designs how an app survives a provider, a region or a bad deploy failing, and proves it with drills.

RPO and RTO from business impactCross-provider replicas and failoverFailback without data lossGame days and runbooks
senior
Site Reliability / Platform Lead

Sets reliability targets with the business and decides where redundancy is worth its cost.

Reliability strategyError budgetsVendor and cost strategyIncident leadership
Reviews

No reviews yet. Reviews come from learners who have taken the course, so this stays empty until someone leaves one.

Sign in to leave a review.

Students also explore