Skip to content
all episodes
Code of Architecture · episode 46

Continuous Architecture in Practice — Episode 4

1:21:30

Episode participants

  • Evgeny Peshkov

    guest

  • Sergey Baranov

    guest

Conversation

What we discussed on the recording

The finale of Continuous Architecture in Practice covers reliability and emerging technology. Sergey Baranov and Evgeny Peshkov distinguish availability, reliability, resilience, and fault tolerance: reachability, failure-free operation, and ways to withstand failure and recover.

High availability is expected far beyond technology companies, but it needs boundaries. Cloud zones and regions, fault isolation, and standby paths help only with deliberate application design. “Five nines” mean little without critical journeys, downtime cost, and a decision about which functions need continuous service.

MTBF and MTTR connect failure frequency with recovery time, while RPO and RTO define acceptable data loss and restoration time. The panel discusses postmortems, prevented failures, and preserving evidence after an incident. Saturation metrics can trigger action before a predictable outage.

The practical loop is recognition, isolation, recovery, and verification, supported by health checks, timeouts, circuit breakers, fallbacks, replication, backups, and compensation. Operability and observability are designed with the system, and disaster recovery rehearsed. A closing discussion of blockchain and AI warns against uncritical hype and reflexive rejection.

Book series
Continuous Architecture in Practice
Murat Erder, Pierre Pureur, Eoin Woods
Book playlist