
Reliability fails in surprising ways—not because engineers don’t care, but because our brains are optimized for speed, not accuracy. In high-pressure IT environments, that leads to familiar patterns: premature conclusions during incidents, “obvious” fixes that don’t fix anything, risky changes that felt safe at the time, and meetings where everyone nods while the real problem stays untouched. This lecture is a sharp, highly relatable look at the human side of reliability: how bias shapes decisions in operations, why smart teams repeat the same mistakes, and what lightweight practices help you improve outcomes without adding bureaucracy. Expect real-world scenarios, humor, and practical takeaways you can apply immediately—without me giving away the full playbook.
I’m Marcel Koert—a Site Reliability / DevOps / Platform Engineering professional focused on turning “reliability theory” into a way of working teams can actually sustain. I’m the author of two books in the Essential SRE series: • Essential SRE: Way of Working — a practical operating model for SRE: how to run reliability day to day, reduce chaos, and make delivery and operations work together (without hero culture). • Essential SRE Articles — a curated, straight-to-the-point collection of the essential knowledge SREs need, grounded in real-world application rather than slideware. Alongside the books, I share talks, playbooks, and field-tested guidance through MeloMar-IT to help engineers and leaders build systems (and teams) that stay stable under pressure. If you care about reliable delivery, calmer on-call, and reliability that survives contact with reality—let’s connect.