Published Jun 22, 2023

SE Radio 569: Vladyslav Ukis on Rolling out SRE in an Enterprise

Vladyslav Ukis delves into the enterprise implementation of Site Reliability Engineering (SRE), highlighting its integration with ITIL, the quantification of service reliability through error budgets, and overcoming the cultural challenges of transformation to enhance IT efficiency.
Episode Highlights
Software Engineering Radio - the podcast for professional software developers logo

Popular Clips

Episode Highlights

  • Stakeholder Alignment

    Aligning stakeholders is crucial for a successful Site Reliability Engineering (SRE) transformation. shares how at Teamplay digital health platform, a combination of bottom-up and top-down approaches was used to gain support. Informal meetings and presentations helped seed initial interest, while leadership was engaged to include SRE in the organization's portfolio management. emphasizes the importance of this dual approach, stating, "That combination of bottom up and top down is absolutely necessary here because one without the other doesn't work." 1 2

       

    Foundation Steps

    Establishing a solid foundation is essential for SRE transformation. outlines three foundational steps: aligning stakeholders to prioritize SRE, assessing the current state of the organization, and planning the transformation. He notes that each step brings visible improvements, making the process rewarding. explains, "Every little step will mean a tangible improvement." 3 4

       

    Implementation Challenges

    Implementing SRE comes with its own set of challenges, particularly in traditional organizations. highlights the cultural shift required, as developers must take on on-call duties, a significant change from traditional roles. To ease this transition, he suggests starting with on-call duties during business hours only. shares, "We are only talking about on call during business hours," which helped developers adapt without disrupting their work-life balance. 5 6

Related Episodes