Nicht aus der Schweiz? Besuchen Sie lehmanns.de
Becoming a Rockstar SRE - Jeremy Proffitt, Rod Anami

Becoming a Rockstar SRE

Electrify your site reliability engineering mindset to build reliable, resilient, and efficient systems
Buch | Softcover
420 Seiten
2023
Packt Publishing Limited (Verlag)
978-1-80323-922-4 (ISBN)
CHF 59,30 inkl. MwSt
Excel in site reliability engineering by learning from field-driven lessons on observability and reliability in code, architecture, process, systems management, costs, and people to minimize downtime and enhance developers’ output
Purchase of the print or Kindle book includes a free eBook in the PDF format

Key Features

Understand the goals of an SRE in terms of reliability, efficiency, and constant improvement
Master highly resilient architecture in server, serverless, and containerized workloads
Learn the why and when of employing Kubernetes, GitHub, Prometheus, Grafana, Terraform, Python, Argo CD, and GitOps

Book DescriptionSite reliability engineering is all about continuous improvement, finding the balance between business and product demands while working within technological limitations to drive higher revenue. But quantifying and understanding reliability, handling resources, and meeting developer requirements can sometimes be overwhelming. With a focus on reliability from an infrastructure and coding perspective, Becoming a Rockstar SRE brings forth the site reliability engineer (SRE) persona using real-world examples.
This book will acquaint you the role of an SRE, followed by the why and how of site reliability engineering. It walks you through the jobs of an SRE, from the automation of CI/CD pipelines and reducing toil to reliability best practices. You’ll learn what creates bad code and how to circumvent it with reliable design and patterns. The book also guides you through interacting and negotiating with businesses and vendors on various technical matters and exploring observability, outages, and why and how to craft an excellent runbook. Finally, you’ll learn how to elevate your site reliability engineering career, including certifications and interview tips and questions.
By the end of this book, you’ll be able to identify and measure reliability, reduce downtime, troubleshoot outages, and enhance productivity to become a true rockstar SRE!What you will learn

Get insights into the SRE role and its evolution, starting from Google’s original vision
Understand the key terms, such as golden signals, SLO, SLI, MTBF, MTTR, and MTTD
Overcome the challenges in adopting site reliability engineering
Employ reliable architecture and deployments with serverless, containerization, and release strategies
Identify monitoring targets and determine observability strategy
Reduce toil and leverage root cause analysis to enhance efficiency and reliability
Realize how business decisions can impact quality and reliability

Who this book is forThis book is for IT professionals, including developers looking to advance into an SRE role, system administrators mastering technologies, and executives experiencing repeated downtime in their organizations. Anyone interested in bringing reliability and automation to their organization to drive down customer impact and revenue loss while increasing development throughput will find this book useful. A basic understanding of API and web architecture and some experience with cloud computing and services will assist with understanding the concepts covered.

Jeremy Proffitt is passionate about solving problems with an unmatched sense of urgency - the definition of a Site Reliability Engineer. A master of solutions and technology knowledge, Jeremy is a rockstar SRE with AWS Professional Certifications in Architecture and DevOps. He has routinely saved millions in potential lost revenue in his career. In his free time, Jeremy enjoys sending time in his rockstar-appropriate man cave and loves venturing into 3D printing, electronics, and Internet of Things (IoT) projects. Jeremy currently manages a team of top SRE and DevOps talent, driving constant improvement, and is often cited in the company as a visionary of observability and emergency response. Rod Anami is a seasoned engineer who works with cloud infrastructure and software engineering technologies. As one of the SREs at the Kyndryl CoE, he coaches other SREs on running IT modernization, transformation, and automation projects for clients worldwide. Rod leads the global SRE guild inside Kyndryl, where he helps plant and grow SRE chapters in many countries. Rod is certified as an SRE, technical specialist, and DevOps engineer professional at the ultimate level. He holds AWS, HashiCorp, Azure, and Kubernetes certifications, among many others. He is passionate about contributing to open source software at large with Node.js libraries.

Table of Contents

SRE Job Role – Activities and Responsibilities
Fundamental Numbers – Reliability Statistics
Imperfect Habits – Duct Tape Architecture and Spaghetti Code
Essential Observability – Metrics, Events, Logs, and Traces (MELT)
Resolution Path – Master Troubleshooting
Operational Framework – Managing Infrastructure and Systems
Data Consumed – Observability Data Science
Reliable Architecture – Systems Strategy and Design
Valued Automation – Toil Discovery and Elimination
Exposing Pipelines – GitOps and Testing Essentials
Worker Bees – Orchestrations of Serverless, Containers, and Kubernetes
Final Exam – Tests and Capacity Planning
First Thing – Runbooks and Low Noise Outage Notifications
Rapid Response – Outage Management Techniques
Postmortem Candor – Long-Term Resolution
Chaos Injector – Advanced Systems Stability
Interview Advice – Hiring and Being Hired
Appendix A The Site Reliability Engineer Manifesto
Appendix B The 12-Factor App Questionnaire

Erscheinungsdatum
Verlagsort Birmingham
Sprache englisch
Maße 191 x 235 mm
Themenwelt Mathematik / Informatik Informatik Software Entwicklung
Mathematik / Informatik Informatik Theorie / Studium
ISBN-10 1-80323-922-0 / 1803239220
ISBN-13 978-1-80323-922-4 / 9781803239224
Zustand Neuware
Informationen gemäß Produktsicherheitsverordnung (GPSR)
Haben Sie eine Frage zum Produkt?
Mehr entdecken
aus dem Bereich
Deterministische und randomisierte Algorithmen

von Volker Turau; Christoph Weyer

Buch | Softcover (2024)
De Gruyter Oldenbourg (Verlag)
CHF 89,95
Grundlagen, Prozesse, Methoden und Werkzeuge

von Jörg Schäuffele; Thomas Zurawka

Buch | Hardcover (2024)
Springer Vieweg (Verlag)
CHF 139,95
Programmieren erlernen und technische Fragestellungen lösen

von Harald Nahrstedt

Buch | Softcover (2023)
Springer Vieweg (Verlag)
CHF 62,95