Google pioneered the discipline of Site Reliability Engineering (SRE), applying reliability to the entire user journey for consumer, enterprise, and infrastructure systems. In the years since, many organizations have followed suit, guided by the tenets laid out in the best-selling first edition of this practical book. Now, a fully revised edition brings Site Reliability Engineering up-to-date with fresh insights on engineering techniques, organizational processes, and case studies that will help you promote and implement greater reliability throughout the engineering lifecycle.
In this edition, key members of Google's SRE team explore the company's current SRE practices and explain how they've evolved in the decade since the book's initial publication. Updates cover the value of reliability, cloud reliability, and the impact of AI. You'll learn the principles and practices that can make your organization's systems scalable, reliable, and efficient.
"synopsis" may belong to another edition of this title.
Betsy Beyer has worked at Google for almost 20 years, most recently in the role of program manager for Site Reliability Engineering. She's the editor of three books in this space, including the best-selling first edition of this one, and also cochairs the editorial board of ACM Queue. She previously worked on Google's datacenter and hardware operations teams. Before moving to New York, Betsy was a lecturer on technical writing at Stanford University. She holds degrees from Stanford and Tulane.
Chris Jones is a Site Reliability Engineer for Google Maps. Based in San Francisco, he has previously been responsible for the care and feeding of Google App Engine, a cloud platform-as-a-service product serving over 28 billion requests per day, and a number of other services. In other lives, Chris has worked in privacy engineering and academic IT, analyzed data for political campaigns, and engaged in some light BSD kernel hacking, picking up degrees in computer engineering, economics, and technology policy along the way. He's also a licensed professional engineer.
Christof Leng has been working as a Site Reliability Engineer for Google for more than 11 years in the Dublin and Munich offices. He has worked on systems in Google's ads, cloud, and internal developer infrastructure and built and managed various teams in these areas over the years. He has been responsible for central Google SRE programs including the SRE engagement model, production excellence (ProdEx), and production launch reviews. Christof holds a PhD in computer science from TU Darmstadt and has been a postdoc at ICSI and UC Berkeley. He has also served as vice president of the German Informatics Society (GI). Christof lives in Darmstadt with his wife, three children, and two cats.
David Huska is a software engineer and Site Reliability Engineer at Google, scaling productivity and reliability through AI. His operational background includes roles with Google's Cloud Incident Response Team, the cross-service escalation point for major incidents on Google Cloud, and Google Maps. David has directly advised Google Cloud's largest enterprise customers on integrating SRE best practices into their services and operations. He is a coauthor and contributing editor of The Site Reliability Workbook and Building Secure and Reliable Systems, both from O'Reilly. Based in California, David has worked at Google for over 15 years on two continents at locations strategically close to great sailing.
Jennifer Petoff is a globally recognized expert on SRE education. She is the lead author of Training Site Reliability Engineers: What Your Organization Needs to Create a Learning Program. Known as The Reliable PgM, she is an advocate for applying SRE principles to a wide range of program management situations. Jennifer is an avid public speaker: she has given talks and keynote presentations and has participated in panel discussions at DevOps, SRE, and other industry conferences in 16 countries (and counting) on four continents. Jennifer joined Google in 2007 after spending eight years in the chemical industry. She holds a PhD in chemistry from Stanford University, along with a BS in chemistry and a BA in psychology from the University of Rochester in Western New York. Jennifer and her husband Scott are avid travelers and have lived in the US, Ireland, and now Portugal. Jennifer loves travel writing and door photography, both of which you can find on Sidewalk Safari.
"About this title" may belong to another edition of this title.
Seller: California Books, Miami, FL, U.S.A.
Condition: New. Seller Inventory # I-9798341607682
Seller: Rarewaves USA United, HEBRON, KY, U.S.A.
Paperback. Condition: New. Google pioneered the discipline of Site Reliability Engineering, applying reliability to the entire user journey for consumer, enterprise, and infrastructure systems. In the years since, many organizations have followed suit, guided by the tenets laid out in this practical book. This fully revised edition brings Site Reliability Engineering up-to-date with fresh insights on engineering techniques, organizational processes, and case studies that will help you promote and implement greater reliability throughout the engineering lifecycle.In this collection of essays and articles, key members of Google's Site Reliability Engineering team explore the company's current SRE practices and explain how they've evolved in the decade since the initial publication. New updates cover the value of reliability, cloud reliability, and the impact of AI. You'll learn the principles and practices that enable Google engineers to make some of the world's largest systems scalable, reliable, and efficient-lessons directly applicable to your organization.Train new Site Reliability Engineers based on the latest practices in the fieldDevelop engineering organizations that support reliability as a featureBuild online services that incorporate reliability principlesUse AI to improve SRE across the organization and optimize critical areas such as automation and incident detection. Seller Inventory # LU-9798341607682
Seller: Rarewaves.com UK, London, United Kingdom
Paperback. Condition: New. Google pioneered the discipline of Site Reliability Engineering, applying reliability to the entire user journey for consumer, enterprise, and infrastructure systems. In the years since, many organizations have followed suit, guided by the tenets laid out in this practical book. This fully revised edition brings Site Reliability Engineering up-to-date with fresh insights on engineering techniques, organizational processes, and case studies that will help you promote and implement greater reliability throughout the engineering lifecycle.In this collection of essays and articles, key members of Google's Site Reliability Engineering team explore the company's current SRE practices and explain how they've evolved in the decade since the initial publication. New updates cover the value of reliability, cloud reliability, and the impact of AI. You'll learn the principles and practices that enable Google engineers to make some of the world's largest systems scalable, reliable, and efficient-lessons directly applicable to your organization.Train new Site Reliability Engineers based on the latest practices in the fieldDevelop engineering organizations that support reliability as a featureBuild online services that incorporate reliability principlesUse AI to improve SRE across the organization and optimize critical areas such as automation and incident detection. Seller Inventory # LU-9798341607682
Quantity: 2 available