This book provides a rich collection of the essential foundation and advanced practices for understanding and running SRE. The first part gives a brief historical trace of how SRE is born, of its roots in DevOps, highlighting its relevance in the context of minimizing downtime and achieving a better software reliability. The book explores the core SRE principles such as service level objectives (SLOs), automation and incident management. The focus is on building resilient systems that can take faults, that will balance it, and mitigate against disasters. Readers will learn what observability is, real time monitoring, and post mortem process. The book also goes on to explain automation, Infrastructure as Code (IaC), CI/CD pipelines and the rise of AI to use in incident response and self-healing systems. Last, it covers organizational adoption of SRE, promotion of collaboration, error budgeting and managing multi cloud environments. Engineers, architects, and leaders who wish to instill reliability and resilience in modern software operations should read this book.
"Sinopsis" puede pertenecer a otra edición de este libro.
Librería: California Books, Miami, FL, Estados Unidos de America
Condición: New. Nº de ref. del artículo: I-9798899843983
Cantidad disponible: Más de 20 disponibles
Librería: PBShop.store US, Wood Dale, IL, Estados Unidos de America
HRD. Condición: New. New Book. Shipped from UK. Established seller since 2000. Nº de ref. del artículo: L2-9798899843983
Cantidad disponible: Más de 20 disponibles
Librería: PBShop.store UK, Fairford, GLOS, Reino Unido
HRD. Condición: New. New Book. Shipped from UK. Established seller since 2000. Nº de ref. del artículo: L2-9798899843983
Cantidad disponible: Más de 20 disponibles
Librería: AHA-BUCH GmbH, Einbeck, Alemania
Buch. Condición: Neu. Neuware - This book provides a rich collection of the essential foundation and advanced practices for understanding and running SRE. The first part gives a brief historical trace of how SRE is born, of its roots in DevOps, highlighting its relevance in the context of minimizing downtime and achieving a better software reliability. The book explores the core SRE principles such as service level objectives (SLOs), automation and incident management. The focus is on building resilient systems that can take faults, that will balance it, and mitigate against disasters. Readers will learn what observability is, real time monitoring, and post mortem process. The book also goes on to explain automation, Infrastructure as Code (IaC), CI/CD pipelines and the rise of AI to use in incident response and self-healing systems. Last, it covers organizational adoption of SRE, promotion of collaboration, error budgeting and managing multi cloud environments. Engineers, architects, and leaders who wish to instill reliability and resilience in modern software operations should read this book. Nº de ref. del artículo: 9798899843983
Cantidad disponible: 2 disponibles