Verwandte Artikel zu Observability in the AI-Native Era: Leveraging AIOps...

Observability in the AI-Native Era: Leveraging AIOps to build, observe, and operate resilient systems - Softcover

Hilliary Lipsig; Andreas Grabner; Robert Rati

 
9781806389599: Observability in the AI-Native Era: Leveraging AIOps to build, observe, and operate resilient systems

Inhaltsangabe

Endorsements

“Observability in the AI-Native Era is a practical, timely guide for DevOps, SRE, platform engineering, and technology leaders looking into the adoption of AI in their teams. It connects observability fundamentals with AIOps, agentic platforms, trust, governance, cost, and security, without treating AI as magic.”- Pini Reznik, CEO & Co-Founder, re:cinq

“This book provides a great, insightful exploration of observability, SRE, and AI, demonstrating how AI can enhance collaboration, accelerate troubleshooting, and effectively summarize events to improve operational tasks.”-Alexandra Franz, GEO Lead at Dynatrace

Key Features

  • Bridges observability and AI into a unified operational approach rather than treating them as separate domains
  • Uses a continuous case study to connect concepts across chapters and reflect real-world engineering scenarios
  • Focuses on evolving operational maturity from reactive to proactive and preventive systems
  • Purchase of the print or Kindle book includes a free PDF eBook

Book Description

Observability is mandatory for building and operating cloud-native distributed systems. Tools like OpenTelemetry have standardized how observability data is sourced, and AI now transforms how we extract value from the vast amounts of observability data generated by modern systems. This book guides you in implementing scalable observability, improving engineering efficiency with AI, and integrating observability throughout the Software Development Lifecycle (SDLC) via modern self-service internal developer platforms.
You'll start with observability basics and learn how AIOps enhances signal correlation, anomaly detection, and root-cause analysis. Using real-world examples, the book demonstrates how to implement AIOps, build proactive detection pipelines, and automate diagnostics and remediation. You'll explore best practices for expanding observability using OpenTelemetry, Prometheus, Grafana, Dynatrace, Datadog, and New Relic alongside machine learning models, ensuring your systems are accurate, efficient, and secure.You'll also learn how to benchmark, measure, and secure your AIOps implementation, and gain a practical understanding of software compliance and how it applies to your systems. By the end of this book, you'll be ready to design and deliver AIOps-enabled observability solutions that make cloud-native systems more resilient, efficient, and secure.

Table of Contents

  1. Observability: The Art of Turning Data into Insights
  2. The Elephant in the Room: Artificial Intelligence
  3. From Observability to AIOps and the Use Cases it Solves Today4.
  4. ACME Financial Services: Implementing AIOps
  5. Democratizing Observability: A Primer to Self-Service Platforms
  6. The Observability Agent: Real-Life Use Cases
  7. ACME Financial Services: How to Move from AIOps to Agentic Platforms
  8. Evolving Operations: Proactive > Preventive > Self-Driven Architecture
  9. No Future Without Challenges
  10. ACME Financial Services: How Will the AI Future Shape Our Company?

Die Inhaltsangabe kann sich auf eine andere Ausgabe dieses Titels beziehen.

Über die Autorin bzw. den Autor

Hilliary Lipsig is an autodidact and start-up veteran who has frequently learned and applied technologies to get a job done. She's had her hand in every part of the application delivery process, honing her skills originally as a quality engineer. Hilliary is an IT polyglot, able to talk the lingo of both the Operations and Development teams. She's currently a senior principal site reliability engineer at Red Hat Inc., working on Kubernetes-based platforms. She's passionate about GitOps, continuous integration, scalable processes, consistency in tooling, and good developer documentation. Her open source activities include contributions to the CNCF Glossary, and she's a member of the Code of Conduct Committee for the Cloud Native Computing Foundation (CNCF).

Andreas Grabner is a technical advocate for making distributed systems observable and making automated data-driven decisions across the software development lifecycle. In his capacity as a CNCF ambassador and a DevRel at Dynatrace, he connects and educates global software engineering communities on building and continuously validating digital services for resiliency, high availability, and security. Since his early days, he has been passionate about software quality and performance engineering, as it results in building excellent digital products. Andi uses his advocacy platforms to share best practices on topics such as observability, progressive delivery, DevOps, site reliability engineering, platform engineering, and digital business operations!

Robert Rati is a software and platform engineer veteran of small, medium, and large corporations in regulated industries ranging from wireless communications to the financial sector. He is passionate about reducing noise and enabling teams to focus on creating business value. He emphasizes maintainability, consistency, user friendliness, and productivity when planning and implementing projects. He is currently an engineering manager with Second Front.

„Über diesen Titel“ kann sich auf eine andere Ausgabe dieses Titels beziehen.