YoVDO

A Dashboard Is Worth a Thousand Words - Better Monitoring for Better Ops

Offered By: USENIX via YouTube

Tags

SREcon Courses Elasticsearch Courses Grafana Courses InfluxDB Courses

Course Description

Overview

Explore a conference talk from SREcon19 Asia/Pacific that delves into the transformative power of effective monitoring systems in large-scale scientific organizations. Learn how CERN implemented a new monitoring system, integrating metrics and logs for infrastructure and services using technologies like Kafka, Grafana, InfluxDB, and Elasticsearch. Discover how this initiative not only improved service operations but also fostered awareness of Site Reliability Engineering (SRE) practices among service managers. Gain insights into the design decisions, operational challenges in scaling the system to tens of thousands of hosts, and strategies for enhancing monitoring practices by introducing concepts such as Service Level Indicators (SLIs) and Service Level Objectives (SLOs). Understand the benefits derived from this approach and how it bridges the gap between traditional IT service operations and modern SRE practices in a scientific context.

Syllabus

SREcon19 Asia/Pacific - A Dashboard Is Worth a Thousand Words...


Taught by

USENIX

Related Courses

Maîtrisez les bases de données NoSQL
CentraleSupélec via OpenClassrooms
Implementando un motor con Alibaba Cloud y ElasticSearch
Coursera Project Network via Coursera
Learn DevOps: Advanced Kubernetes Usage
Udemy
Big Data on Amazon web services (AWS)
Udemy
Building an Elasticsearch Cluster with Amazon Elasticsearch Service on AWS
Pluralsight