DevOps.Academy
CoursesCurriculumPricing

    Loading…

    ↑ ↓ move · ↵ open · Ctrl K anywhere

    ◔My learningSign in
    Join the waitlist
    DevOps.Academy

    Free, hands-on DevOps lessons. Made with care in India.

    Learn

    All coursesLinux roadmapCurriculumMy learning

    Academy

    PricingCertificates

    Legal

    PrivacyTermsContact

    devops.rajeev.pro

    course 08 · intermediate

    Monitoring

    Know something is wrong before your users do, and know where to look when it is.

    Start lesson 1 →See the roadmap ↓
    Level
    Intermediate
    Roadmap
    9 stages · 29 topics
    Lessons
    1 of 9 ready
    Done by you
    0/9 stages
    Before this
    Linux & Bash
    Certificate
    Preview →
    before you start

    The Monitoring roadmap

    Everything this course covers, in the order you'll learn it. Open a stage to see why it matters and jump to its lesson.

    lesson ready coming soon done by you
    • Metrics
    • Logs
    01Observability basics+start here

    The three signals and the questions each one answers.

    Lesson 1: Metrics, logs and traces →
    • Traces
    • Golden signals
    • Architecture
    • Scraping & targets
    02Prometheus+

    The standard for metrics in cloud-native systems.

    Lesson 2: Prometheus · coming soon
    • Exporters (node_exporter)
    • PromQL
    • Data sources
    03Grafana+

    Turn raw metrics into dashboards people actually read.

    Lesson 3: Grafana · coming soon
    • Dashboards
    • Variables
    • Alert rules
    • Alertmanager
    04Alerting+

    Alerts that wake you up only when something is truly broken.

    Lesson 4: Alerting · coming soon
    • Routing & silences
    • Writing useful alerts
    • Loki
    05Logs+

    Search every server's logs from one place.

    Lesson 5: Logs · coming soon
    • Elastic Stack
    • Graylog & Splunk
    • OpenTelemetry
    06Tracing+

    Follow one request through many services.

    Lesson 6: Tracing · coming soon
    • Jaeger
    • kube-prometheus-stack
    07Kubernetes monitoring+

    Watch the cluster and every pod on it.

    Lesson 7: Kubernetes monitoring · coming soon
    • Cluster metrics
    • Pod metrics
    • SLIs & SLOs
    08SLOs & on-call+

    Decide how reliable is reliable enough, and how to respond when you're not.

    Lesson 8: SLOs & on-call · coming soon
    • Error budgets
    • Runbooks
    • Datadog
    09Hosted tools+

    What the paid platforms add, and when they're worth it.

    Lesson 9: Hosted tools · coming soon
    • New Relic
    • Dynatrace
    Course complete · ready for the next one

    Stage order follows the community roadmap at roadmap.sh/devops, trimmed to what DevOps work needs.