Logging, Monitoring, and Observability in Google Cloud
In this 2-day course we’re taking a look at how to monitor and improve infrastructure and application performance in Google Cloud.
The course combines presentations, demos, hands-on labs and real-world case studies to give you experience with full-stack monitoring, real-time log management and analysis, debugging code in production, tracing application performance bottlenecks and profiling CPU and memory usage. Topics include creating dashboards, uptime checks and alerting policies, defining SLIs and SLOs, logs-based metrics and log sinks, Cloud Audit Logs, monitoring GKE with Managed Service for Prometheus, and monitoring the network with VPC Flow Logs. The course also looks at how to control the cost of observability. It includes seven labs.
SRE concepts, SRE best practices and incident response are not covered.
The course is aimed at cloud architects, administrators and SysOps personnel, as well as cloud developers and DevOps personnel. It is recommended that attendees have completed Google Cloud Fundamentals: Core Infrastructure or have equivalent experience, have basic scripting or coding familiarity, and be comfortable with command-line tools and Linux.