What problem does it solve? Designing reliable, resilient, and highly available systems on Google Cloud requires navigating extensive architectural guidance. This Skill distills the Reliability pillar of the Google Cloud Well-Architected Framework into actionable principles, assessment questions, and validation checklists so you can evaluate and improve workload reliability without reading the full framework documentation. ## Core Features & Use Cases - Reliability Design Guidance: Covers nine core principles including SLO definition, resource redundancy, horizontal scalability, observability, graceful degradation, and blameless postmortems. - Workload Assessment: Provides structured questions to understand an organization's reliability requirements, constraints, and current practices. - Validation Checklist: Offers a concrete checklist to verify architectures against reliability recommendations such as cross-zone redundancy, autoscaling, backup testing, and circuit breakers. - Use Case: When reviewing a GKE-based application architecture, use this Skill to check for single points of failure, confirm SLOs are defined, and verify disaster recovery procedures are tested against RTO and RPO targets. ## Quick Start Evaluate the reliability of my Google Cloud architecture and recommend improvements for high availability and disaster recovery.