Automation & Ops ✓ Google · Official google-cloudreliability-engineeringdisaster-recoverywell-architected-frameworkobservabilityhigh-availability

Google Cloud Reliability Architect

Evaluate and improve Google Cloud workload reliability using the Well-Architected Framework.

FollowSkills review · FSRS-2.0
Not recommended
57/ 100 5-point scale 2.9 / 5
Trust19 / 25 · 3.8/5

The skill only supplies Google Cloud reliability assessment questions, principles, and a checklist; it executes no commands, accesses no credentials, changes no resources, and exposes no data-exfiltration path. It also links the Google Cloud Well-Architected Framework as grounding material. Six points are deducted because sensitive-data handling, user confirmation, data-flow disclosure, rollback boundaries, and dependency security are not specified.

Reliability8 / 20 · 2.0/5

The objective, principles, questions, and validation checklist are broadly consistent, and there are no scripts or runtime dependencies. Eight points are awarded because there is no test coverage, abnormal-input handling, failure feedback, or key-path reproduction evidence, and this review did not execute the skill; the static calibration cap therefore applies.

Adaptability10 / 15 · 3.3/5

The intended audience and scenario are clear: reliability assessment for Google Cloud workloads across design, deployment, and operations. Five points are deducted because invocation triggers, non-fit boundaries, input/output formats, Chinese-language support, and mainland-China network reachability are not defined.

Convention10 / 15 · 3.3/5

The documentation is well organized with metadata, overview, principles, product examples, assessment questions, and a validation checklist. Repository context supplies installation, Apache-2.0 licensing, contribution, and support paths. Five points are deducted for missing skill-specific versioning, changelog, named maintenance responsibility, dependency notes, output examples, and troubleshooting guidance.

Effectiveness6 / 15 · 2.0/5

The questions and checklist can support an initial reliability review and cover SLOs, redundancy, scaling, observability, graceful degradation, recovery testing, and postmortems. Six points are awarded because the material is generic and does not define output format, prioritization, evidence requirements, or complete architecture-specific results; static review also provides no execution evidence of directly usable outputs.

Verifiability4 / 10 · 2.0/5

Each core principle includes a Google Cloud documentation path, providing some primary-source traceability. Four points are awarded because there are no committed tests, CI coverage, representative assessment outputs, or cross-source corroboration; this conclusion is based only on static source review.

Evidence confidence:Low Reviewed Jul 20, 2026 Reviewed revision 513a7a51e85f
The upstream repository has new commits since this review. The score still applies to the reviewed revision shown and may not cover the latest changes.
Before you use it
  • This is guidance based on a framework, not a completed architecture audit or reliability guarantee.
  • Before relying on it, provide workload-specific SLOs, RTO/RPO targets, data sensitivity, permission boundaries, and recovery evidence.
  • The grounding links point to Google Cloud documentation; confirm that the target users' network environment can reach those sources or provide accessible local copies.
Review evidence [1][2][3][4]
See the full review method →

What does this skill do, and when should you use it?

This skill focuses on the Reliability pillar of the Google Cloud Well-Architected Framework. It provides guidance for reliability, resilience, availability, redundancy, fault tolerance, and disaster recovery in Google Cloud workloads. Its coverage includes user-focused SLIs and SLOs, resource redundancy, horizontal scalability, observability, graceful degradation, recovery testing, data-loss recovery, and blameless postmortems. It also uses Google Cloud product examples and a validation checklist to support architecture reviews.

Generates guidance from the embedded reliability principles and recommendations for Google Cloud workloads; asks assessment questions about reliability targets, redundancy, scaling, monitoring, alerting, graceful degradation, failure recovery, data recovery, and postmortems; and applies a validation checklist covering SLIs/SLOs, cross-zone or cross-region redundancy, autoscaling, health checks, backups, circuit breakers, retries, chaos practices, and postmortem processes.

  1. A cloud architect designing a new Google Cloud workload needs guidance on availability, redundancy, and disaster recovery.
  2. A platform engineering team reviewing an existing system needs to identify single points of failure, scalability risks, and failover gaps.
  3. An SRE team defining SLOs, error budgets, monitoring, alerts, and user-experience measures needs a reliability framework.
  4. An operations team preparing regional failover, release rollback, or data-recovery exercises needs a readiness review.
  5. An engineering leader conducting an incident review needs a structured approach to root-cause analysis and recurrence prevention.

What are this skill's strengths and limitations?

Pros
  • Covers reliability design, operations, recovery testing, and organizational learning.
  • Includes concrete assessment questions and a validation checklist for architecture reviews.
  • Provides examples across compute, networking, storage, databases, operations, and disaster recovery.
  • Uses the Reliability pillar of the Google Cloud Well-Architected Framework as its basis.
Limitations
  • Focused on Google Cloud workloads rather than cloud reliability in general.
  • The source provides no test suite, automated assessment tool, or implementation scripts.
  • The source does not specify a platform compatibility matrix, permission requirements, or operating cost.
  • It requires adaptation to the workload’s architecture, business objectives, and organizational constraints; it does not replace live recovery exercises.

How do you install this skill?

Use the command provided in the repository README: npx skills add google/skills. During installation, select skills/cloud/google-cloud-waf-reliability. The README does not specify a more precise destination folder or a separate installation command for this skill.

How do you use this skill?

After installation, ask for a Google Cloud workload reliability assessment, for example: “Evaluate this Google Cloud architecture’s reliability and disaster recovery readiness, and identify improvements using the validation checklist.” Provide architecture details, SLOs, redundancy, observability, backup, and recovery information for more targeted guidance.

FAQ

Does this skill automatically change my Google Cloud environment?
The source describes guidance generation, assessment questions, and checklist-based evaluation; it does not describe modifying cloud resources.
Does it require payment or special permissions?
The source does not state any cost for the skill or any special permission requirements.
Which cloud platform does it support?
It explicitly targets Google Cloud workloads. The source does not describe adaptation for other cloud platforms.
Can it guarantee that a system meets its reliability targets?
No. It provides principles, questions, and checks; outcomes still depend on implementation, monitoring configuration, and actual failure and data-recovery testing.

More skills from this repository

All from google/skills

Automation & Ops ✓ Google · Official

Google Cloud Operational Excellence Advisor

Assesses Google Cloud workloads and recommends improvements using the WAF Operational Excellence pillar.

Automation & Ops ✓ Google · Official

Google Cloud Performance Advisor

Assess and improve workload performance using the Google Cloud WAF.

Automation & Ops ✓ Google · Official

Google Cloud WAF Security Advisor

Assesses Google Cloud workloads against Well-Architected security principles and produces actionable improvement guidance.

Automation & Ops ✓ Google · Official

Google Cloud Sustainability Advisor

Evaluates Google Cloud workloads against the WAF Sustainability pillar and recommends practical emissions-reduction actions.

Automation & Ops ✓ Google · Official

Google Cloud Cost Optimization Advisor

Generates actionable Google Cloud cost guidance using the Well-Architected Framework.

Automation & Ops ✓ Google · Official

GKE Production Readiness Review

Assess whether GKE clusters and workloads are ready for production.

Dev & Engineering ✓ Google · Official

Google Agents CLI Agent Development Guide

Guides agents from specification through development, deployment, and monitoring.

Automation & Ops ✓ Google · Official

Google Cloud Network Observability

Investigate Google Cloud network behavior with logs, metrics, and path diagnostics for VPC, NAT, and firewall issues.

Automation & Ops ✓ Google · Official

Google Cloud Authentication Guide

Choose secure Google Cloud authentication and authorization for local, production, and cross-cloud workloads.

Automation & Ops ✓ Google · Official

Google Cloud Solution Architect

Plan, validate, and package end-to-end architectures for complex multi-product Google Cloud workloads.

Automation & Ops ✓ Google · Official

Google Cloud Workload Manager Evaluator

Evaluate Google Cloud workloads against best-practice rules and review actionable findings.

Automation & Ops ✓ Google · Official

Google Cloud Live Multimodal Streaming Architect

Design and deploy Google Cloud solutions for live, bidirectional multimodal streams.

Automation & Ops ✓ Google · Official

GKE Production Golden Path

Set production-oriented GKE defaults, readiness checks, and decision guardrails for cluster design.

Automation & Ops ✓ Google · Official

Google Cloud AI Agent Builder

Design, implement, deploy, and validate AI agents and multi-agent systems on Google Cloud.

Automation & Ops ✓ Google · Official

GKE Enterprise RAG Search Architect

Designs and validates enterprise RAG search systems built on GKE and AlloyDB.

Dev & Engineering ✓ Google · Official

Gemini LiveAPI Client Service Skill

Generate a resumable, multimodal, bidirectional Gemini LiveAPI client over WebSockets.

Data & Analysis ✓ Google · Official

Google Cloud Agent Evaluation Flywheel

Improve model and agent quality through structured evaluation, failure analysis, and iteration.

Automation & Ops ✓ Google · Official

Google Cloud gcloud CLI Safety Skill

Helps agents manage and troubleshoot Google Cloud resources through validated, scoped gcloud commands.

Data & Analysis ✓ Google · Official

Cross-Cloud Agentic Analytics Architect

Design governed, secure agentic analytics for distributed data

Automation & Ops ✓ Google · Official

Agent Platform Endpoint Management

Manage Agent Platform serving endpoints and troubleshoot common endpoint failures.

Related skills