/google-cloud-waf-operational-excellence
Generates operations-focused guidance for Google Cloud workloads based on the design principles and recommendations in the Operational Excellence pillar of the Google Cloud Well-Architected Framework (WAF). Use this skill to evaluate a workload, identify operational
$ npx -y skills add google/skills --skill google-cloud-waf-operational-excellence --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/google-cloud-waf-operational-excellence
Context preview
The summary Claude sees to decide when to auto-load this skill.
Generates operations-focused guidance for Google Cloud workloads based on the design principles and recommendations in the Operational Excellence pillar of the Google Cloud Well-Architected Framework (WAF). Use this skill to evaluate a workload, identify operational
SKILL.md
google-cloud-waf-operational-excellence.SKILL.mdname: google-cloud-waf-operational-excellence
metadata:
category: WellArchitectedFramework
description: >-
Generates operations-focused guidance for Google Cloud workloads based on the
design principles and recommendations in the Operational Excellence pillar of
the Google Cloud Well-Architected Framework (WAF). Use this skill to evaluate
a workload, identify operational requirements, and provide actionable
recommendations for deployment, monitoring, and incident management.
Google Cloud Well-Architected Framework skill for the Operational Excellence pillar
Overview
The operational excellence pillar in the Google Cloud Well-Architected Framework provides recommendations to operate workloads efficiently on Google Cloud. Operational excellence in the cloud involves designing, implementing, and managing cloud solutions that provide value, performance, security, and reliability. The recommendations in this pillar help you to continuously improve and adapt workloads to meet the dynamic and ever-evolving needs in the cloud.
Core principles
The recommendations in the operational excellence pillar of the Well-Architected Framework are aligned with the following core principles:
- **Ensure operational readiness**: Define and measure criteria for a workload
to be considered ready for production, including staffing, processes, and governance. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/operational-readiness-and-performance-using-cloudops.md.txt
- **Manage incidents and problems**: Establish structured processes for
incident response, communication, and root cause analysis to minimize impact and prevent recurrence. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/manage-incidents-and-problems.md.txt
- **Manage and optimize cloud resources**: Monitor resource utilization and
right-size environments to maintain performance while ensuring operational efficiency. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/manage-and-optimize-cloud-resources.md.txt
- **Automate and manage change**: Use Infrastructure as Code (IaC) and CI/CD
pipelines to ensure consistent, repeatable, and low-risk deployments and configuration changes. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/automate-and-manage-change.md.txt
- **Continuously improve and innovate**: Regularly review architectures,
monitor industry trends, and adapt operations to meet evolving business needs. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/continuously-improve-and-innovate.md.txt
Relevant Google Cloud products
The following are _examples_ of Google Cloud products and features that are relevant to operational excellence:
- **Observability and monitoring**
- **Cloud Monitoring**: Full-stack observability for Google Cloud and
hybrid environments.
- **Cloud Logging**: Real-time log management and analysis at scale.
- **Error Reporting**: Aggregates and displays errors for running cloud
services.
- **Service Monitoring**: Tools for defining and tracking Service Level
Objectives (SLOs).
- **Automation and CI/CD**
- **Cloud Build**: Serverless platform for building, testing, and
deploying software.
- **Cloud Deploy**: Managed continuous delivery service for GKE, Cloud
Run, and GCE.
- **Terraform / Infrastructure Manager**: Managed service for
Infrastructure as Code (IaC) automation.
- **Artifact Registry**: Central repository for managing build artifacts
and container images.
- **Resource management and optimization**
- **Recommender (Active Assist)**: Automatically identifies idle resources
and right-sizing opportunities.
- **Resource Manager**: Hierarchical management of resources across
organizations, folders, and projects.
- **Incident response**
- **Incident response & management (IRM)**: Structured tools and processes
for managing operational disruptions.
Workload assessment questions
Ask appropriate questions to understand operations-related requirements and constraints of the workload and the user's organization. Choose questions from the following list:
- **Operational readiness and performance**
- How do you define and measure operational readiness for your cloud
workloads and what specific criteria or metrics do you use?
- Describe your process for defining, tracking, and achieving SLOs for
your critical workloads.
- **Incident and problem management**
- Describe your incident management process, including roles,
responsibilities, and communication channels.
- How do you conduct post-incident reviews (PIRs) to identify root causes
and implement preventive measures?
- **Resource management and optimization**
- How do you ensure that your cloud resources are right-sized for your
workloads, and what tools or techniques do you use?
- **Change automation**
- Describe your change management process, including approval workflows,
testing procedures, and deployment strategies.
- How do you automate deployments, ensure their consistency and manage
configuration?
- **Continuous improvement**
- How do you ensure that your cloud operations are continuously adapting
to meet evolving business needs and technological advancements?
Validation checklist
Use the following checklist to evaluate the architecture's alignment with operational excellence recommendations:
- **Operational readiness**
- [ ] A formal framework or set of criteria exists to assess operational
readiness before production deployment
Read more
name: google-cloud-waf-operational-excellence metadata: category: WellArchitectedFramework description: >- Generates operations-focused guidance for Google Cloud workloads based on the design principles and recommendations in the Operational Excellence pillar of the Google Cloud Well-Architected Framework (WAF). Use this skill to evaluate a workload, identify operational requirements, and provide actionable recommendations for deployment, monitoring, and incident management.
Google Cloud Well-Architected Framework skill for the Operational Excellence pillar
Overview
The operational excellence pillar in the Google Cloud Well-Architected Framework provides recommendations to operate workloads efficiently on Google Cloud. Operational excellence in the cloud involves designing, implementing, and managing cloud solutions that provide value, performance, security, and reliability. The recommendations in this pillar help you to continuously improve and adapt workloads to meet the dynamic and ever-evolving needs in the cloud.
Core principles
The recommendations in the operational excellence pillar of the Well-Architected Framework are aligned with the following core principles:
- **Ensure operational readiness**: Define and measure criteria for a workload
to be considered ready for production, including staffing, processes, and governance. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/operational-readiness-and-performance-using-cloudops.md.txt
- **Manage incidents and problems**: Establish structured processes for
incident response, communication, and root cause analysis to minimize impact and prevent recurrence. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/manage-incidents-and-problems.md.txt
- **Manage and optimize cloud resources**: Monitor resource utilization and
right-size environments to maintain performance while ensuring operational efficiency. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/manage-and-optimize-cloud-resources.md.txt
- **Automate and manage change**: Use Infrastructure as Code (IaC) and CI/CD
pipelines to ensure consistent, repeatable, and low-risk deployments and configuration changes. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/automate-and-manage-change.md.txt
- **Continuously improve and innovate**: Regularly review architectures,
monitor industry trends, and adapt operations to meet evolving business needs. Grounding document: https://docs.cloud.google.com/architecture/framework/operational-excellence/continuously-improve-and-innovate.md.txt
Relevant Google Cloud products
The following are _examples_ of Google Cloud products and features that are relevant to operational excellence:
- **Observability and monitoring**
- **Cloud Monitoring**: Full-stack observability for Google Cloud and
hybrid environments.
- **Cloud Logging**: Real-time log management and analysis at scale.
- **Error Reporting**: Aggregates and displays errors for running cloud
services.
- **Service Monitoring**: Tools for defining and tracking Service Level
Objectives (SLOs).
- **Automation and CI/CD**
- **Cloud Build**: Serverless platform for building, testing, and
deploying software.
- **Cloud Deploy**: Managed continuous delivery service for GKE, Cloud
Run, and GCE.
- **Terraform / Infrastructure Manager**: Managed service for
Infrastructure as Code (IaC) automation.
- **Artifact Registry**: Central repository for managing build artifacts
and container images.
- **Resource management and optimization**
- **Recommender (Active Assist)**: Automatically identifies idle resources
and right-sizing opportunities.
- **Resource Manager**: Hierarchical management of resources across
organizations, folders, and projects.
- **Incident response**
- **Incident response & management (IRM)**: Structured tools and processes
for managing operational disruptions.
Workload assessment questions
Ask appropriate questions to understand operations-related requirements and constraints of the workload and the user's organization. Choose questions from the following list:
- **Operational readiness and performance**
- How do you define and measure operational readiness for your cloud
workloads and what specific criteria or metrics do you use?
- Describe your process for defining, tracking, and achieving SLOs for
your critical workloads.
- **Incident and problem management**
- Describe your incident management process, including roles,
responsibilities, and communication channels.
- How do you conduct post-incident reviews (PIRs) to identify root causes
and implement preventive measures?
- **Resource management and optimization**
- How do you ensure that your cloud resources are right-sized for your
workloads, and what tools or techniques do you use?
- **Change automation**
- Describe your change management process, including approval workflows,
testing procedures, and deployment strategies.
- How do you automate deployments, ensure their consistency and manage
configuration?
- **Continuous improvement**
- How do you ensure that your cloud operations are continuously adapting
to meet evolving business needs and technological advancements?
Validation checklist
Use the following checklist to evaluate the architecture's alignment with operational excellence recommendations:
- **Operational readiness**
- [ ] A formal framework or set of criteria exists to assess operational
readiness before production deployment
This repository contains Agent Skills for Google products and technologies, including Google Cloud. This repository is under active development.
Repo: google/skills
Other skills on google-skills.
- /data-manager-api-audience-ingestion
Guides developers through managing (adding, removing, and clearing) audience members for Google products using the Data Manager API and its associated client libraries. Use this skill when the user wants to upload audience members, remove specific users, or clear/replace an
Open skill - /data-manager-api-event-ingestion
Guides developers through implementing event and conversion ingestion to Google products using the Data Manager API /v1/events/ingest endpoint and its associated client libraries. Use this skill when the user wants to upload offline conversions, enhanced conversions for leads,
Open skill - /data-manager-api-setup
Guides developers through client library installation and authentication setup steps for the Data Manager API. Use this skill when a user is getting started with the Data Manager API and needs to setup their local environment, install the client library, or setup access to the
Open skill - /google-ads-api-account-diagnostics
Diagnoses Google Ads account performance issues such as conversion loss (value or volume), low lead flow/volume, and lost impression share (opportunities) due to ad rank, bids, or budgets. Use when troubleshooting sudden performance drops, analyzing campaign impression share
Open skill - /google-ads-api-mcp-setup
Guides developers through downloading, configuring, and installing the official open-source Google Ads MCP Server. Use this skill when a user wants to connect their AI assistant (such as Gemini, Claude Code, or Cursor) to their Google Ads account to query campaigns or retrieve
Open skill - /google-ads-api-quickstart
Guides developers through Google Ads API quickstart: credential setup, choosing from 6 client libraries/REST, configuring environments, and running a "retrieve campaigns" script. Troubleshoots common setup errors: USER_PERMISSION_DENIED, login_customer_id issues, and
Open skill

