Site Reliability Engineering (SRE) helps keep large computer systems running without problems. This skill helps you set clear goals for system uptime and speed. It also shows how to track mistakes and automate daily tasks.
You can design incident response plans for when things go wrong. Use chaos engineering to test how systems fail and recover. Reduce repetitive work by automating it. This makes your systems more stable and your team more efficient.
Anyone who runs production services can benefit. You will learn to balance new features with system reliability. The tools and templates make it easy to apply right away.
Global
mkdir -p ~/.claude/skills/sre-engineerProject
mkdir -p .claude/skills/sre-engineerSource Repository
Google Agents Cli Deploygoogle/agents-cli
Deploy Google ADK agents to Agent Runtime, Cloud Run, or GKE with CI/CD
Turborepovercel/turborepo
Speed up your monorepo builds with Turborepo's caching and parallel task runs
Expo Cicd Workflowsexpo/skills
Write and validate EAS workflow YAML files for Expo CI/CD automation
Golang Continuous Integrationsamber/cc-skills-golang
Automate testing linting security and releases for Go projects with GitHub Actions
Docker Expertsickn33/antigravity-awesome-skills
Optimize, secure, and deploy Docker containers like a pro with proven techniques
Multi Stage Dockerfilegithub/awesome-copilot
Create efficient, secure Docker images using multi-stage builds and best practices
Diagnose Ci Failureswarpdotdev/common-skills
Diagnose CI failures fast and get a clear fix plan without touching code
Github Actions Templateswshobson/agents
Create production-ready GitHub Actions workflows for automated testing and deployment