infra-sre

Generates SRE and DevOps interview questions covering Linux troubleshooting, Kubernetes, CI/CD, and SLO design.

23|1|Updated Aug 3, 2026
One-click install
npx skills add https://github.com/yuecao365/OfferCome --skill infra-sre-yuecao365
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: infra-sre
Source: https://github.com/yuecao365/OfferCome/tree/main/src/lib/mock-interviews/skills/infra-sre
Command: npx skills add https://github.com/yuecao365/OfferCome --skill infra-sre-yuecao365

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve? Interviewers and hiring teams often struggle to design realistic, scenario-based questions for SRE, DevOps, and infrastructure roles, defaulting to trivia that fails to reveal real operational experience. ## Core Features & Use Cases - Scenario-Based Question Banks: Provides graduated question ladders across 13 domains including Linux troubleshooting, TCP/HTTP networking, containers, Kubernetes operations, CI/CD, IaC, monitoring, SLO, incident response, capacity planning, change management, HA/DR, and platform automation. - Signal Evaluation Rubrics: Each topic lists red-flag answers and strong-signal answers so interviewers can calibrate candidate scoring consistently. - Resume-Triggered Probing: Maps resume claims (e.g., "built CI/CD platform", "99.9% availability") to targeted follow-up questions about scale, incidents, and measurable outcomes. - Use Case: When a candidate's resume mentions Kubernetes cluster management, load this Skill to generate questions about Pod CrashLoopBackOff diagnosis, request/limit governance, and upgrade risk handling. ## Quick Start Load this Skill when the candidate's role or resume mentions SRE, DevOps, platform engineering, or cloud infrastructure, then ask it to generate a troubleshooting question about a high-load Linux server with full evaluation criteria.

Frequently Asked Questions about infra-sre

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I interview SRE candidates with scenario-based questions?▼

Present a concrete symptom such as high load with low CPU usage, then ask the candidate for a diagnostic sequence and at least three possible causes with verification methods. Strong candidates layer their investigation across process states, iowait, and cgroup limits rather than guessing.

What topics should a DevOps interview cover?▼

Cover Linux and network troubleshooting, container and Kubernetes operations, CI/CD and release system design, infrastructure as code, monitoring and alerting design, SLO and error budgets, incident response with full timelines, capacity planning, and change management.

How to evaluate Kubernetes interview answers?▼

Check whether the candidate rolls back before investigating, reads events via kubectl describe, compares nodes and image versions, and understands how resource requests and limits affect scheduling and QoS. Candidates who only know kubectl get pods signal shallow experience.

What are red flags in SRE interview answers?▼

Red flags include inability to reconstruct an incident timeline, treating buffer/cache as memory leaks, blaming TIME_WAIT without analysis, alerting on raw CPU thresholds, and defining SLOs without stating the numerator and denominator of the SLI.

When should this Skill be loaded during mock interviews?▼

Load it when the job description or resume mentions SRE, DevOps, platform engineering, cloud-native infrastructure, or operations development. It is a domain-layer Skill activated by role keywords rather than a general-purpose interviewer.