SkillsLib.ai

Site Reliability Engineer (standalone)

0 skills available

$35
SRE Incident Management Framework

Rapid incident response, severity triage, and blameless postmortems

Incident Diagnosis & Root Cause Analysis via Observability Signals
$40
Incident Diagnosis & Root Cause Analysis via Observability Signals

Diagnose production incidents by correlating observability signals

Chaos Engineering Experiment Designer
$35
Chaos Engineering Experiment Designer

Design chaos experiments with controlled risks and validated hypotheses

Log Analyzer
$25
4.3(41)
Log Analyzer

Diagnose production incidents by correlating logs across services

SRE Incident Response Orchestration
$30
SRE Incident Response Orchestration

Orchestrate production incidents with structured response frameworks

Incident Investigation & Root Cause Analysis Using Observability Signals
$40
Incident Investigation & Root Cause Analysis Using Observability Signals

Correlate observability signals to pinpoint incident root causes

$30
Capacity Planning & Scaling Decision Framework for SREs

Turn infrastructure metrics into scaling decisions with forecasting analysis

Incident Response & Root Cause Analysis
$35
Incident Response & Root Cause Analysis

Correlate logs, traces, and metrics to surface root causes

SRE Capacity Planner
$35
SRE Capacity Planner

Forecast infrastructure needs and model scaling decisions with data analysis

Task Observer
$30
4.4(46)
Task Observer

Monitor task execution and capture skill improvement opportunities in real-time

Incident Response
$40
3.5(22)
Incident Response

Execute structured incident response from detection through blameless post-mortems

Postmortem Report Writer
$45
4.2(46)
Postmortem Report Writer

Generate blameless postmortem reports following Google SRE culture