
Network Troubleshooting & Incident Response Assistant
Diagnose network issues and generate incident resolution workflows
What You Can Do
You can quickly diagnose network problems by providing symptoms and context, and Claude will generate structured troubleshooting workflows tailored to your infrastructure. The skill creates detailed incident documentation, root cause analysis templates, and solution guides that you can immediately apply or add to your knowledge base for future reference.
Features
asks clarifying questions to narrow down root cause from symptoms
generates step-by-step diagnostics organized by layer (physical, data link, network, transport, application)
documents findings, contributing factors, and preventive measures
creates clear, actionable fixes with rollback procedures and validation steps
automates notification, escalation, and communication templates
formats incident timeline and calculates mean time to resolution
auto-generates searchable postmortem summaries for future reference
recommends when and how to escalate based on severity and scope
Example Output
Example 1: Connectivity Issue Diagnosis
## Issue Summary
Loss of connectivity to production database server
## Diagnostic Questions
1. Is the connectivity issue affecting all clients or specific ones?
2. When did the issue start and was there a recent change?
3. What is the network path (direct LAN, VPN, cloud)?
## Troubleshooting Workflow
### Layer 1: Physical
- [ ] Verify interface status: `ip link show` / `ifconfig`
- [ ] Check cable connections and switch ports
- [ ] Review interface errors: `ethtool -S eth0`
### Layer 3: Network
- [ ] Test connectivity: `ping 10.0.1.50`
- [ ] Verify routing: `traceroute 10.0.1.50`
- [ ] Check ARP table: `arp -a`
Example 2: Root Cause Analysis Output
## Root Cause: STP Topology Change
After a switch firmware upgrade, rapid spanning tree convergence caused a 45-second broadcast storm.
## Contributing Factors
- No BPDU guard configured on access ports
- Network monitoring did not alert on topology changes
- Change window did not include connectivity validation
## Preventive Measures
1. Enable BPDU guard on all access ports
2. Add STP convergence alerts to monitoring
3. Require pre/post connectivity tests in change procedures
Example 3: Solution Documentation
## Resolution Steps
1. **Isolate affected segment** — Move impacted VLAN to isolated switch port
2. **Clear MAC table** — `clear mac-address-table dynamic`
3. **Reset STP** — Reboot switch in maintenance window
4. **Validate connectivity** — `ping` and `traceroute` all critical paths
5. **Restore production** — Merge VLAN back to production switch
What's Included
- SKILL.md: Complete troubleshooting skill with diagnostic workflows
- Incident Response Template: Structured format for documenting issues and resolutions
- Root Cause Analysis Worksheet: Guided template for identifying contributing factors
- Troubleshooting Checklists: Layer-by-layer diagnostics for common scenarios (connectivity, DNS, latency, packet loss)
- Solution Documentation Format: Steps, rollback procedures, and validation checklist
- Postmortem Template: Incident summary, timeline, findings, and preventive actions
- Escalation Decision Tree: Criteria for when to escalate and to whom
- Knowledge Base Entry Template: Auto-formatted summaries for searchable reference
Who It's For
- Network Engineers — Quickly diagnose complex multi-layer issues and document solutions
- IT Support Specialists — Provide structured troubleshooting to end-users with consistent documentation
- Infrastructure/DevOps Teams — Generate runbooks and postmortems to reduce MTTR and prevent recurrence
- System Administrators — Build institutional knowledge by creating searchable incident archives
- On-call Engineers — Access guided troubleshooting during incidents to reduce diagnosis time
Best For
- Network outage diagnosis and resolution — Root cause complex connectivity and routing issues
- Incident documentation and postmortems — Create thorough incident records with preventive actions
- Creating runbooks and troubleshooting guides — Generate layer-by-layer diagnostic workflows for common scenarios
- Root cause analysis — Identify contributing factors and systemic improvements
- Knowledge base building — Auto-generate searchable incident summaries for future reference






