Network Troubleshooting & Diagnostic Analyzer
Diagnose complex network issues with structured analysis and actionable solutions
What You Can Do
You systematically isolate network problems by analyzing symptoms, logs, and infrastructure context through a structured diagnostic framework. Claude performs multi-layer analysis across physical, data link, network, and application layers to identify root causes and provide prioritized remediation steps. From intermittent connectivity drops to performance degradation to configuration drift, you get specific, actionable solutions with verification and rollback strategies.
Features
Parse your problem description and systematically narrow down possibilities across all network layers
Extract meaningful patterns from syslog, event logs, packet captures, and device output
Contextualize your infrastructure (routers, switches, firewalls, services) to pinpoint scope of impact
Analyze Layer 1 (physical), Layer 2 (switching/VLAN), Layer 3 (routing/BGP), and Layer 7 (application) simultaneously
Compare actual device configs against expected settings and flag misconfigurations
Identify degradation patterns and anomalies relative to historical metrics
Ranked actions with impact estimates and rollback procedures for each fix
Build structured reports for compliance, knowledge transfer, and future reference
Example Output
Example 1: Intermittent DNS Failures
## Root Cause: Primary DNS server degradation
FINDINGS:
- DNS response time: 450ms (baseline: 15ms, 30x slower)
- Primary DNS server memory: 92% utilization
- Secondary failover threshold: 1000ms (not triggered)
REMEDIATION (in priority order):
1. Restart primary DNS service (impact: immediate, risk: low)
2. Increase DNS cache size 1GB → 2GB (impact: medium-term, risk: minimal)
3. Lower failover timeout to 500ms (impact: prevents future flaps, risk: low)
4. Deploy tertiary DNS in separate AZ (impact: redundancy, effort: medium)
Example 2: BGP Route Instability
## Root Cause: BGP session flapping + suboptimal routing
SYMPTOM: Traffic to Region-2 experiencing 150ms latency spike
FINDINGS:
- BGP session on R2 dropped 47 times in 2 hours
- AS_PATH preference: primary (3 hops) vs alternate (2 hops)
- Router R2 config shows aggressive timers (keepalive 10s)
ACTIONS:
1. Increase BGP keepalive 30s → 60s on R2 (immediate stability)
2. Adjust LOCAL_PREF to favor shorter AS_PATH (optimal routing)
What's Included
- SKILL.md: Diagnostic framework with symptoms-to-causes decision trees and checklists
- Diagnostic worksheet: Structured form for capturing symptoms, timeline, scope, and infrastructure context
- Log parsing reference: Common syslog formats, NetFlow fields, packet capture interpretation
- Network topology template: YAML/diagram format for documenting your infrastructure
- Remediation playbook: Prioritization matrix and rollback procedures for safe fixes
- Common issues guide: RCA patterns by symptom (connectivity, performance, config, security)
- Escalation matrix: When to involve network ops, vendor support, or security teams
Who It's For
- Network engineers — Daily troubleshooting of connectivity, routing, and performance issues
- Systems administrators — Diagnosing problems in on-prem or hybrid infrastructure
- DevOps engineers — Analyzing cloud networking, service mesh, and container networking issues
- Infrastructure architects — RCA for capacity planning and post-incident reviews
- Security engineers — Investigating potential network security incidents or policy violations
Best For
- Intermittent connectivity and timeout issues — Users randomly losing connection or experiencing brief outages
- Performance degradation investigations — Unexplained latency spikes or throughput reduction
- Configuration drift detection — Validating device settings after changes or migrations
- Multi-layer problem diagnosis — Issues that span hardware, routing, firewall, and application layers
- Post-incident root cause analysis — Understanding what failed during an outage for prevention






