SkillsLib.ai

Summarize Experiment

Extract and summarize experiment results into lightweight markdown reports

4.3(24 reviews)
100+ downloads
Updated Sep 2026
Verified SafeSecurity VerifiedThis skill was analyzed by our AI security scanner for harmful content including data exfiltration, system manipulation, credential theft, and prompt injection. No threats were detected.

What You Can Do

This skill parses experiment metadata and outputs to extract critical metrics from completed machine learning experiments. It reads experiment_summary.yaml to identify runs, pulls final training loss from SLURM logs, extracts accuracy metrics from inspect-ai evaluation files, and generates a clean summary.md report—all while logging the extraction process for reproducibility.

Features

Parse experiment_summary.yaml to identify fine-tuned and control runs with their hyperparameters
Extract final training loss from SLURM stdout logs automatically
Pull accuracy and performance metrics from inspect-ai .eval files
Generate summary.md with structured metrics and run metadata
Log all extraction steps to logs/summarize-experiment.log for audit trails
Support partial experiment results—handle incomplete runs gracefully
Map evaluation tasks and epochs to their corresponding results

Example Output

Example summary.md output:

Experiment Summary: bert-classification-v2

Run Status

Run NameTypeModelStatusTraining LossAccuracy
run-001-lr0.001fine-tunedbert-baseCompleted0.2450.924
run-002-lr0.0001fine-tunedbert-baseCompleted0.1980.931
control-baselinecontrolbert-baseCompleted—0.847

Key Findings

  • Best accuracy: 0.931 (run-002-lr0.0001, learning rate 0.0001)
  • Improvement over baseline: +8.4 percentage points
  • Training converged in all runs

Log entry example:

code
[2024-01-15 14:32:01] Parsing experiment_summary.yaml...
[2024-01-15 14:32:02] Found 3 runs (2 fine-tuned, 1 control)
[2024-01-15 14:32:03] Extracted training loss from run-001: 0.245
[2024-01-15 14:32:04] Extracted accuracy from eval/logs/run-002.eval: 0.931
[2024-01-15 14:32:05] summary.md generated successfully

What's Included

  • `summarize-experiment.md` instruction file with full workflow documentation:
  • YAML parsing template for reading experiment_summary.yaml structures:
  • Python extraction script (parse_eval_log.py) for inspect-ai eval file parsing:
  • summary.md template with markdown table and findings sections:
  • SLURM log extraction patterns and regular expressions:

Who It's For

  • Machine learning researchers conducting hyperparameter tuning experiments
  • ML engineers documenting model fine-tuning results for team review
  • Research scientists tracking multiple experimental runs for publications
  • AI labs automating experiment result documentation workflows
  • Data scientists creating reproducible experiment reports

Best For

  • Summarizing multi-run fine-tuning experiments with varied hyperparameters
  • Extracting metrics from completed inspect-ai evaluation workflows
  • Generating quick reference documents comparing control vs. fine-tuned models
  • Automating post-experiment documentation after run-experiment completion
  • Creating audit trails of experiment execution and metric extraction

You might also like

Kaizen Event Facilitator
$40
Kaizen3.3(6)
Kaizen Event Facilitator

You can structure and run Kaizen events from start to finish, guiding your team through waste analysis, root-cause identification, and rapid improvement design. The skill generates detailed event agendas, countermeasure documentation, and performance metrics—enabling you to close improvement opportunities within days rather than weeks.

ISO/Audit Companion for Manufacturing Quality Engineers
$30
ISO/Audit3.3(3)
ISO/Audit Companion for Manufacturing Quality Engineers

Prepare comprehensive audit documentation and manage ISO 9001/14001 compliance faster. You can generate tailored audit checklists, analyze nonconformance findings systematically, and develop corrective action plans that meet ISO standards — all in hours instead of days. This skill ensures consistent documentation across your manufacturing operations and maintains audit readiness year-round.

Plant Manager P&L Analysis & Variance Explanation
$35
P&L3.3(3)
Plant Manager P&L Analysis & Variance Explanation

Transform raw manufacturing P&L data into actionable insights by automatically identifying cost variances, pinpointing root causes, and generating executive-ready narratives. You upload your plant's P&L statement, provide context about recent operational changes, and Claude analyzes the numbers to explain why costs moved and what actions drive profitability.

Chemical Process Troubleshooting & Parameter Analysis
$25
Chemical3.3(3)
Chemical Process Troubleshooting & Parameter Analysis

This skill systematically analyzes out-of-spec chemical manufacturing batches to identify root causes and recommend parameter adjustments. You provide process data, operating conditions, and quality metrics; Claude performs hypothesis testing, statistical correlation analysis, and multi-variate optimization to pinpoint the failure mechanism. You get actionable recommendations ranked by impact and implementation risk.

Film Production Coordination & Problem Solver
$40
Film3.5(6)
Film Production Coordination & Problem Solver

You can leverage Claude to streamline every aspect of film production coordination—from managing crew schedules and budgets to troubleshooting on-set logistics and creative challenges. This skill helps you create call sheets, coordinate locations, track production expenses, and quickly solve problems that arise during shooting without disrupting your workflow.

Food & Beverage Process Optimization & Root Cause Analysis
$40
Food & Beverage Process Optimization & Root Cause Analysis

You provide production data and process details, and this skill systematically analyzes the information to identify root causes of quality issues, yield losses, or equipment failures. It develops prioritized corrective actions with implementation timelines, compliance documentation, and performance metrics to track improvement effectiveness.

Research Synthesis & Gap Analysis
$25
Research Synthesis & Gap Analysis

Accelerate literature reviews by automatically synthesizing research papers, extracting key methodologies and findings, and identifying research gaps. This skill analyzes multiple papers simultaneously to spot unexplored areas, methodological blind spots, and emerging opportunities in your field. You'll get structured analysis, comparison matrices, and research landscape summaries that would normally take weeks to produce manually.

ISO Audit Preparation & Non-Conformance Management
$25
ISO/Audit3.4(5)
ISO Audit Preparation & Non-Conformance Management

Prepare comprehensive audit documentation, track non-conformance issues from identification through resolution, and generate corrective action plans. You can document audit findings, prioritize risks by severity, assign accountability, and create formal reports that demonstrate compliance readiness to auditors and stakeholders.

$36.00$45.00