SkillsLib.ai

Experiment Design Assistant for Data Scientists

Design statistically rigorous experiments with preregistration and power analysis

0.0(0 reviews)
100+ downloads
Updated Sep 2026

What You Can Do

Design experiments with statistical rigor from the ground up. You'll generate power analyses, create randomization strategies, write preregistration documents, and get AI-assisted reviews of your experimental design to catch biases before data collection. Whether you're planning a clinical trial, A/B test, or academic study, you'll have a complete, defensible experimental protocol.

Features

Power analysis calculator

Determine sample size needed for 80% statistical power with customizable effect sizes and significance levels

Randomization strategy generator

Create stratified, blocked, or clustered randomization schemes tailored to your study design

Preregistration template

Auto-generate Open Science Framework (OSF)-compatible preregistration documents with all required sections

Sample size recommendations

Get quick reference tables for common experimental designs (t-tests, ANOVA, correlation)

Statistical test advisor

Get recommendations on which statistical tests match your hypotheses and experimental structure

Bias detection checklist

Identify selection bias, confounding, measurement bias, and other threats to validity

Hypothesis framework builder

Structure your research questions using the SMART hypothesis framework

Experimental design reviewer

Get detailed feedback on your study design with specific, actionable improvements

Example Output

Power Analysis Output

code
Study: A/B Test on Email Subject Lines
Effect Size (Cohen's d): 0.3 (small-to-medium)
Target Power: 80%
Significance Level (α): 0.05

Sample Size Per Arm: 176
Total Sample Size: 352
Duration: ~4 weeks (assuming 45 conversions/day)

Randomization Schema

code
Method: Stratified Random Assignment
Strata: By user region (US, EU, APAC)
Block Size: 4 (to maintain balance within regions)
Software: Use Python's random.seed(42) or R's blockrand

Assignment Table:
Region | Control | Treatment | Total
US     | 88      | 88        | 176
EU     | 58      | 58        | 116
APAC   | 30      | 30        | 60
TOTAL  | 176     | 176       | 352

Preregistration Snippet (OSF Format)

Study Title: Impact of Subject Line Personalization on Email Open Rates

Primary Hypothesis: Personalized subject lines increase open rates by ≥10% vs. generic lines.

Statistical Test: Two-sample proportion test (χ²)

Analysis Plan: Intent-to-treat analysis with logistic regression controlling for user region and account age.

What's Included

  • SKILL.md: Core experiment design system prompt and workflows
  • Power Analysis Calculator: Template with effect size interpretations and sample size tables
  • Preregistration Template: OSF-ready document covering all 22 AsPredicted.org fields
  • Randomization Strategy Guide: Stratified, blocked, and clustered assignment methods with code examples
  • Bias Detection Checklist: 15-point validity review covering internal, external, construct, and statistical conclusion validity
  • Statistical Test Reference: Quick matrix of tests by hypothesis type (correlation, comparison, regression)
  • Hypothesis Framework Worksheet: SMART hypothesis structure guide

Who It's For

  • Academic researchers designing dissertation studies or journal submissions
  • Data scientists in healthcare and biotech planning clinical or observational studies
  • Product managers and UX researchers designing A/B tests and user experiments
  • Clinical trial coordinators creating protocols that meet regulatory standards
  • Social scientists conducting behavioral experiments or field studies

Best For

  • Planning experiments before data collection — Set up your study design and get feedback before you invest time and resources
  • Calculating sample sizes — Ensure you have enough statistical power to detect your effect of interest
  • Creating randomization procedures — Eliminate selection bias with defensible assignment methods
  • Writing preregistration documents — Meet journal requirements and boost credibility with Open Science practices
  • Reviewing designs for validity threats — Catch confounds, measurement bias, and generalizability issues early

You might also like

Experiment Design & Statistical Analysis for Research Engineers
$45
Experiment Design & Statistical Analysis for Research Engineers

You can design statistically valid experiments with proper power analysis, choose the right statistical tests for your data type, analyze results while controlling for multiple comparisons, and generate publication-ready reports with accurate interpretation of findings. Claude helps you avoid common statistical pitfalls and ensures your experimental claims are well-supported by evidence.

Analytics Documentation Generator
$25
Analytics Documentation Generator

This skill automatically documents your entire analytics infrastructure by analyzing data sources, transformations, and outputs. You'll generate production-ready data dictionaries with field definitions, lineage maps showing data flow across systems, and transformation documentation that explains logic and dependencies. Save weeks of manual documentation work while keeping your analytics stack discoverable as it evolves.

BI Data Quality Investigator
$30
BI Data Quality Investigator

You'll systematically diagnose data quality problems by developing structured root cause analysis frameworks, calculating the true business impact, and creating reproducible validation tests. This skill walks you through hypothesis-driven investigation, data lineage analysis, and remediation planning — turning data issues into documented fixes and preventive measures.

RAG Pipeline Optimization with Claude
$40
RAG Pipeline Optimization with Claude

You can systematically evaluate and improve your RAG pipelines using Claude as a design partner. You'll analyze retrieval quality, identify bottlenecks in your embedding and chunking strategies, and receive actionable recommendations to reduce hallucinations and improve context relevance. By the end, you'll have a data-driven optimization plan tailored to your specific use case and performance metrics.

Data Quality Test Framework Builder
$35
Data Quality Test Framework Builder

You'll build comprehensive data quality test suites that validate transformations, detect anomalies, and document standards across your dbt and SQL pipelines. This skill generates production-ready test configurations, anomaly detection protocols, and validation rules that catch data issues before they impact analytics.

HEOR Evidence Synthesis & Dossier Builder
$35
HEOR3.6(5)
HEOR Evidence Synthesis & Dossier Builder

You can rapidly compile, organize, and format disparate health economic evidence—from clinical trials to cost-effectiveness analyses—into structured, regulatory-compliant dossiers. This skill maps your evidence to specific payer and HTA requirements, automatically generates evidence hierarchies, and produces submission-ready dossier outlines with formatting that meets regulatory standards for NICE, EUnetHTA, and other major bodies.

Payer Evidence Synthesis & HTA Builder
$30
Payer4.0(3)
Payer Evidence Synthesis & HTA Builder

You can structure comprehensive health technology assessments (HTAs) that organize clinical evidence, economic analyses, and regulatory considerations into evidence-based coverage recommendations. This skill helps you synthesize clinical trial data, health economic models, and real-world evidence into clear, defensible payer coverage determinations. You'll generate professional HTA reports that align with major frameworks like ICER, CADTH, and NICE standards.

ML Infrastructure Failure Analysis & Optimization
$25
ML Infrastructure Failure Analysis & Optimization

This skill helps you systematically diagnose failures in distributed ML training and serving infrastructure. You provide system logs, metrics, and error traces, and Claude performs structured root-cause analysis to identify the underlying issue—whether it's resource exhaustion, distributed system deadlock, data pipeline corruption, or model serving misconfiguration. You get a detailed diagnosis with remediation steps ranked by likelihood and implementation effort.

$35.00