SkillsLib.ai

Evaluate

Review Claude's execution steps and provide structured feedback

3.7(29 reviews)
100+ downloads
Updated Oct 2026
Verified SafeSecurity VerifiedThis skill was analyzed by our AI security scanner for harmful content including data exfiltration, system manipulation, credential theft, and prompt injection. No threats were detected.

What You Can Do

You can trace Claude's decision-making process step-by-step through the Agent Execution Loop, seeing exactly which skills were matched, what was expected, which tools were called, how results were verified, and what was learned. Generate an interactive HTML review page where you write targeted feedback next to each execution step, then feed that structured feedback back into Claude for iterative improvements.

Features

Step-by-step execution tracing

view MATCH, THINK, ACT, VERIFY, LEARN phases with actual values and tool calls

Interactive HTML review interface

write feedback directly next to each step and export structured comments

Skill matching transparency

understand why a specific skill was chosen or why none matched

Tool call audit

see which tools were invoked, their order, and whether they ran in parallel or sequence

Verification checkpoint visibility

review what checks passed or failed and identify audit issues

Learning loop documentation

track what was updated in the skill or why no update occurred

Feedback parsing

Claude automatically parses your comments per step and applies targeted fixes

Session-based review

capture execution context from the current session without external logging

Example Output

MATCH Step: Selected 'Professional Writing' skill because user asked for a client email. No matching skills found for technical documentation.

THINK Step: Expected output: formal tone, 3-paragraph structure, call-to-action. Verification criteria: tone check via sentiment analysis, paragraph count validation.

ACT Step: Called email-generator tool (sequential), then grammar-check tool (parallel). Generated 285 words, processed through Hemingway API.

VERIFY Step: Passed tone check (formal: 92%), failed paragraph count (2 instead of 3). Skill audit: no security issues detected.

LEARN Step: Updated skill prompt to enforce paragraph structure with numbered examples.

What's Included

  • SKILL.md instruction file with Agent Execution Loop definitions:
  • HTML template with embedded CSS/JavaScript for interactive feedback collection:
  • Reflection worksheet for mapping MATCH, THINK, ACT, VERIFY, LEARN steps:
  • Feedback parser logic to extract and apply per-step improvements:
  • Timestamp-based file naming convention for managing multiple evaluations:

Who It's For

  • Professional writers refining client deliverables and optimizing writing workflows
  • Content strategists auditing content generation processes and skill effectiveness
  • AI prompt engineers debugging complex multi-step agent behaviors
  • Learning & development specialists analyzing how Claude processes training content
  • Quality assurance teams validating consistency and accuracy of automated writing tasks

Best For

  • Troubleshooting unsatisfactory writing outputs with precise, actionable feedback
  • Debugging skill selection logic when the wrong writing style was applied
  • Auditing tool chains to optimize writing workflows and eliminate redundant steps
  • Validating verification checkpoints in content quality assurance pipelines
  • Iteratively improving writing skills through structured evaluation cycles

You might also like

Manual QA Test Strategy & Execution Acceleration
$35
Manual QA Test Strategy & Execution Acceleration

You get AI-powered guidance for creating comprehensive test plans, designing effective test cases, and accelerating your manual QA execution. Claude helps you prioritize testing efforts based on risk, identify coverage gaps, standardize bug reporting, and optimize your QA workflows—so you test smarter, not just harder.

skill-platform-architecture-decisions
$30
Platform3.4(5)
skill-platform-architecture-decisions

This skill guides you through systematic evaluation of platform architecture options by analyzing trade-offs, scalability implications, and long-term maintenance costs. You'll generate decision matrices, risk assessments, and implementation roadmaps for critical architecture choices like microservices vs monolith, cloud providers, database strategies, and caching layers.

Skill for Manual QA Professional
$30
Skill for Manual QA Professional

Organize your manual testing process with AI-assisted test case generation, bug report templates, and quality assurance checklists. Get structured guidance on testing strategies, device/browser combinations to prioritize, and edge cases to cover—so you catch regressions before they ship and spend less time on documentation.

Manual QA Test Case & Coverage Planning
$35
Manual QA Test Case & Coverage Planning

You'll systematically create detailed test cases with clear preconditions, steps, and expected outcomes. This skill helps you map test coverage across features and requirements, identify gaps in your testing strategy, and organize test scenarios by risk and priority. You'll generate test matrices, data requirements, and dependency chains to ensure thorough manual QA before release.

$30.00