Automated end-to-end testing for AI agent skills(agentskills.io). Launches Claude Code and Cursor as subprocesses, runs scenarios in real workspaces, and asserts what the model actually does.
-
Updated
Apr 12, 2026 - Python
Automated end-to-end testing for AI agent skills(agentskills.io). Launches Claude Code and Cursor as subprocesses, runs scenarios in real workspaces, and asserts what the model actually does.
B/O/E evaluation framework for SKILL.md-based LLM operating systems and agent hubs
Validate, pressure-test, and release-gate SKILL.md packages for OpenAI and portable agent runtimes.
Test and validate OpenClaw skills locally before publishing to ClawHub — checks triggers, env vars, scripts, and SKILL.md completeness without needing a live agent. Zero dependencies.
Agent Skill Infrastructure: quality check, behavior test runner, version awareness
Add a description, image, and links to the skill-testing topic page so that developers can more easily learn about it.
To associate your repository with the skill-testing topic, visit your repo's landing page and select "manage topics."