Compatibility
Does it work in your stack?
Separate declared requirements, static findings, and what the lab has actually tested.PRE-PRODUCTION TESTING FOR AGENT SKILLS
Run a skill across models and agent environments. Compare it with the base model. Inspect every run. Blind-test subjective outputs. Know what you're installing before it reaches production.
For vibe coders, agent builders, and the people shipping their work.
Available now: text + bundled-reference lab environments.
Native runtime execution and image testing are not yet available.
Illustrative sequence only. No AI calls, measurements, or benchmark evidence.
FOR VIBE CODERS AND AGENT BUILDERS
A promising SKILL.md can still be a poor fit for your model, tools or taste. Stop downloading skills and debugging blind. Compare a matched baseline, inspect requirements and traces, and see what actually changed before spending hours tweaking.
Import from GitHub →Use Discover Skills → Import skill in the workspace. Native runtimes are not executed by this build.FOUR QUESTIONS. SEPARATE ANSWERS.
Does it work in your stack?
Separate declared requirements, static findings, and what the lab has actually tested.Does it work repeatedly?
Repeat scenarios. Look for failed checks, execution errors, and inconsistent outputs.Does the skill actually improve the base model?
Compare matched conditions. Equal check scores can mean your checks missed the difference.Would you actually use the output?
Make the human call separately. A compliant response can still be the wrong response for you.THE TEST PATH
Start with a small text comparison. Package size, model response times and test scope affect how long it takes.
SOURCE → SNAPSHOTImport a public GitHub directory, ClawHub version, SKILL.md or ZIP. Keep the instructions, supporting files and immutable provenance together.
PUBLIC SKILL LIBRARY
Start with a genuine indexed package.
Then decide whether it belongs in your rig.
Official means the publisher’s own repository; it is not a SkillRig endorsement. Empty filters mean no indexed match, not no such skills exist.
Current harness: text + bundled references. Ecosystem tags are source/static assessments, not native runtime test results.
Applies Anthropic's official brand colors and typography to any sort of artifact that may benefit from having Anthropic's look-and-feel. Use it when brand colors or style guidelines, visual formatting, or company design standards apply.
Detected capabilities: text
Use when writing or editing prose humans will read — documentation, commit messages, error messages, READMEs, reports, UI text, explanations. Triggers on clarity concerns, wordiness, passive voice, or vague language.
Detected capabilities: text, references
HUMAN PREFERENCE / BLIND COMPARISON
Hide the identities. Read the outputs.
Choose what works for you.
TASK / Announce a tool for testing Agent Skills before installation. Keep it clear and concise.
Evaluate Agent Skills before installation. Compare outputs across models, review execution traces, and record your preferences to inform deployment decisions.
Identity hiddenBefore you add a skill to your agent, put it on the rig. Compare what changes, inspect the run, and decide whether you'd actually use the output.
Identity hiddenTry a choice. There is no universal winner.
Handwritten examples, not model results. This interaction saves no preference profile. Human preference remains separate from automated compliance.
OBSERVABILITY / RUN TRACE
Follow the instructions supplied, reference reads, model responses and explicit checks. Missing usage stays unavailable. Errors stay visible.
Testing availability ↗Current tool access is limited to imported text references. The lab does not run shell commands, browse the web, or execute imported code.
SkillLoadedSKILL.mdExact package snapshot recorded
ContextPreparedMatched task + reference manifestPrimary instructions added to this condition
ReferenceReadreferences/voice.mdRead-only · imported package path
ModelResponseOutput capturedModel identity and reported usage retained when available
EvaluationConfigured checks evaluatedHuman preference assessed separately
Illustrative event sequence; no real invocation or timing measurements.
SKILL CONTRIBUTION / MATCHED CONDITIONS
Same task. Same model. Same references and tool access. The primary skill instructions are the variable.
Meet your all-in-one AI testing platform. Generate images, automate every workflow, and guarantee production-ready results.
SkillRig helps you inspect Agent Skills and compare text outputs before installation. Review the evidence, then decide what to use.
An illustration of a difference worth checking, not measured uplift. Real tests may show improvement, no difference, or regression. Literal checks alone do not establish factual accuracy.
REPRODUCIBLE INPUTS / HONEST LIMITS
Keep the source version and test configuration with the result. A report is evidence for your decision, not security certification.
Source provenance is from a real indexed package. Model and run settings below are illustrative; this public example has not been executed.
QUICK START / CURRENT BUILD
For vibe coders and experienced builders alike: replace “will this skill work for me?” with a comparison you can inspect.
Testing availability, choose a skill in Discover Skills, and select Test skill. Start with one model, one scenario and one repetition. Use DEMO to explore the workflow without AI calls.
Connect your own OpenRouter key in Settings, select LIVE, choose exact model IDs, and review the maximum budget before starting. API usage is billed to your OpenRouter account. A ChatGPT subscription does not fund these calls.
Not yet. These are skill ecosystem labels. Current tests use bundled-context and emulated reference-reader lab environments. Native runtime compatibility remains untested.
This is the public catalogue and product demonstration. Hosted testing and sign-in are not open yet. All examples are illustrative; this site makes no model calls.
INSPECT FIRST. SHIP WITH EVIDENCE.
Find out how it behaves before your production agent does.
The local lab supports free demos and live tests with your own OpenRouter key. Hosted testing is not open yet.TRY LOCALLY
Use your local SkillRig build to run tests. No tests or API keys are accepted on this public website. Local builds are currently shared privately; there is no public download or sign-up yet.
Already have the build? With Node.js 22 or later installed, double-click Start-Lab.cmd on Windows. Alternatively, run node server.mjs from the project folder and open the local address printed in the terminal. Start with Demo; Live requires your own OpenRouter key.