Academic Experiments
Audit and verify experimental evidence for research papers with reproducibility checks.
Installation
- Make sure Claude is on your device and in your terminal.
Skills load from
~/.claude/skills/when Claude Code starts up β so you need it on your machine first. If you don't have it yet, install it once with the command below, then runclaudein any terminal to verify.One-time setupnpm i -g @anthropic-ai/claude-codeAlready have it? Skip ahead.
- Paste into Claude Code or into your terminal.
This copies the whole skill folder into
~/.claude/skills/academic-experiments-joshua-zyy/β the SKILL.md plus any scripts, reference docs, or templates the skill ships with. Safe default: works for every skill.Faster alternative (instruction-only skills)
Skips the clone and grabs only the SKILL.md file. Don't use this if the skill ships Python scripts, reference markdowns, or asset templates β they won't be downloaded and the skill will fail when it tries to load them.
Quick install (SKILL.md only)Sign up to copy - Restart Claude Code.
Quit and reopen Claude Code (or any other agent that loads from
~/.claude/skills/). New skills are picked up on startup. - Just ask Claude.
Skills auto-activate when your request matches the skill's description β no slash command needed. Trigger phrases live in the skill's own frontmatter; you can read them in the βWhat this skill doesβ section above.
Prefer to read the source first? Open on GitHub.
When Claude uses it
Use when auditing, running, or verifying the experimental evidence behind a CS/AI/ML paper β builds an evidence inventory, flags protocol risks like data leakage or missing baselines, and runs minimal reproducible checks without full retraining.
What this skill does
Academic Experiments
skill ""οΌοΌ
Router Protocol
- Read
manifest.yaml. It declaresalways_loadfiles,axes, andreferences.on_demand. - Read every file listed under
always_load. These are the skill's binding rules β not reference material. - Apply the loaded material as constraints:
stance.mddefines non-negotiable rules, evidence type semantics, failure degradation, and scope.red-lines.mddefines absolute prohibitions. Do not negotiate these.output-contract.mddefines deliverables per mode and claim-readiness classification.anti-patterns.mddefines known failure modes and their correct alternatives.
- Detect the mode using the manifest's
modeaxis:experiment-evidence-pass,evidence-inventory-only, orminimal-reproducible-run. Align evidence type semantics to../shared/core/evidence-policy.md. - Echo the selected mode to the user before executing.
- Reach for
references/only when the manifest'sreferences.on_demandcondition is satisfied.
Modes
| Mode | Use when |
|---|---|
experiment-evidence-pass | Full audit: inventory + run + record + risk analysis |
evidence-inventory-only | Inventory existing artifacts only, no execution |
minimal-reproducible-run | Execute minimal reproducible command (e.g. eval existing checkpoint) |
Agent Dispatch
agents/experiment_agent.md is dispatched by academic-paper-writer orchestrator at Step 4. The agent may run experiments but must not modify project source code or data files, nor write paper prose independently.
Independent Use
| Input | Mode | Priority | Behavior |
|---|---|---|---|
repo_path + no run mode | experiment-evidence-pass | 2 (path trigger) | Full audit: inventory β env β minimal run β risk |
repo_path + "inspect only" | evidence-inventory-only | 1 (explicit) | Inventory only, no commands |
repo_path + specific command | minimal-reproducible-run | 1 (explicit) | Verify env β execute β record |
No repo_path | β | 3 (no input) | Ask path, or auto-detect entry files |
| Scenario | Recommended |
|---|---|
| Just auditing/reproducing evidence | This skill (standalone) |
| Writing results into paper prose | academic-paper-writer orchestrator |
| Draft results need verification | This skill β academic-reviser |
Related skills
Claude API Helper
anthropics
Build, debug, and optimize Claude API applications with caching and model migration support.
Customer Health Scorer
alirezarezvani
Analyze customer accounts to predict churn risk and identify expansion opportunities.
Adaptyv Protein Lab
foryourhealth111-pixel
Submit protein sequences for automated lab testing and validation experiments.
CLAUDE.md Optimizer
daymade
Optimize your CLAUDE.md file for clarity, efficiency, and maintainability.