Screened · automated checks passed

experiment-audit

from wanshuiyin/Auto-claude-code-research-in-sleep · ★ 13,472 on GitHub · found by our crawler 2026-07-07

What the author says it does

Audit experiment integrity before claiming results. Uses cross-model review (external reviewer backend) to check for fake ground truth, score normalization fraud, phantom results, and insufficient scope. Use when user says \"审计实验\", \"check experiment integrity\", \"audit results\", \"实验诚实度\", or after experiments complete before writing claims.

Quoted from the skill's own SKILL.md trigger description — this is what tells Claude when to activate it. Not yet verified by us.

Automated screening

95/100 validator score

Scored by the same rules as our free SKILL.md validator: trigger description quality, body substance, structure. Automated — a human bench test is the next step in the pipeline.

Install (unverified — review first)

git clone https://github.com/wanshuiyin/Auto-claude-code-research-in-sleep
# skill lives at: skills/experiment-audit/SKILL.md

SkillProof status

This skill is in our test queue. We install every skill in a clean environment, run a trigger battery and score output against a baseline before it earns a catalog page — the full protocol is public. Until then, treat it like any unreviewed dependency: read the SKILL.md and any scripts before installing.

Already tested in Writing & Content

  • Humanizer ★ 9.6/10 Strips 33 documented AI-writing tells from text via a draft-audit-final rewrite loop.
  • Blog Outline Generator ★ 9.6/10 SERP-informed H2/H3 outline generator; skeleton only, not a full content brief.
  • Academic Paper Reviewer ★ 9.6/10 Simulates rigorous academic peer review across originality, methodology, results, and writing.
  • PR Writing Review ★ 9.6/10 Extracts before/after writing fixes and style lessons from GitHub PR review comments.

Top 10 Writing skills, ranked →