Screened · automated checks passed

acl-reproducibility

from brycewang-stanford/Awesome-Journal-Skills · ★ 808 on GitHub · found by our crawler 2026-07-07

What the author says it does

Use when strengthening reproducibility evidence for an ACL paper reviewed through ACL Rolling Review, covering the Responsible NLP checklist end to end, hyperparameter and compute reporting, prompt and decoding disclosure for LLM experiments, data contamination auditing, variance across runs, and checklist-to-paper consistency.

Quoted from the skill's own SKILL.md trigger description — this is what tells Claude when to activate it. Not yet verified by us.

Automated screening

100/100 validator score

Scored by the same rules as our free SKILL.md validator: trigger description quality, body substance, structure. Automated — a human bench test is the next step in the pipeline.

Install (unverified — review first)

git clone https://github.com/brycewang-stanford/Awesome-Journal-Skills
# skill lives at: ACL-Skills/skills/acl-reproducibility/SKILL.md

SkillProof status

This skill is in our test queue. We install every skill in a clean environment, run a trigger battery and score output against a baseline before it earns a catalog page — the full protocol is public. Until then, treat it like any unreviewed dependency: read the SKILL.md and any scripts before installing.

Already tested in Data & Analytics

  • Statistical Analysis ★ 10.0/10 Enforces frame-inspect-check-assumptions-effect-size-APA-report pipeline for hypothesis tests; bundled Shapiro-Wilk/Levene's script actually runs.
  • CocoIndex ★ 10.0/10 Turns vague 'build me a RAG pipeline' requests into working CocoIndex flow code with the real API syntax.
  • Crawl4AI ★ 10.0/10 Wraps the Crawl4AI browser scraper for JS-heavy sites, batch crawls, and LLM-free schema extraction.
  • National Team Position ★ 10.0/10 Fetches real Shanghai Exchange ETF-share data to estimate China's 'national team' broad-base ETF positioning.

Top 10 Data skills, ranked →