Screened · automated checks passed

finetuning

from evo-hq/evo · ★ 1,337 on GitHub · found by our crawler 2026-07-07

What the author says it does

This skill should be used when picking or diagnosing a training move (SFT, LoRA, DPO/KTO/ORPO, RFT, GRPO/PPO/RLOO, RLHF), or when the user mentions fine-tuning, post-training, training recipe, reward design, or weight updates. Decision tree by reward shape, smoke-run gate, three failure diagnostics, five false-progress patterns. Provider recipes and I/O contract in references/.

Quoted from the skill's own SKILL.md trigger description — this is what tells Claude when to activate it. Not yet verified by us.

Automated screening

100/100 validator score

Scored by the same rules as our free SKILL.md validator: trigger description quality, body substance, structure. Automated — a human bench test is the next step in the pipeline.

Install (unverified — review first)

git clone https://github.com/evo-hq/evo
# skill lives at: plugins/evo/skills/finetuning/SKILL.md

SkillProof status

This skill is in our test queue. We install every skill in a clean environment, run a trigger battery and score output against a baseline before it earns a catalog page — the full protocol is public. Until then, treat it like any unreviewed dependency: read the SKILL.md and any scripts before installing.

Already tested in Design & Frontend

  • Codex GPT Image ★ 10.0/10 Generates gpt-image-2 images via local Codex OAuth, no OPENAI_API_KEY needed.
  • Slide Wright ★ 10.0/10 Proposes an invented theme and a real 2-slide preview before it ever builds the full deck.
  • Compose App Icon ★ 10.0/10 uv-bundled CLI pair that authors and JSON-Schema-validates Apple Icon Composer .icon packages
  • HTML Explainer ★ 10.0/10 Builds polished, source-grounded, playable single-page or slide-deck web explainers with a real quality loop.

Top 10 Design skills, ranked →