Seed 42
Clean stabilization: early accepted edits, then skips once the admission rule stops improving.
SkillOpt recipe drilldown
This second-level page exposes the experiment behind the UDL story: SkillOpt edits a Markdown admission rule, evaluates each candidate through validation gates, and keeps the best behavior when later edits drift.
Headline metrics
The learned skills turn ambiguous news and scientific causal surfaces into TICKET-compatible admission bundles without promoting hypotheses into interventional proof. Seed 44 adds a useful stricter generalization signal.
Optimization trace
Run comparison
Seed 42 stabilizes with skips, seed 43 keeps proposing tie-score patches, and seed 44 rejects every post-step-5 candidate. All three preserve the same UDL validation winner.
Held-out examples
Seed drilldowns
Clean stabilization: early accepted edits, then skips once the admission rule stops improving.
Exploration pressure: many later candidate patches are rejected by the gate.
Sharper generalization check: validation converges, while valid-unseen exposes the harder edge.