Evaluation loops
Loops that grade other work: eval suites and acceptance gates with explicit criteria.
5 graded loops in this category.
Keep only the lessons that help
Test one recorded lesson per run, keep evidence across runs, and drop guidance that stops paying off.
Attack a design until it holds
A critic hammers the design and a builder answers — every objection tracked, and none closed without evidence.
Separate fact from assumption
Split facts from assumptions, test falsifiable hypotheses, update confidence, and pick the next highest-information experiment.
Autonomous overnight ML research loop with stall detection (ARIS)
Framework-agnostic (Claude Code, Codex, OpenClaw, or any LLM agent), markdown-only skill bundle (79+ skills) for running ML research unattended overnight: literature search, idea generation, experiment execution, and cross-model paper review, with a silent-death watchdog and a stall/pivot mechanism so a stuck loop changes approach instead of looping forever on minor variants.
Check active goals against rubric
Verify each goal in active.md has evidence attached, stopping after 25 turns or completion.