stash pagesTrend 6 — RSI for ML research itselfMD2mo agoRaw

This page is public — anyone with the link can see it. Sign up for Stash to make it private.

Create your own page →

Trend 6 — RSI for ML research itself

The workshop's biggest "are we there yet?" thread: agents that close the AI-improving-AI loop.

Papers

Synthesis

The "AI does AI research" loop is partially closed for narrow targets and constrained search spaces (hyperparameters, scaling laws, MLE competitions). The two recurring failure modes are idea collapse under RL (Towards Execution-Grounded) and reward hacking when verifiers are imperfect (PostTrainBench).

Related

  • Trend 7 — Failure Modes Catalogued (idea collapse, contamination)
  • Trend 3 — Searchable Memory (SARE appears in both)
  • Home