improvement

12 stories and discussions about improvement, aggregated from every source we track.

2.

Contribute to google-research/rrsi development by creating an account on GitHub.

3 points•simonpure•2 days ago•0 comments•
3.

Regularize the search, not the harness. Harness evolution that transfers out of distribution on a third fewer policy tokens.

2 points•jonbaer•about 23 hours ago•0 comments•
4.

On-policy distillation (OPD) trains a student model by having it generate trajectories, then matching its next-token predictions with an external teacher's next-token predictions. This provides dense, token-level…

2 points•simonpure•2 days ago•0 comments•
5.

AI agents are beginning to automate research and development across the AI stack, from improving training efficiency to optimizing inference. A natural next step is to improve the research efficiency of the agents…

2 points•handfuloflight•7 days ago•1 comment•
6.

AI self-improvement may not be powerful enough to overcome diminishing returns. Even RSI may not lead to a fast takeoff to ASI (superintelligence).

1 points•thevises•2 days ago•0 comments•
7.
1 points•taivare•3 days ago•0 comments•
8.

AI agents are beginning to automate research and development across the AI stack, from improving training efficiency to optimizing inference. A natural next step is to improve the research efficiency of the agents…

1 points•sbulaev•8 days ago•0 comments•
9.
1 points•paulpauper•9 days ago•0 comments•
10.

An LLM agent's capability is largely magnified by its harness, namely the prompts, control flow, tooling, memory, and context management surrounding the frozen backbone model. Recent methods increasingly automate this…

1 points•Betelbuddy•9 days ago•0 comments•
11.
1 points•gps372•9 days ago•0 comments•
12.

“We never want to be in a situation again where we underestimate the AI.”

1 points•gmays•10 days ago•0 comments•

Related topics