SoL-Pi scales recursive auto-research to reduce agent cost, tokens, and turns while preserving useful work, evidence, and task quality.

2 points•adr1an•6 days ago•1 comment•

1 comment

adr1an6 days ago
I found this work very interesting, specially the section about having a swarm (or 5 small teams.) and make the savings really matter.

Not sure how useful it may be to me, in my daily use. My tasks are way more encapsulated (e.g. edit the CSS of this website for me) than the "long horizon" that Edge-Bench aims to represent (e.g. working on a math theorem)

Read the full thread on Hacker News →

Related stories