Solving the game with reasoning, not reinforcement learning. An interactive walkthrough of AI agents running a research loop to beat my game.
1 comment
dustinlakin10 days ago
One of the most interesting parts was having this benchmark setup available over the last couple of months. It has made each new release something fun I can spend my night tokens on. The newest models really show the consistency I hadn't seen until Fable 5.1 and Astra 6.
Around to answer any questions or hear any feedback on the visualizations.
Read the full thread on Hacker News →
Related stories
- The Nine-Person Game Team Will Soon Outperform a AAA Game Studioclairvoyanceai.comHacker News · 2 points · 4 days ago
- Hacker News · 1 points · 1 day ago
- Show HN: Built an online multiplayer game in 2 daysbigbeanbattle.comHacker News · 1 points · 1 day ago
- Hacker News · 1 points · 2 days ago
- Show HN: Trade Lord, inspired by Drug Lord 2, quick daily leaderboard gametradelord.theboyvr.comHacker News · 1 points · 2 days ago
- Hacker News · 1 points · 2 days ago