I trained a small JEPA-style world model on Pokémon Red.
18 comments
Show HN: Jev Plays Pokémon Red - https://news.ycombinator.com/item?id=49845172
For reference, standard PPO has been able to beat the game end-to-end with a relatively small network https://drubinstein.github.io/pokerl/
You could compose multiple network, one with the goal of learning how to interact with the world through the screenshots. This would probably require a lot of frames of the game from many diverse locations and scenarios--as the network would need to learn how battles works, items (though technically not strictly necessary), and general dialogue and menu interactions.
Then other networks could then try to learn to encode different intermediary goals trained on a bunch of noisy runs that complete that goal. Collecting this data seems tedious and difficult though and is a whole project in itself.
So my intuition is yes, but I don't think simply making my current model bigger would get us there. I see you have experience with world models--what would you try? ;)
Read the full thread on Hacker News →
Related stories
- Show HN: Divot.css – a simple skeuomorphic button stylenlaz.github.ioHacker News · 2 points · 2 days ago
- Hacker News · 1 points · 6 days ago
- Hacker News · 2 points · 10 days ago
- Hacker News · 1 points · 10 days ago
- Hacker News · 115 points · 11 days ago
- Hacker News · 4 points · 11 days ago