I trained a small JEPA-style world model on Pokémon Red.

24 points•stmonty•5 days ago•18 comments•

18 comments

dang5 days ago
Dare we have two AI-plays-Pokemon threads at the same time?

Show HN: Jev Plays Pokémon Red - https://news.ycombinator.com/item?id=49845172

stmonty5 days ago
I think seeing that thread on the front page gave me the courage to post my own post!
brudgers4 days ago
YOLO.
stmonty5 days ago
Author here, feel free to ask me any questions. Though I am no expert in this, just someone trying to learn and have some fun doing it.
ainch5 days ago
Curious on whether you think this JEPA-style approach could scale to a full game completions?

For reference, standard PPO has been able to beat the game end-to-end with a relatively small network https://drubinstein.github.io/pokerl/

stmonty5 days ago
I think it is possible but seems quite difficult to me from this experiment. The hard part is that Pokemon has a lot of different goals, and has many strategies to complete it. You would need a way to encode these intermediary goals in some way temporally that the model needs to learn to complete in order. For example, you need to collect all eight gym badges before you can challenge the Elite Four to win.

You could compose multiple network, one with the goal of learning how to interact with the world through the screenshots. This would probably require a lot of frames of the game from many diverse locations and scenarios--as the network would need to learn how battles works, items (though technically not strictly necessary), and general dialogue and menu interactions.

Then other networks could then try to learn to encode different intermediary goals trained on a bunch of noisy runs that complete that goal. Collecting this data seems tedious and difficult though and is a whole project in itself.

So my intuition is yes, but I don't think simply making my current model bigger would get us there. I see you have experience with world models--what would you try? ;)

singularity20015 days ago
World models don't need to be taught again. All the current top AIs are capable of solving Pokémon just from what they've read on the internet or their general knowledge.
password543214 days ago
At the end of rabbit hole you meet The Bitter Lesson like everyone else.
hobo1234 days ago
We really need AI to play Final Fantasy Legend *.

Read the full thread on Hacker News →

Related stories