Hi everyone!
After spending hundreds of hours, we're excited to finally share our progress in developing a reinforcement learning system to beat Pokémon Red. Our system successfully completes the game using a policy under 10M parameters, PPO, and a few novel techniques. With the release of Claude Plays Pokémon, now feels like the perfect time to showcase our work.
We'd love to get feedback!
Comments URL: https://news.ycombinator.com/item?id=43269330
Points: 41
# Comments: 26
Zaloguj się, aby dodać komentarz
Inne posty w tej grupie
Article URL: https://packagephobia.com/
Comments URL: https://news.ycombinator.com/item?id=4

Article URL: https://platform.openai.com/docs/models/o1-pro
Article URL: https://canopylabs.ai/model-releases
Comments URL: https://news.ycomb
Article URL: https://szymanowiczs.github.io/bolt3d
Comments URL: https://news.yco
