I trained a 1.57B-parameter Dreamer 4 World Model from scratch for under $150
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| My first attempt didn't work. I built on Genie's architecture and the videos looked great, but the controls barely did anything. The effect of a keypress was basically zero. Genie learns its actions unsupervised into 8 codes, and that was too loose a grip for us. So I scrapped it and started again with Dreamer 4. The second attempt: Tokenizer at 40.41 PSNR (Genie's paper reports 35.7) Two important learnings: (1) One is that $150 is enough. You don't need a frontier lab to do this anymore, and I don't think enough people have noticed. (2) The other is the data. We generated every frame ourselves with Procgen instead of scraping video. We know the true action at every step, so we can actually check whether the model is responding to us or just making pretty motion. Website: https://worldmodel-platformer.vizuara.ai/ [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.