Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • I wish jev took in images so we could do this generically for any game, without memhacks. I'm sure that's coming.

    You could front this with an image -> text model but that would be much lower quality vs latency, and the whole point of doing it with a decision model is remove the latency.

    Games are a really interesting testing ground for robotics; if we can solve game playing (incl 3d) we could embody "system one" intelligence into robots that have something emulating general reflexes without needing to fine tune.

  • I'm watching this. After 6 hours it managed to solve the rock boulder puzzle and now it's trying to break the Elite Four wall with its face. I'm no pokemon expert, but I'd say that no way in hell it beats them with this team composition. It has to take a step back, re-compose or at least re-train for overall higher levels. Will see if it manages to do it, my expectations are low.
  • The seminal (lol) Twitch Plays Pokémon was twelve years ago, so just posting this amazing moment of internet history/lore just in case folks don’t know or have forgotten: https://en.wikipedia.org/wiki/Twitch_Plays_Pok%C3%A9mon
  • with such a fat harness, this is more like watching a walk thru play the game.
  • Considering it just made Charizard forget its only fire-type move "Ember" to learn "Counter", I note no signs of intelligence.
  • This is kinda chill to have in the background. I wish there were livestreams showing live reasoning of top models which are currently trying to solve cancer or whatever. Imagine the pogs in chat when it does.
  • Cool project, comes with a little too much guidance in the harness though IMO (pathfinding, textual milestones etc). (The author is very upfront about this in their README though)

    I think if it was combined with a regular vLLM it could be really interesting, especially watching the reasoning logs.

    Bonus points if it was one of the latest open models that somehow had all prior training knowledge of Pokemon abliterated so it was reasoning as an intelligent persona that had no knowledge of even the concept of Pokemon.

    by ac2u
  • This is so interesting to watch. For a couple minutes I was in awe of how quick and cheap it was. Then I saw just how bad the decision are and how it would get stuck in strange loops of going in and out of the same door to no end.

    This seems like a technology heading in the right direction but not quiet there yet. Excited for what they are cooking up but probably won't start building around it yet.

Explore Birbla archives