Join the discussion
Write your take first — we'll ask for email only when you're ready to publish.
- Hacker News
- A demo video would be useful, so one can see what it does without having to run oneself.by croemer
- I did not understand what this is all about. Anyone with more brain than me can explain please?by Surac
- It seems like the more honest comparison would be to OCR the screen and send that as input to the LLM?by aruss
- How does it do on OSWorld-verified? Recently read that even Fable 5 is just at 85% .by Zaraif13
- Super cool. Hope waitlist will move soon. I have a use case for it too.
are you the author? If so - what are your notes on using Jev in this scenario?
by john_minsk - I'd be interested to see if using DiffusionGemma-as-Jev helps as you can feed the image directly into the model and it'll make decisions based on the image embeddings.by mmastrac
- not sure how this is innovative they show the System-1 model can play Doom right in the announcement [1] :
>Doom >We love how this doomo doomonstrates real-time intelligence and what can be doone with code + AI. The engineer behind it was worried about making 10 queries a second (which ends up costing ~$7/hour), but the rest of us agreed that was lower than expected! This is so fun we intend to not only release an in-depth walkthrough, but also host some events to hack on this.
[1] https://typesafe.ai/blog/introducing-system-one-models-and-j...
by jimmySixDOF - I trailed off a few lines into the README. No human ever edited any of this. « LLM detected, project rejected ».by jdkoeck