Typesafe-computer-use drives a Mac toward a goal for 1/50th of a cent per step
73 points - last Wednesday at 10:01 PM
SourceComments
jdkoeck today at 5:55 AM
I trailed off a few lines into the README. No human ever edited any of this. « LLM detected, project rejected ».
jimmySixDOF today at 6:58 AM
not sure how this is innovative they show the System-1 model can play Doom right in the announcement [1] :
>Doom >We love how this doomo doomonstrates real-time intelligence and what can be doone with code + AI. The engineer behind it was worried about making 10 queries a second (which ends up costing ~$7/hour), but the rest of us agreed that was lower than expected! This is so fun we intend to not only release an in-depth walkthrough, but also host some events to hack on this.
[1] https://typesafe.ai/blog/introducing-system-one-models-and-j...
mmastrac today at 5:35 AM
I'd be interested to see if using DiffusionGemma-as-Jev helps as you can feed the image directly into the model and it'll make decisions based on the image embeddings.
Surac today at 7:13 AM
I did not understand what this is all about. Anyone with more brain than me can explain please?
Zaraif13 today at 5:13 AM
How does it do on OSWorld-verified? Recently read that even Fable 5 is just at 85% .
aruss today at 7:33 AM
It seems like the more honest comparison would be to OCR the screen and send that as input to the LLM?
john_minsk today at 4:50 AM
Super cool. Hope waitlist will move soon. I have a use case for it too.
are you the author? If so - what are your notes on using Jev in this scenario?
curtisblaine today at 7:51 AM
"The honest caveat"
pulvinar today at 5:41 AM
[dead]