Typesafe-computer-use drives a Mac toward a goal for 1/50th of a cent per step
37 points by rahimnathwani 3 days ago | 9 comments
jdkoeck 7 minutes ago
I trailed off a few lines into the README. No human ever edited any of this. « LLM detected, project rejected ».
replyalex_suzuki 6 minutes ago
I stopped reading at “The honest caveat”…
replymmastrac 27 minutes ago
I'd be interested to see if using DiffusionGemma-as-Jev helps as you can feed the image directly into the model and it'll make decisions based on the image embeddings.
replynowittyusername 16 minutes ago
I had a long talk with chat gpt about this today as well. I think its duable and prolly not too hard either, also you could do lotsa funky stuff with stitched frames of a video in one 4x4 grid for example and send that as one image for analysis. that way temporal understanding can be had for fractions of a second by jev... also because vlm works in pixel space you can get around the whole state machine issue as well, so many possibilities...
replyZaraif13 49 minutes ago
How does it do on OSWorld-verified? Recently read that even Fable 5 is just at 85% .
replyjohn_minsk 2 hours ago
Super cool. Hope waitlist will move soon. I have a use case for it too.
replyare you the author? If so - what are your notes on using Jev in this scenario?
jasonjmcghee 2 hours ago
Did you do the follow up questions? I was invited within a few hours of joining today.
replyIt's also now on openrouter and cloudflare