Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Human brain is two separate organs, Stanford Medicine-led research finds(stanford.edu ↗)
    28comments
  2. San Francisco Onion Futures Company(onionfutures.com ↗)
    53comments
  3. If math is more than proof, we need to better celebrate the rest of it(terrytao.wordpress.com ↗)
    9comments
  4. Android 17 is the first since 3.x to add new APIs without releasing to the AOSP(grapheneos.social ↗)
    369comments
  5. GPT-6 Astra Solves a WWI German Radio Cipher(prinzai.com ↗)
    discuss
  6. Typesafe-computer-use drives a Mac toward a goal for 1/50th of a cent per step(github.com/awlevin ↗)
    22comments
  7. SDCC – Small Device C Compiler(sourceforge.net ↗)
    15comments
  8. Science Is Open Software(jepedersen.dk ↗)
    26comments
  9. Cloudflare Quick Tunnels(cloudflare.com ↗)
    275comments
  10. Saving another 100TB of RAM(cloudflare.com ↗)
    64comments
  11. Why building a Rust LSP is hard(rust-glancer.github.io ↗)
    24comments
  12. NASA-IBM Lunar Foundation open-Source Geospatial AI Model(usra.edu ↗)
    1comments
  13. How to Write with an LLM(sockpuppet.org ↗)
    316comments
  14. How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip(ieee.org ↗)
    77comments
  15. You can run Git on object storage if you re-make packfiles(tigrisdata.com ↗)
    6comments
  16. Ctenophores: Wonders of Biology(quantamagazine.org ↗)
    4comments
  17. Goroutine Leak Profiles(go.dev ↗)
    2comments
  18. The first new cat species discovered in 100 years(nationalgeographic.com ↗)
    90comments
  19. Veronese's Dogs(publicdomainreview.org ↗)
    discuss
  20. OpenJev(openjev.com ↗)
    257comments
  21. Show HN: Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash(cactuscompute.com ↗)
    87comments
  22. Photon-Emission-Guided Laser Fault Injection Enables RP2350 Secure Debug(ledger.com ↗)
    66comments
  23. Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)(arxiv.org ↗)
    12comments
  24. Stepfun Step 5 Preview (LLM): On AA Pareto frontier(artificialanalysis.ai ↗)
    1comments
  25. The Farnese letter(simonklee.dk ↗)
    6comments
  26. Minimal Phone 2(minimalcompany.com ↗)
    212comments
  27. Xcode 27.1 Beta Release Notes(developer.apple.com ↗)
    96comments
  28. Cyclomatic Complexity in C#(ndepend.com ↗)
    18comments
  29. Inside ZCode: Silently uploading your Git history to the cloud(ferstar.org ↗)
    97comments
  30. Warez: The Infrastructure and Aesthetics of Piracy (2021)(archive.org ↗)
    49comments

Typesafe-computer-use drives a Mac toward a goal for 1/50th of a cent per step

62 pointsby 2d agogithub.com
22 comments
2h agoHN ↗

Super cool. Hope waitlist will move soon. I have a use case for it too.

are you the author? If so - what are your notes on using Jev in this scenario?

2h agoHN ↗

Did you do the follow up questions? I was invited within a few hours of joining today.

It's also now on openrouter and cloudflare

1h agoHN ↗

Curious about people’s experience here. I am working on a small model, verify by jev, and escalate to big model. Some cases, the small model is not a model but some regex.

cheap-confirm-escalate

Using jev as the confirm step.

2h agoHN ↗

How does it do on OSWorld-verified? Recently read that even Fable 5 is just at 85% .

1h agoHN ↗

This probably throws a spanner in the wheels there:

Every piece of reasoning the frontier model does for free has to be rebuilt here as deterministic state.

EDIT: Not to shit on this though. I totally believe that some smart mixture of LLM-reasoning + Jev-style + determinism is going to be pretty amazing.

1h agoHN ↗

I'd be interested to see if using DiffusionGemma-as-Jev helps as you can feed the image directly into the model and it'll make decisions based on the image embeddings.

1h agoHN ↗

I had a long talk with chat gpt about this today as well. I think its duable and prolly not too hard either, also you could do lotsa funky stuff with stitched frames of a video in one 4x4 grid for example and send that as one image for analysis. that way temporal understanding can be had for fractions of a second by jev... also because vlm works in pixel space you can get around the whole state machine issue as well, so many possibilities...

1h agoHN ↗

I trailed off a few lines into the README. No human ever edited any of this. « LLM detected, project rejected ».

1h agoHN ↗

I stopped reading at “The honest caveat”…

1h agoHN ↗

Every third post in HN has this same complaint comment. Everybody knows.

53m agoHN ↗

And every comment on HN calling out vibe-coded slop has this same complaint comment in turn. OP has an actual complaint, your complaint is just "stop complaining".

47m agoHN ↗

I appreciate it when people say something is slop. Saves me from wasting my time looking at it.

How many slop ideas have come to the front page, never to be heard from again because the execution isn’t actually any good? I’m guessing most of them.

37m agoHN ↗

Just because the ReadMe is slop, doesn’t mean the code is slop. People are starting to make apps for themselves now and open-sourcing them so they’re not putting much thought into the ReadMe or distribution.

This just means ReadMe’s are less important now. I just have my terminal agent dig into the code and tell me what features are there. If the app is actually useful.

Even before AI, there were so many projects with subpar ReadMes, no screenshots, etc. But once you use the software, you realize how good it is.

Source: I maintain a massive collection of open-source alternatives and quality of open-source alternatives have increased a lot

35m agoHN ↗

If you do not respect your project to write your readme yourself chances are i will not care though.

26m agoHN ↗

For many new coders, LLMs are so good at writing the code, asking it to write the ReadMe sounds like a good idea. Clearly it’s the first impression your project makes so handwriting it is important.

Being turned off by the project because of the ReadMe is your prerogative. I’m just suggesting you dig into the code sometimes, the ReadMe is not the be all, end all.

32m agoHN ↗

You’re right that a poor quality readme doesn’t mean a poor-quality product, but it seems more likely than not to me.

Slop is an instant tab close for me. If something’s good, it’ll come around again. I’ll catch it when there's some evidence that it’s worth my time.

33m agoHN ↗

It's better that everyone is loud about it then everyone giving up and being silently irritated. At least if people complain it's possible to read the room

17m agoHN ↗

I thought it was a fine informative readme, starts with the problem, outlines the core of the solutions, and some limitations. Everything I want to know in the first few paragraphs. No need to spend human time to improve it.

30m agoHN ↗

not sure how this is innovative they show the System-1 model can play Doom right in the announcement [1] :

Doom >We love how this doomo doomonstrates real-time intelligence and what can be doone with code + AI. The engineer behind it was worried about making 10 queries a second (which ends up costing ~$7/hour), but the rest of us agreed that was lower than expected! This is so fun we intend to not only release an in-depth walkthrough, but also host some events to hack on this.

[1] https://typesafe.ai/blog/introducing-system-one-models-and-j...

9m agoHN ↗

Well nothing about the Doom demo or this entire model is new new either, I don't think even Typesafe themselves are claiming anything novel, they say that they're focusing on practicality instead of chasing big numbers and AGI. Classifiers are older than generative models and are used everywhere. Fast classifiers are used in sampling layers of every big model and for automation in agentic game plugins for years, except they're usually small and finetuned for the task, not general-use.

I think many people wondered why non-generative models are so underused on a big scale, well here's a long overdue attempt to market that which evidently goes well with people being interested in this again. The field has been captured by the vibe coding hype and valuation-goes-up a bit. AI has a ton of low hanging fruits that are much more practical than using one tool for everything.

14m agoHN ↗

I did not understand what this is all about. Anyone with more brain than me can explain please?