Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Recalld – A memory layer for AI agents that returns only relevant facts (recalld.ai)
    1comments
  2. Claude Sonnet 5.5 is out! (claude.dev)
    —discuss
  3. Summary Advice (Churchill, 1940) (futilitycloset.com)
    —discuss
  4. Walmart promises to never use AI to charge shoppers higher prices (walmart.com)
    —discuss
  5. Oil Rises as Strait of Hormuz Talks Stall (coinmarketcap.com)
    1comments
  6. Bioenergetic Landscapes (worldsensorium.com)
    —discuss
  7. Blade Runner (blue-continuum.com)
    —discuss
  8. The AI as the New Compiler (buildwithdc.co)
    —discuss
  9. Claude and Jev Play Pokémon (singlepageapp.co)
    —discuss
  10. Show HN: What a VM for your AI agent costs across 13 sandbox providers (sf.tools)
    —discuss
  11. Show HN: Serverless, cheater-proof MTG engine using zero-knowledge play (chiplis.com)
    —discuss
  12. ESP32S3 cluster running 1.58-bit (BitNet) Language model (github.com/low-zi-hong)
    —discuss
  13. DrawDB: Online database diagram editor and SQL generator (github.com/drawdb-io)
    —discuss
  14. OpenAI, Hugging Face Attack – Music Video (reddit.com)
    1comments
  15. Women Behind North Korea's IT Worker Operations (haydenmckenzie.com)
    —discuss
  16. Remembering the Chicago Pile, the First Nuclear Reactor (newyorker.com)
    —discuss
  17. Spec Driven Development in 14mins (productmindset.substack.com)
    2comments
  18. Cap Snatching (wikipedia.org)
    —discuss
  19. Deutsche Bahn "joke" is no longer funny (jonworth.eu)
    —discuss
  20. End of Hand-Written Code: A Conversation with DHH (thoughteconomics.com)
    3comments
  21. Stanford CME295 Transformers n LLMs – Autumn 2026 – Lecture 1 – Transformers [video] (youtube.com)
    —discuss
  22. Starship Achieved Orbit (twitter.com/spacex)
    5comments
  23. Machine-Generated, Machine-Checked Proofs for a Verified Compiler (ICFP'26) [video] (youtube.com)
    —discuss
  24. Authors Guild calls on publishers to share Anthropic wealth (publishersweekly.com)
    —discuss
  25. The Internet Is Dead (code-bude.net)
    —discuss
  26. Conway's Law cuts both ways (contrapeso.xyz)
    —discuss
  27. Florida asks for order to halt ChatGPT development (axios.com)
    1comments
  28. Researchers proved AI has deleted every reason universities exist (twitter.com/thesupermannx)
    1comments
  29. Full Self Hauling: Volvo moves 3 million tonnes of Earth, autonomously (electrek.co)
    —discuss
  30. Dutch 'reformed hacker' arrested in ShinyHunters investigation, police say (reuters.com)
    —discuss

MicroLLM Lab – Try 7 tiny LLM's in the browser

83 pointsby 2h agostateofutopia.com
29 comments
2h agoHN ↗

You can try 7 different tiny LLM's in your browser.

1h agoHN ↗

What I missed from the title was this: Can I try 7 different tiny LLMs in my browser?

1h agoHN ↗

i have a bromine tub, give me simple instructions for what to do since i just filled it up w fresh water

1. Add 1 tablespoon of water to the water bath. 2. Place the tub into the water bath and let it sit for about 5 minutes. 3. After 5 minutes, remove the tub and let it cool down. 4. Now, fill the tub with water and let it sit for about

uhmm completely unusable ?

1h agoHN ↗

No, that is not what SLMs are useful for. They have very little knowledge, and typically lack reasoning skills.

Useful applications include sentiment analysis, text classification, entity extraction, etc.

They certainly can be useful, but you shouldn't compare them with LLMs such as Opus or Fable.

1h agoHN ↗

Completely unusable for that, yeah, that's a 100M model, not an assistant

1h agoHN ↗

The website literally says they are SLM's .. derp af

1h agoHN ↗

I did the default arithmetic with PetitGPT research-v1

What is 2+2?

Answer

To find 2 + 2, we need to add 2 to both sides of the equation.

2 + 2 = 4

So, 2 + 2 = 4 + 2.

Brilliant

29m agoHN ↗

That's not incorrect.

LLMs produce semantically correct sentences, not factually correct statements. Have we forgotten this so soon?

22m agoHN ↗

It is, in every sense, incorrect. Which statement in this short snippet is "semantically correct"? (Better LLMs get this right, of course.)

13m agoHN ↗

GPT-2 124M is terrible...

  > what is 2+2?

Answer:

  > 3+3? 4+4? 5+6? 7+8?
  > Reply ~18000 0 ~10 min 2 By : 1-1: I'm a beginner. 3x2 is my best option, but if you're not sure about the other options then just go for it and try again
10m agoHN ↗

It's not instruct tuned looks like? It's closer to a base model rather than a chatbot.

1h agoHN ↗

Cool project, but I'd really suggest looking at the UI.

The text is too small and it's way too dense with information in general. Considering how simple this product is to use, it's kinda crazy that I have to scroll through over a page length of (mostly useless, AI-generated) information before getting to the actual interface.

Also what is going on with the footer (why does it link back to the site itself, why is it telling me to "serve over HTTP").

42m agoHN ↗

This is the future of software, sloppy ui.

25m agoHN ↗

Thank you for the feedback, I'll think about how to incorporate the changes you've suggested.

1h agoHN ↗

PetitGPT told me that

"2+2 is 2."

Otherwise, a very neat demo. As others have said, the UI is VERY confusing, way too much stuff going on.

1h agoHN ↗

Really cool project. Giving web apps direct access to on-device models is something I’m excited about, and it’s cool to see the different approaches.

I’ve been working on a related proposal called the Web Models API, which explores a browser standard for an API that runs open-weight models on-device. Would love your thoughts: https://www.webmodels.dev

52m agoHN ↗

I read your proposal, I think it's great! Where will the navigator get the model if the user agrees to download it? For this demonstration I just serve the models on my own server, but for larger models it may be an issue as they may not have direct download links even if they are open weights.

46m agoHN ↗

Great question. Right now there isn’t a definitive answer, but it’s something that needs to be worked out. There would likely be a registry. The question is how to keep model IDs consistent: does each browser manage its own registry, or is there one shared across browsers?

6m agoHN ↗

Since you're asking for some kinds of permissions anyway, you could ask if the user is willing to also seed the model, p2p. (However, seeding files is not as popular as it used to be, many residential Internet connections don't have good upload.) If you have the capacity for it, your site webmodels.dev could act as a tracker and initial seed for any models. Then it could be the one central registry. It might get to be too much for you though, a lot of the open weights models are huge.

3m agoHN ↗

Interesting idea. I hadn't considered that. Thanks

59m agoHN ↗

unfortunately on firefox: Uncaught ReferenceError: GPUShaderStage is not defined