Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Escalate hard decisions with the advisor tool (claude.com)
    —discuss
  2. NYT head of engineering and gaming experience allegedly shot by elderly in-laws (nbcnews.com)
    —discuss
  3. Made a thing, if you use Grok and Pi check it out. Or scoll on bros (github.com/jangman-j)
    1comments
  4. HouseOS – self-hosted homesystem platform, designed to be operated by agents (github.com/boitech-dev)
    1comments
  5. Found in translation: the selfless legacy of Eien Ni Hen (magnvsrpgjourney.substack.com)
    —discuss
  6. Range – open a 1 TB AI model in 3 seconds without downloading it (getrange.sh)
    —discuss
  7. How I Made a Zero Dependency Terminal Text Editor in C Using POSIX System Calls (111nation.github.io)
    —discuss
  8. Openprinter (crowdsupply.com)
    —discuss
  9. Recalld – A memory layer for AI agents that returns only relevant facts (recalld.ai)
    1comments
  10. When to choose Sonnet over Opus, what it costs, and how to tune it (claude.dev)
    —discuss
  11. Summary Advice (Churchill, 1940) (futilitycloset.com)
    —discuss
  12. Walmart promises to never use AI to charge shoppers higher prices (walmart.com)
    1comments
  13. Oil Rises as Strait of Hormuz Talks Stall (coinmarketcap.com)
    1comments
  14. Bioenergetic Landscapes (worldsensorium.com)
    —discuss
  15. Blade Runner (blue-continuum.com)
    —discuss
  16. The AI as the New Compiler (buildwithdc.co)
    —discuss
  17. Claude and Jev Play Pokémon (singlepageapp.co)
    —discuss
  18. Show HN: What a VM for your AI agent costs across 13 sandbox providers (sf.tools)
    —discuss
  19. Show HN: Serverless, cheater-proof MTG engine using zero-knowledge play (chiplis.com)
    —discuss
  20. ESP32S3 cluster running 1.58-bit (BitNet) Language model (github.com/low-zi-hong)
    —discuss
  21. DrawDB: Online database diagram editor and SQL generator (github.com/drawdb-io)
    —discuss
  22. OpenAI, Hugging Face Attack – Music Video (reddit.com)
    1comments
  23. Women Behind North Korea's IT Worker Operations (haydenmckenzie.com)
    —discuss
  24. Remembering the Chicago Pile, the First Nuclear Reactor (newyorker.com)
    —discuss
  25. Spec Driven Development in 14mins (productmindset.substack.com)
    3comments
  26. Cap Snatching (wikipedia.org)
    —discuss
  27. Deutsche Bahn "joke" is no longer funny (jonworth.eu)
    —discuss
  28. Stanford CME295 Transformers n LLMs – Autumn 2026 – Lecture 1 – Transformers [video] (youtube.com)
    —discuss
  29. Starship Achieved Orbit (twitter.com/spacex)
    5comments
  30. Machine-Generated, Machine-Checked Proofs for a Verified Compiler (ICFP'26) [video] (youtube.com)
    —discuss

MicroLLM Lab – Try 7 tiny LLM's in the browser

88 pointsby 3h agostateofutopia.com
32 comments
3h agoHN ↗

You can try 7 different tiny LLM's in your browser.

1h agoHN ↗

What I missed from the title was this: Can I try 7 different tiny LLMs in my browser?

2h agoHN ↗

i have a bromine tub, give me simple instructions for what to do since i just filled it up w fresh water

1. Add 1 tablespoon of water to the water bath. 2. Place the tub into the water bath and let it sit for about 5 minutes. 3. After 5 minutes, remove the tub and let it cool down. 4. Now, fill the tub with water and let it sit for about

uhmm completely unusable ?

1h agoHN ↗

No, that is not what SLMs are useful for. They have very little knowledge, and typically lack reasoning skills.

Useful applications include sentiment analysis, text classification, entity extraction, etc.

They certainly can be useful, but you shouldn't compare them with LLMs such as Opus or Fable.

1h agoHN ↗

Completely unusable for that, yeah, that's a 100M model, not an assistant

2h agoHN ↗

The website literally says they are SLM's .. derp af

2h agoHN ↗

I did the default arithmetic with PetitGPT research-v1

What is 2+2?

Answer

To find 2 + 2, we need to add 2 to both sides of the equation.

2 + 2 = 4

So, 2 + 2 = 4 + 2.

Brilliant

45m agoHN ↗

That's not incorrect.

LLMs produce semantically correct sentences, not factually correct statements. Have we forgotten this so soon?

37m agoHN ↗

It is, in every sense, incorrect. Which statement in this short snippet is "semantically correct"? (Better LLMs get this right, of course.)

29m agoHN ↗

GPT-2 124M is terrible...

  > what is 2+2?

Answer:

  > 3+3? 4+4? 5+6? 7+8?
  > Reply ~18000 0 ~10 min 2 By : 1-1: I'm a beginner. 3x2 is my best option, but if you're not sure about the other options then just go for it and try again
26m agoHN ↗

It's not instruct tuned looks like? It's closer to a base model rather than a chatbot.

1h agoHN ↗

Cool project, but I'd really suggest looking at the UI.

The text is too small and it's way too dense with information in general. Considering how simple this product is to use, it's kinda crazy that I have to scroll through over a page length of (mostly useless, AI-generated) information before getting to the actual interface.

Also what is going on with the footer (why does it link back to the site itself, why is it telling me to "serve over HTTP").

58m agoHN ↗

This is the future of software, sloppy ui.

40m agoHN ↗

Thank you for the feedback, I'll think about how to incorporate the changes you've suggested.

1h agoHN ↗

PetitGPT told me that

"2+2 is 2."

Otherwise, a very neat demo. As others have said, the UI is VERY confusing, way too much stuff going on.

1h agoHN ↗

Really cool project. Giving web apps direct access to on-device models is something I’m excited about, and it’s cool to see the different approaches.

I’ve been working on a related proposal called the Web Models API, which explores a browser standard for an API that runs open-weight models on-device. Would love your thoughts: https://www.webmodels.dev

1h agoHN ↗

I read your proposal, I think it's great! Where will the navigator get the model if the user agrees to download it? For this demonstration I just serve the models on my own server, but for larger models it may be an issue as they may not have direct download links even if they are open weights.

1h agoHN ↗

Great question. Right now there isn’t a definitive answer, but it’s something that needs to be worked out. There would likely be a registry. The question is how to keep model IDs consistent: does each browser manage its own registry, or is there one shared across browsers?

21m agoHN ↗

Since you're asking for some kinds of permissions anyway, you could ask if the user is willing to also seed the model, p2p. (However, seeding files is not as popular as it used to be, many residential Internet connections don't have good upload.) If you have the capacity for it, your site webmodels.dev could act as a tracker and initial seed for any models. Then it could be the one central registry. It might get to be too much for you though, a lot of the open weights models are huge.

18m agoHN ↗

Interesting idea. I hadn't considered that. Thanks

1h agoHN ↗

unfortunately on firefox: Uncaught ReferenceError: GPUShaderStage is not defined

10m agoHN ↗

I tested it on Firefox on windows, version 156.0.1 and didn't get that error.

What version of Firefox are you using and what is your operating system and graphics card, please? Can you also try it without WebGPU? (Reload the page and uncheck "Prefer WebGPU" and try a prompt.)

15m agoHN ↗

What browser are you using? I tested it on Windows, Mac, and iPhone. I tested it in Chrome, Firefox, Edge, and Safari. Everything works on the three machines and phone I tested it on.

(It's a little bit slow at the moment - you have to wait a few seconds for the models to load - as it's currently on the HN front page. The server is on a 1 gigabit unmetered network connection so it can serve all the weights - around 600 megabytes - to one person every few seconds, there are several concurrent users now.)