Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. You Are No Longer Invited to Dinner (derekthompson.org)
    126comments
  2. AI companies leak data to advertisers [pdf] (jorgegarciaherrero.com)
    52comments
  3. Jeeves. Reasoning improves Jev-like decision models (github.com/posthog)
    25comments
  4. 500k facial scans at UK stations yield no arrests, 1 false positive (theguardian.com)
    53comments
  5. Using any C++ library in Godot (conan.io)
    21comments
  6. Delhi Cut Electricity Loss from 50 to 5 Percent (ieee.org)
    —discuss
  7. US sanctions force The Netherlands off Microsoft and toward alternative NixOS (tomshardware.com)
    33comments
  8. Phyllotaxis: An audio-reactive LED display (jagi.studio)
    23comments
  9. Booted up in 1993, this server still runs – but not for much longer (2017) (computerworld.com)
    42comments
  10. The systems that no one will test (christianperone.com)
    36comments
  11. Startup Nights 2026 is comming up on 5-6 Nov. in Switzerland (startup-nights.ch)
    11comments
  12. Show HN: Raven – The harness of harnesses, built for RSI (github.com/evermind-ai)
    5comments
  13. California farmers are struggling to sell grapes as demand for wine drops (kqed.org)
    677comments
  14. Software occlusion culling in Block Game (enikofox.com)
    4comments
  15. Pirating the Pirates (mubi.com)
    316comments
  16. Four CHI '26 papers I wish I wrote (countingfromzero.blog)
    3comments
  17. Tank Body Problem (jimsitu.com)
    30comments
  18. The Beatles have permeated research papers across academic disciplines (phys.org)
    —discuss
  19. MicroLLM Lab – Try 7 tiny LLM's in the browser (stateofutopia.com)
    90comments
  20. ESP32S3 cluster running 1.58-bit (BitNet) Language model (github.com/low-zi-hong)
    25comments
  21. 12,000-year-old Göbeklitepe burials explain scattered bones (archaeologymag.com)
    42comments
  22. Optimizing x264 settings and per-title ladders (streaminglearningcenter.com)
    —discuss
  23. Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms (github.com/firelex)
    199comments
  24. Evan Doorbell's Phone Tapes – Brought to You by Telephone World (evan-doorbell.com)
    1comments
  25. New Cyber-OSINT model released (twitter.com/0x0sojalsec)
    6comments
  26. Updated Google Maps shows destruction of the city of Rafah (twitter.com/aliabunimah)
    578comments
  27. Nvidia wants to put a watchdog chip next to every AI agent (cnbc.com)
    250comments
  28. Hijacking the PS5's RTMP stream (yashgarg.dev)
    88comments
  29. Sonnet 5.5 (anthropic.com)
    561comments
  30. Kids turned low-traffic NPR Spotify comments into a secret group chat (thisamericanlife.org)
    235comments

Jeeves. Reasoning improves Jev-like decision models

62 pointsby 1h agogithub.com
24 comments
1h agoHN ↗

Just curious, where has this term 'noul' come from for yes/no ansers?

/a bit more digging and..

A Noul performs a Bernoulli trial—an experiment with exactly two outcomes (yes or no)—but instead of picking one, it returns the calibrated probability (ranging from 0.0 to 1.0) that the statement is true.

I hate it :)

1h agoHN ↗

If you want to get super pedantic about what’s happening in a transistor every digital Boolean is actually this

56m agoHN ↗

Not quite.. that boolean is about whether the voltage exceeds some threshold. It's not about how close the voltage is to the circuit's maximum possible threshold, or how much it exceeds the threshold.

In an analog circuit, maybe.

28m agoHN ↗

The whole "no hallucinations" premise is based on that.

Like, yeah, you don't hallucinate, but only because you force the user to decide in the end.

16m agoHN ↗

Wait til you hear about how digital circuits work at die level

14m agoHN ↗

That seems like the only possible way to eliminate hallucination, short of a model that is never wrong.

11m agoHN ↗

force the user to decide in the end

And that's...bad?

18m agoHN ↗

I like it. It is short and distinct which is a good fit for a primitive. It describes its fundamental meaning and draws a connotation with Boolean.

55m agoHN ↗

Are there "good" Open source Decision models built on Gemma-4 and also trainiable on own data?

50m agoHN ↗

Jeeves, that's a name I haven't heard in a long time...

44m agoHN ↗

If Jeeves returned as an AI chat bot it would be the most brilliant resurgence of nostalgia

35m agoHN ↗

jeeves is currently the name of my local hosted assistant, in its context there are rules that tell it to behave like good old jeeves.

soon I'll make sure that my home assistant pod answers to "Hey jeeves"

30m agoHN ↗

Personally I’m happy that after a 30 year effort and hundreds of billions spent, AskJeeves finally works as intended.

13m agoHN ↗

We locked him in the basement with Clippy, Bob and BonziBuddy. Who opened the damn basement door???

36m agoHN ↗

This isn't really surprising. LLM reasoning and before that, chain of thought prompting are essentially forms of test-time compute scaling.

32m agoHN ↗

Seems like this is the way, a hybrid approach where some of the pipeline will be jev like and some traditional LLM depending on the nature of the work.

14m agoHN ↗

Ask Jeeves - Only took us 30 years to come full circle.

14m agoHN ↗

Super dope. If it would ship as prod ready code supporting mps as well that would be even doper.

But funny that jev is getting its lunch eaten apparently in under two weeks?

13m agoHN ↗

interesting bench list, what about benchmark against smaller or bigger models? 9B looks too huge for small like laya, and too small for llm-level decisions.

11m agoHN ↗

See if it can beat Jev's Pokémon benchmark

8m agoHN ↗

what about benchmark against smaller or bigger models? 9B looks too small for llm-level decisions.