Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. You Are No Longer Invited to Dinner (derekthompson.org)
    124comments
  2. AI companies leak data to advertisers [pdf] (jorgegarciaherrero.com)
    52comments
  3. Jeeves. Reasoning improves Jev-like decision models (github.com/posthog)
    24comments
  4. Using any C++ library in Godot (conan.io)
    21comments
  5. 500k facial scans at UK stations yield no arrests, 1 false positive (theguardian.com)
    53comments
  6. US sanctions force The Netherlands off Microsoft and toward alternative NixOS (tomshardware.com)
    32comments
  7. Delhi Cut Electricity Loss from 50 to 5 Percent (ieee.org)
    —discuss
  8. Phyllotaxis: An audio-reactive LED display (jagi.studio)
    22comments
  9. Booted up in 1993, this server still runs – but not for much longer (2017) (computerworld.com)
    41comments
  10. The systems that no one will test (christianperone.com)
    36comments
  11. Show HN: Raven – The harness of harnesses, built for RSI (github.com/evermind-ai)
    5comments
  12. Startup Nights 2026 is comming up on 5-6 Nov. in Switzerland (startup-nights.ch)
    11comments
  13. California farmers are struggling to sell grapes as demand for wine drops (kqed.org)
    676comments
  14. Software occlusion culling in Block Game (enikofox.com)
    4comments
  15. Pirating the Pirates (mubi.com)
    316comments
  16. Four CHI '26 papers I wish I wrote (countingfromzero.blog)
    3comments
  17. Tank Body Problem (jimsitu.com)
    30comments
  18. MicroLLM Lab – Try 7 tiny LLM's in the browser (stateofutopia.com)
    90comments
  19. ESP32S3 cluster running 1.58-bit (BitNet) Language model (github.com/low-zi-hong)
    25comments
  20. 12,000-year-old Göbeklitepe burials explain scattered bones (archaeologymag.com)
    42comments
  21. Optimizing x264 settings and per-title ladders (streaminglearningcenter.com)
    —discuss
  22. Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms (github.com/firelex)
    199comments
  23. Evan Doorbell's Phone Tapes – Brought to You by Telephone World (evan-doorbell.com)
    1comments
  24. New Cyber-OSINT model released (twitter.com/0x0sojalsec)
    6comments
  25. Updated Google Maps shows destruction of the city of Rafah (twitter.com/aliabunimah)
    574comments
  26. Nvidia wants to put a watchdog chip next to every AI agent (cnbc.com)
    250comments
  27. Hijacking the PS5's RTMP stream (yashgarg.dev)
    87comments
  28. Sonnet 5.5 (anthropic.com)
    561comments
  29. Kids turned low-traffic NPR Spotify comments into a secret group chat (thisamericanlife.org)
    235comments
  30. Scientists solve 1840s space weather mystery (arstechnica.com)
    48comments

Jeeves. Reasoning improves Jev-like decision models

62 pointsby 1h agogithub.com
24 comments
1h agoHN ↗

Just curious, where has this term 'noul' come from for yes/no ansers?

/a bit more digging and..

A Noul performs a Bernoulli trial—an experiment with exactly two outcomes (yes or no)—but instead of picking one, it returns the calibrated probability (ranging from 0.0 to 1.0) that the statement is true.

I hate it :)

1h agoHN ↗

If you want to get super pedantic about what’s happening in a transistor every digital Boolean is actually this

54m agoHN ↗

Not quite.. that boolean is about whether the voltage exceeds some threshold. It's not about how close the voltage is to the circuit's maximum possible threshold, or how much it exceeds the threshold.

In an analog circuit, maybe.

26m agoHN ↗

The whole "no hallucinations" premise is based on that.

Like, yeah, you don't hallucinate, but only because you force the user to decide in the end.

14m agoHN ↗

Wait til you hear about how digital circuits work at die level

13m agoHN ↗

That seems like the only possible way to eliminate hallucination, short of a model that is never wrong.

10m agoHN ↗

force the user to decide in the end

And that's...bad?

16m agoHN ↗

I like it. It is short and distinct which is a good fit for a primitive. It describes its fundamental meaning and draws a connotation with Boolean.

53m agoHN ↗

Are there "good" Open source Decision models built on Gemma-4 and also trainiable on own data?

48m agoHN ↗

Jeeves, that's a name I haven't heard in a long time...

42m agoHN ↗

If Jeeves returned as an AI chat bot it would be the most brilliant resurgence of nostalgia

34m agoHN ↗

jeeves is currently the name of my local hosted assistant, in its context there are rules that tell it to behave like good old jeeves.

soon I'll make sure that my home assistant pod answers to "Hey jeeves"

28m agoHN ↗

Personally I’m happy that after a 30 year effort and hundreds of billions spent, AskJeeves finally works as intended.

11m agoHN ↗

We locked him in the basement with Clippy, Bob and BonziBuddy. Who opened the damn basement door???

34m agoHN ↗

This isn't really surprising. LLM reasoning and before that, chain of thought prompting are essentially forms of test-time compute scaling.

30m agoHN ↗

Seems like this is the way, a hybrid approach where some of the pipeline will be jev like and some traditional LLM depending on the nature of the work.

12m agoHN ↗

Ask Jeeves - Only took us 30 years to come full circle.

12m agoHN ↗

Super dope. If it would ship as prod ready code supporting mps as well that would be even doper.

But funny that jev is getting its lunch eaten apparently in under two weeks?

12m agoHN ↗

interesting bench list, what about benchmark against smaller or bigger models? 9B looks too huge for small like laya, and too small for llm-level decisions.

9m agoHN ↗

See if it can beat Jev's Pokémon benchmark

6m agoHN ↗

what about benchmark against smaller or bigger models? 9B looks too small for llm-level decisions.