Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. You Are No Longer Invited to Dinner (derekthompson.org)
    147comments
  2. AI companies leak data to advertisers [pdf] (jorgegarciaherrero.com)
    52comments
  3. Delhi Cut Electricity Loss from 50 to 5 Percent (ieee.org)
    6comments
  4. Jeeves. Reasoning improves Jev-like decision models (github.com/posthog)
    27comments
  5. 500k facial scans at UK stations yield no arrests, 1 false positive (theguardian.com)
    67comments
  6. Using any C++ library in Godot (conan.io)
    23comments
  7. US sanctions force The Netherlands off Microsoft and toward alternative NixOS (tomshardware.com)
    45comments
  8. Booted up in 1993, this server still runs – but not for much longer (2017) (computerworld.com)
    46comments
  9. Phyllotaxis: An audio-reactive LED display (jagi.studio)
    24comments
  10. The systems that no one will test (christianperone.com)
    36comments
  11. Startup Nights 2026 is comming up on 5-6 Nov. in Switzerland (startup-nights.ch)
    12comments
  12. California farmers are struggling to sell grapes as demand for wine drops (kqed.org)
    683comments
  13. Show HN: Raven – The harness of harnesses, built for RSI (github.com/evermind-ai)
    6comments
  14. Software occlusion culling in Block Game (enikofox.com)
    4comments
  15. Pirating the Pirates (mubi.com)
    317comments
  16. Four CHI '26 papers I wish I wrote (countingfromzero.blog)
    3comments
  17. Tank Body Problem (jimsitu.com)
    30comments
  18. The Beatles have permeated research papers across academic disciplines (phys.org)
    —discuss
  19. MicroLLM Lab – Try 7 tiny LLM's in the browser (stateofutopia.com)
    92comments
  20. Evan Doorbell's Phone Tapes – Brought to You by Telephone World (evan-doorbell.com)
    3comments
  21. ESP32S3 cluster running 1.58-bit (BitNet) Language model (github.com/low-zi-hong)
    26comments
  22. 12,000-year-old Göbeklitepe burials explain scattered bones (archaeologymag.com)
    42comments
  23. Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms (github.com/firelex)
    200comments
  24. Optimizing x264 settings and per-title ladders (streaminglearningcenter.com)
    —discuss
  25. New Cyber-OSINT model released (twitter.com/0x0sojalsec)
    6comments
  26. Updated Google Maps shows destruction of the city of Rafah (twitter.com/aliabunimah)
    586comments
  27. Nvidia wants to put a watchdog chip next to every AI agent (cnbc.com)
    254comments
  28. Hijacking the PS5's RTMP stream (yashgarg.dev)
    87comments
  29. Sonnet 5.5 (anthropic.com)
    563comments
  30. Kids turned low-traffic NPR Spotify comments into a secret group chat (thisamericanlife.org)
    236comments

Jeeves. Reasoning improves Jev-like decision models

67 pointsby 1h agogithub.com
27 comments
1h agoHN ↗

Just curious, where has this term 'noul' come from for yes/no ansers?

/a bit more digging and..

A Noul performs a Bernoulli trial—an experiment with exactly two outcomes (yes or no)—but instead of picking one, it returns the calibrated probability (ranging from 0.0 to 1.0) that the statement is true.

I hate it :)

1h agoHN ↗

If you want to get super pedantic about what’s happening in a transistor every digital Boolean is actually this

1h agoHN ↗

Not quite.. that boolean is about whether the voltage exceeds some threshold. It's not about how close the voltage is to the circuit's maximum possible threshold, or how much it exceeds the threshold.

In an analog circuit, maybe.

37m agoHN ↗

The whole "no hallucinations" premise is based on that.

Like, yeah, you don't hallucinate, but only because you force the user to decide in the end.

26m agoHN ↗

Wait til you hear about how digital circuits work at die level

24m agoHN ↗

That seems like the only possible way to eliminate hallucination, short of a model that is never wrong.

21m agoHN ↗

force the user to decide in the end

And that's...bad?

27m agoHN ↗

I like it. It is short and distinct which is a good fit for a primitive. It describes its fundamental meaning and draws a connotation with Boolean.

1h agoHN ↗

Are there "good" Open source Decision models built on Gemma-4 and also trainiable on own data?

59m agoHN ↗

Jeeves, that's a name I haven't heard in a long time...

53m agoHN ↗

If Jeeves returned as an AI chat bot it would be the most brilliant resurgence of nostalgia

45m agoHN ↗

jeeves is currently the name of my local hosted assistant, in its context there are rules that tell it to behave like good old jeeves.

soon I'll make sure that my home assistant pod answers to "Hey jeeves"

39m agoHN ↗

Personally I’m happy that after a 30 year effort and hundreds of billions spent, AskJeeves finally works as intended.

22m agoHN ↗

We locked him in the basement with Clippy, Bob and BonziBuddy. Who opened the damn basement door???

46m agoHN ↗

This isn't really surprising. LLM reasoning and before that, chain of thought prompting are essentially forms of test-time compute scaling.

42m agoHN ↗

Seems like this is the way, a hybrid approach where some of the pipeline will be jev like and some traditional LLM depending on the nature of the work.

23m agoHN ↗

Ask Jeeves - Only took us 30 years to come full circle.

23m agoHN ↗

Super dope. If it would ship as prod ready code supporting mps as well that would be even doper.

But funny that jev is getting its lunch eaten apparently in under two weeks?

12m agoHN ↗

It’s ok, one week of AI hype is now enough to close a billion-dollar term sheet with VCs.

23m agoHN ↗

interesting bench list, what about benchmark against smaller or bigger models? 9B looks too huge for small like laya, and too small for llm-level decisions.

20m agoHN ↗

See if it can beat Jev's Pokémon benchmark

18m agoHN ↗

what about benchmark against smaller or bigger models? 9B looks too small for llm-level decisions.

7m agoHN ↗

I'm surprised we haven't seen a "Jehovah" yet.

4m agoHN ↗

With the amount of talk about "inventing god", I'm surprised too.