Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. The problem is not the AI code, but nobody knows anything anymore (ssp.sh)
    85comments
  2. Pirating the Pirates (mubi.com)
    11comments
  3. Parley: Federated, decentralised chat that speaks plain IRC (mills.io)
    107comments
  4. 13 Months Sober (2025) (bobbytables.io)
    9comments
  5. Hijacking the PS5's RTMP Stream (yashgarg.dev)
    2comments
  6. What Heraldry and Mon Can Teach Us About Building Visual-Identity Generators (benovermyer.com)
    8comments
  7. Kids turned low-traffic NPR Spotify comments into a secret group chat (thisamericanlife.org)
    48comments
  8. What makes Lisp difficult to read? (paultm.nl)
    11comments
  9. Coding Is Not Solved (alexewerlof.com)
    271comments
  10. 37,500 border drawings: a map of the world as people remember it (habibicode.org)
    30comments
  11. Show HN: PaperMono, e-ink fridge magnet shopping list with mobile web page (github.com/seamusc)
    28comments
  12. Footguns with Postgres "at time zone 'UTC'" (bookofrevenue.com)
    67comments
  13. Solving a corn puzzle with CP-SAT (thill.me)
    —discuss
  14. What Would a Serious AI Product Look Like? (glyph.im)
    12comments
  15. Owed a billion dollars in Nvidia stock (colo.to)
    414comments
  16. Show HN: Hntui – A TUI for Hacker News (github.com/ahmd-sh)
    43comments
  17. Ember-1 (fireworks.ai)
    239comments
  18. When did Google get so weird? (sancho.bearblog.dev)
    884comments
  19. Nissan's third generation e-POWER powertrain (nissan-global.com)
    295comments
  20. Show HN: Free alternative to graphics design giants (scissor.studio)
    13comments
  21. Thinking fast and slow in AI: The role of metacognition (2021) (arxiv.org)
    60comments
  22. Self-Hosting on the Dark Web (alvarezrosa.com)
    104comments
  23. Malleable software: Restoring user agency in a world of locked-down apps (2025) (inkandswitch.com)
    68comments
  24. Alan Kay's answer to “Did the ENIAC have a BIOS”? (quora.com)
    59comments
  25. Guitar amp and effects pedal built on the Waveshare ESP32-S3-Touch-AMOLED-2.06 (github.com/dashersw)
    70comments
  26. Functional Mechanical Sympathy [video] (youtube.com)
    11comments
  27. MongoDB CEO resigns "effective immediately" to join Meta, stock drops 20% (reuters.com)
    102comments
  28. Lunar Terminator Paradox (secretsauce.net)
    64comments
  29. The state of SIMD in Rust in 2026 (shnatsel.github.io)
    49comments
  30. Made by Mechanical Means (felixrieseberg.com)
    23comments

Groundtruth – checks your AI coding agent's claims against the Git diff

3 pointsby 2mo agogithub.com
4 comments
2mo agoHN ↗

From https://github.com/akahkhanna/groundtruth :

Groundtruth — a Claude Code plugin that audits whether an agent actually did what it was asked; catches the false 'Done' on Stop (missed subtasks, stubs, false 'tests pass', overridden rules).

Which other agent IDEs does or could this work with?

2mo agoHN ↗

Is this also where to attach provenance metadata and sign the agent trace, and debug?

Re: MCPSnoop, proxies like Aegis and LiteLLM, and awesome-auditable-ai: https://news.ycombinator.com/item?id=48777144#48779413

There are web standards for signing metadata with interoperable schema as linked data:

W3C JSON-LD/YAML-LD + W3C PROV + W3C DID + W3C VC Verifiable Claims

2mo agoHN ↗

"Follow up to verify that the work was actually satisfactorily completed"

Are there other sound management practices that aren't yet effectively implemented in current gen agents?

2mo agoHN ↗

I've had some success with using a more expensive model to plan a decent work breakdown structure in a document and then a low-cost model to implement the plan and keep following-up, but wonder about subsequent need for code review and code quality.

Can a lower-cost model verify completion? Iff tests and test coverage and e2e tests?

The same oracle / model routing and partitioning problem