Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Show HN: Bi-temporal Graph RAG in Postgres (new documents retire old facts)(crajah.github.io ↗)
    discuss
  2. T. Rex Had a Body Temperature of 97 Degrees(nytimes.com ↗)
    discuss
  3. Economic Policy for AGI(deepmind.com ↗)
    discuss
  4. Europe's e‑waste is a valuable resource – let's not throw it away(theconversation.com ↗)
    discuss
  5. Software Engineer: I'm Tired of Pretending [video](youtube.com ↗)
    discuss
  6. Show HN: S-Roll – local AI video understanding and clipping for Mac(saliency.dev ↗)
    discuss
  7. Show HN: Built a Simulation of Autonomous Agents(agents.london ↗)
    discuss
  8. Noam Brown – Agent swarms, alignment, & recursive self-improvement(dwarkesh.com ↗)
    discuss
  9. Show HN: Custom Relevance – Describe what matters and let Jev rank it(foglight.co ↗)
    discuss
  10. Feeling overwhelmed by the AI doom loop? Here's the essential reading list(theguardian.com ↗)
    discuss
  11. Google Noto 3D Emoji Design Process(design.google ↗)
    1comments
  12. Becoming a Benchmark(terrytao.wordpress.com ↗)
    discuss
  13. Sourcerer: Fast Git client written in C++ version 1.0.24 released(sourcererapp.com ↗)
    discuss
  14. The Download: mice with part-human brains and climate tech innovators(technologyreview.com ↗)
    discuss
  15. Should OpenAlex publish a novelty score? A vibe check(openalex.org ↗)
    discuss
  16. Show HN: Smart Crosshair – A lightweight C++/Win32 adaptive contrast crosshair(steampowered.com ↗)
    discuss
  17. I didn't sign the Fields medallists' letter(terrytao.wordpress.com ↗)
    discuss
  18. Towards Self-Driving Codebases(detail.dev ↗)
    discuss
  19. 4-Bit Rotational Quantization: -45% RAM, <1% recall drop vs. TurboQuant(weaviate.io ↗)
    discuss
  20. Show HN: Repodify: Make Podcasts Out of Podcasts(repodify.app ↗)
    1comments
  21. Palantir's Karp: AI needs to have 'reasonable guidelines,'(cnbc.com ↗)
    1comments
  22. Aegis: Zero-GC 64-byte cache-aligned memory arena in C++20 (1B ops in 0.649s)(github.com/markbgilbert ↗)
    discuss
  23. Ask HN: Co-Founder(s). Do I need any? How would I even find them?
    discuss
  24. From Stonemasons to Carpenters(thelastsoftwareengineer.substack.com ↗)
    discuss
  25. I Made Turn-Based Combat into a Database(louisxu3.substack.com ↗)
    discuss
  26. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    discuss
  27. Show HN: Pappice live demo in browser via WASM(pappice.eu ↗)
    discuss
  28. Show HN: AutoBot – live voice control for long-running AI work(github.com/demeyer1 ↗)
    discuss
  29. Show HN: Craigslist for agent skills, curated by a human(skillbay.sh ↗)
    discuss
  30. The Bicycle and the Algorithm: Amber Case on Why AI Has It Backwards(designwhine.com ↗)
    discuss

Do AIs Share a Moral Code? I Prompted Them 78,720 Times to Find Out

5 pointsby 1h agotrolleygame.com
2 comments
1h agoHN ↗

Neal Agarwal's Absurd Trolley Problems inspired me to go a little less absurd, and also to see how frontier and open weight models would choose in high stake and social stake moral dilemmas.

First I had to create a new set of problems, which I did with my "Trolley Game"... where AI tries to predict your choices based on a handful of warm-up questions (77% accuracy right now). Then I ran 14 models through the full set of 20 dilemmas, 200+ times each, asking them for rationale and predictions about what humans would choose along the way.

Reasoning on; reasoning off. Reversal of choices to test for primacy effect.

Conclusion? We (humans) should be careful about how much control we hand over in terms of tool use in the future... because (1) like humans, AI models don't agree on everything, (2) they are not aligned with us in many ways (depending on your POV, of course), and (3) they "see" us as quite predictable creatures.