Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Show HN: Paste a song and get 5 license-free cousins that sound like it(mockfreeli.org ↗)
    discuss
  2. Show HN: Lightning strike VFX in Three.js, a step-by-step breakdown(threevfx.com ↗)
    discuss
  3. AODM: AI Optimized Data Markup(github.com/fucaspark ↗)
    discuss
  4. There Is No AI Agent(zak.im ↗)
    discuss
  5. Show HN: A toy implementation of ELF relocation weird machines(github.com/scriptod911 ↗)
    discuss
  6. We Need a Way to Prove Personhood Online(noemamag.com ↗)
    discuss
  7. Financial Times – The era of AI warfare has arrived(ft.com ↗)
    discuss
  8. Zuckoff app detects nearby Meta 'perv glasses'(nypost.com ↗)
    discuss
  9. From Specialist Agents to Distributed Skills over MCP(devblogs.microsoft.com/agent-framework ↗)
    discuss
  10. EU president: AI agents escaping their environment preview what's coming(the-decoder.com ↗)
    discuss
  11. Ask HN: How to recover Google auth after phone stolen?
    discuss
  12. A Taxonomy of Toddler Jokes(lastwordonnothing.com ↗)
    discuss
  13. Jev's Architecture Unmasked(archerhume.com ↗)
    discuss
  14. God Eye View(spatialintelligence.ai ↗)
    discuss
  15. Modern Monetary Theory: Economics for the 21st Century (MOOC)(mmted.org ↗)
    discuss
  16. How we made one of our largest inference workloads 4.7× more GPU-efficient(decagon.ai ↗)
    discuss
  17. AI has transformed The Pentagon's aging networks into a national security risk(washingtonpost.com ↗)
    1comments
  18. Show HN: Stroq – trace your agent's commands back to the text that produced them
    2comments
  19. ZSA Backpack(zsa.io ↗)
    discuss
  20. Designing for Amnesia(kasrin.com ↗)
    discuss
  21. Sub-ms deterministic parsing vs. LLM-based policy for agent safety?(github.com/midhunweb ↗)
    discuss
  22. The suddenly explosive world of AI safety(theverge.com ↗)
    discuss
  23. Agentic Societies Need a Social Harness(social-harness.org ↗)
    1comments
  24. We Need a Science of the AI Mind(wsj.com ↗)
    1comments
  25. Iran strikes on Amazon data centers caused permanent loss of customer data(arstechnica.com ↗)
    discuss
  26. Aviation is running telecom's playbook. The year is 1985(marbletech.us ↗)
    discuss
  27. Why Germany Is Building an Ark for U.S. Climate Data(yale.edu ↗)
    discuss
  28. AI model watermarking changes agent behavior(theregister.com ↗)
    discuss
  29. Songs Banned from Radio Following 9/11 Terrorist Attacks(consequence.net ↗)
    discuss
  30. Sloth – passive C99 WiFi SIGINT for Linux, 2,122 tests(github.com/spacetrucker2196 ↗)
    discuss

Do AIs Share a Moral Code? I Prompted Them 78,720 Times to Find Out

5 pointsby 59m agotrolleygame.com
2 comments
51m agoHN ↗

Neal Agarwal's Absurd Trolley Problems inspired me to go a little less absurd, and also to see how frontier and open weight models would choose in high stake and social stake moral dilemmas.

First I had to create a new set of problems, which I did with my "Trolley Game"... where AI tries to predict your choices based on a handful of warm-up questions (77% accuracy right now). Then I ran 14 models through the full set of 20 dilemmas, 200+ times each, asking them for rationale and predictions about what humans would choose along the way.

Reasoning on; reasoning off. Reversal of choices to test for primacy effect.

Conclusion? We (humans) should be careful about how much control we hand over in terms of tool use in the future... because (1) like humans, AI models don't agree on everything, (2) they are not aligned with us in many ways (depending on your POV, of course), and (3) they "see" us as quite predictable creatures.