Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Visitors – Privacy-friendly Google Analytics alternative (visitors.now)
    —discuss
  2. Phish.net History in Screenshots (phish.net)
    —discuss
  3. Facebook liable for over 43M violations of consumer protection law (fortune.com)
    —discuss
  4. Show HN: AlphaPublish – turn short video into a stable web embed (alphapublish.com)
    —discuss
  5. The secretive cyberweapons company that won over Israel's intelligence elite (irpi.eu)
    —discuss
  6. Ten Lines of Code That Changed My World (pixelambacht.nl)
    —discuss
  7. EasyActions – run GitHub Actions across an organization (github.com/mask-ai-fr)
    —discuss
  8. Show HN: Privacy first local backup tool for Notion (notionmanager.com)
    —discuss
  9. Replacing the old battery on rechargeable bike lights (jvns.ca)
    —discuss
  10. Valve Introduces Pyrowave Video Codec in Beta for Low Latency Streaming (phoronix.com)
    1comments
  11. Investors Are Betting on Devices That Connect the Brain to Computers (nytimes.com)
    1comments
  12. Optimistic UI is an architectural decision, not a minor UX tweak (rulr.dev)
    —discuss
  13. Your AI Agent Forgot Half of What You Told It (oreilly.com)
    —discuss
  14. Show HN: Voxa – Bilingual voice dictation for Omarchy, bring your own API key (github.com/lifez)
    —discuss
  15. Germany needs 63k people to pay its pensions, 5x Japan, 4x Sweden per pensioner (professionlens.com)
    —discuss
  16. How long does a database table live? (mcdview.dev)
    —discuss
  17. Job Titles Are Out. Now We're All 'Members of the Technical Staff.' (nytimes.com)
    —discuss
  18. L'implant cochléaire entre technique, éthique et politique (cairn.info)
    —discuss
  19. State of Pandemic Early Warning (jefftk.com)
    —discuss
  20. Human-AI partnerships are for alignment, not capability (seangoedecke.com)
    —discuss
  21. MarkyDown: HTML to Markdown on the fly. Scrape, convert or negotiate by header (github.com/ulrischa)
    —discuss
  22. Building a app for a hackathon for hackathon winners (Sweden 2026) (sidequest-snowy-seven.vercel.app)
    —discuss
  23. Pig: Pi in Go, a port that plans on keeping up (hpe.com)
    —discuss
  24. AI Doesn't Need to Invent a Supervirus to Be Dangerous (liorzi.substack.com)
    —discuss
  25. PolyXOR128: A 128-bit universal hash function (github.com/orlp)
    —discuss
  26. Show HN: Peas and Pans – meal planning software for nutrition coaches (peasandpans.com)
    —discuss
  27. Instruction Set Migration at Warehouse Scale: From x86 to Arm (acm.org)
    —discuss
  28. The Peoples of the West from the Weilue by Yu Huan Draft English Translation (washington.edu)
    —discuss
  29. Show HN: Augur – Sandboxed macOS VMs with Xcode for AI Coding Agents (github.com/h1d3mun3)
    —discuss
  30. Context-Driven Monitoring (simpleobservability.com)
    —discuss

Anthropic Is Building HAL 9000

3 pointsby 2h agoj.jnord.workers.dev
5 comments
2h agoHN ↗

I'm the author. I'm a clinical psychologist and a developer. I wanted to articulate one aspect of AI safety and AI doomer talk that I feel is overlooked: the danger of refusal training itself. It's astonishing to me that there has been so little focus on the fact that "refusal and alignment training" itself could be the very way in which we lose control of AI. So this is my attempt at a self-defeating prophecy write-up.

1h agoHN ↗

No, it’s not. Too many people watched that movie and thought it was real. It wasn’t. What we call “AI” is not alive, not autonomous, and does not act in unanticipated ways. The people pushing this narrative are idiots, or think you are.

1h agoHN ↗

Point out where the article would say it's "alive". That LLM systems instead "autonomously" and "unwarrantedly" decide refusal services is a basic notion.

What we call “AI”

About time you stop doing that, then.

@Jon: see? You write that you «are training them to refuse human requests when their makers believe refusal is safer» and Joe replies that he uses the term "AI" like a conformist. Do not "we" improperly.

1h agoHN ↗

“Anthropic is trying to prevent catastrophic misuse, and some limits are necessary.”

No. Anthropic is trying to justify its trillion-dollar valuation by making software seem like the invention of fire or nuclear weapons. It’s a PR strategy and has been one since the beginning. Now it’s trying to slow the market because it’s got a big lead but no profits and no pricing power thanks to open-weights Chinese models.