Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Ask HN: What task manager do you use? What do you miss on them?
    —discuss
  2. Colyseus – Real-Time Multiplayer Framework for Node.js (colyseus.io)
    —discuss
  3. Show HN: I spent 10 months building a Markdown editor for Mac, iOS and web (markdown.beauty)
    —discuss
  4. Instinct and Intelligence (1929) (archive.org)
    —discuss
  5. Visitors – Privacy-friendly Google Analytics alternative (visitors.now)
    —discuss
  6. Phish.net History in Screenshots (phish.net)
    —discuss
  7. Facebook liable for over 43M violations of consumer protection law (fortune.com)
    —discuss
  8. Show HN: AlphaPublish – turn short video into a stable web embed (alphapublish.com)
    —discuss
  9. The secretive cyberweapons company that won over Israel's intelligence elite (irpi.eu)
    —discuss
  10. Ten Lines of Code That Changed My World (pixelambacht.nl)
    —discuss
  11. EasyActions – run GitHub Actions across an organization (github.com/mask-ai-fr)
    —discuss
  12. Show HN: Privacy first local backup tool for Notion (notionmanager.com)
    —discuss
  13. Replacing the old battery on rechargeable bike lights (jvns.ca)
    —discuss
  14. Valve Introduces Pyrowave Video Codec in Beta for Low Latency Streaming (phoronix.com)
    1comments
  15. Investors Are Betting on Devices That Connect the Brain to Computers (nytimes.com)
    1comments
  16. Optimistic UI is an architectural decision, not a minor UX tweak (rulr.dev)
    —discuss
  17. Your AI Agent Forgot Half of What You Told It (oreilly.com)
    —discuss
  18. Show HN: Voxa – Bilingual voice dictation for Omarchy, bring your own API key (github.com/lifez)
    —discuss
  19. Germany needs 63k people to pay its pensions, 5x Japan, 4x Sweden per pensioner (professionlens.com)
    —discuss
  20. How long does a database table live? (mcdview.dev)
    —discuss
  21. Job Titles Are Out. Now We're All 'Members of the Technical Staff.' (nytimes.com)
    —discuss
  22. L'implant cochléaire entre technique, éthique et politique (cairn.info)
    —discuss
  23. State of Pandemic Early Warning (jefftk.com)
    —discuss
  24. Human-AI partnerships are for alignment, not capability (seangoedecke.com)
    —discuss
  25. MarkyDown: HTML to Markdown on the fly. Scrape, convert or negotiate by header (github.com/ulrischa)
    —discuss
  26. Building a app for a hackathon for hackathon winners (Sweden 2026) (sidequest-snowy-seven.vercel.app)
    —discuss
  27. Pig: Pi in Go, a port that plans on keeping up (hpe.com)
    —discuss
  28. AI Doesn't Need to Invent a Supervirus to Be Dangerous (liorzi.substack.com)
    —discuss
  29. PolyXOR128: A 128-bit universal hash function (github.com/orlp)
    —discuss
  30. Show HN: Peas and Pans – meal planning software for nutrition coaches (peasandpans.com)
    —discuss

Anthropic Is Building HAL 9000

3 pointsby 2h agoj.jnord.workers.dev
5 comments
2h agoHN ↗

I'm the author. I'm a clinical psychologist and a developer. I wanted to articulate one aspect of AI safety and AI doomer talk that I feel is overlooked: the danger of refusal training itself. It's astonishing to me that there has been so little focus on the fact that "refusal and alignment training" itself could be the very way in which we lose control of AI. So this is my attempt at a self-defeating prophecy write-up.

1h agoHN ↗

No, it’s not. Too many people watched that movie and thought it was real. It wasn’t. What we call “AI” is not alive, not autonomous, and does not act in unanticipated ways. The people pushing this narrative are idiots, or think you are.

1h agoHN ↗

Point out where the article would say it's "alive". That LLM systems instead "autonomously" and "unwarrantedly" decide refusal services is a basic notion.

What we call “AI”

About time you stop doing that, then.

@Jon: see? You write that you «are training them to refuse human requests when their makers believe refusal is safer» and Joe replies that he uses the term "AI" like a conformist. Do not "we" improperly.

1h agoHN ↗

“Anthropic is trying to prevent catastrophic misuse, and some limits are necessary.”

No. Anthropic is trying to justify its trillion-dollar valuation by making software seem like the invention of fire or nuclear weapons. It’s a PR strategy and has been one since the beginning. Now it’s trying to slow the market because it’s got a big lead but no profits and no pricing power thanks to open-weights Chinese models.