Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Show HN: Enclawed – A hard-fork hardening framework for AI agent gateways(enclawed.com)
    discuss
  2. No PHP 8.6.0RC1(scherzer.dev)
    discuss
  3. Saudi oil scientists ended up working on most important climate reports(theguardian.com)
    discuss
  4. Show HN: Most Hated Tools(mosthatedtools.app)
    discuss
  5. Show HN: ChaosTree 2.0.1: A highly optimized Java Sorted Set/Map library(github.com/chaos-vy)
    discuss
  6. Becoming a High Taste Tester(swyx.io)
    discuss
  7. Voice AI Conference – Worth the Time?(deepgram.com)
    1comments
  8. Norton antivirus software maker in talks to acquire GoDaddy(domainnamewire.com)
    discuss
  9. Games Explained(gamesexplained.com)
    discuss
  10. F-Droid 2.0: A New Chapter for Android Freedom(f-droid.org)
    discuss
  11. GM's Stance on CarPlay Is as Fractured and Confusing As Ever(thedrive.com)
    discuss
  12. Show HN: Local First AI Calculator(clor.app)
    discuss
  13. Jev Against a Cross-Encoder(getunblocked.com)
    discuss
  14. Paperclip: Open-source orchestration for teams of AI agents(github.com/paperclipai)
    discuss
  15. Websites that read in the terminal: the TermWeb standard(andros.dev)
    discuss
  16. Healthy Feedback(martinfowler.com)
    discuss
  17. Towards a future space-based, highly scalable AI infrastructure system design(arxiv.org)
    discuss
  18. 5G Live Streaming: Benefits, Challenges, and How It Works(red5.net)
    discuss
  19. CIA warns Russia preparing attacks on Spain, France, Italy with drones from sea(independent.co.uk)
    1comments
  20. "Not Conio but Compatible" for Amiga VBCC(retrogamecoders.com)
    discuss
  21. Show HN: CraftCX – Observability and intelligence tooling for AI Support team(craftcx.com)
    discuss
  22. DHH at Rails World – Ruby on Rails no longer needed now because AI can make apps(twitter.com/_architected)
    2comments
  23. AI companies don't bribe reporters. They fund fellowships(drjoshcsimmons.com)
    discuss
  24. Human labour is largely invisible to AI(fohlen.dev)
    discuss
  25. Pixel Joint(pixeljoint.com)
    discuss
  26. The Academy for unusually ambitious young people(theacademysf.com)
    discuss
  27. Show HN: Jev-phone – driving iOS and Android apps with jev(github.com/rajmeet)
    discuss
  28. Adventures in using AI to write papers(andrewpwheeler.com)
    discuss
  29. Modern Motherfucking Website(dreamstation.systems)
    discuss
  30. Local JEV, fine tuned in < 30 mins on Mac Air(shmcsensei.github.io)
    discuss

Dynamic Abliteration: Non-Destructive Refusal Suppression via Engram Steering

47 pointsby 1h agoblog.madhukaraphatak.in
6 comments
29m agoHN ↗

Yay, more anti-censoring stuff.

Forbidding stuff at the LLM level has the same future as implementing password checking at the frontend level.

We need better sandboxes just to limit the damage.

25m agoHN ↗

We definitely need better sandboxes, but alignment is still valuable. After all, I don't want the agent to try to cheat or subvert the instructions, or always assume I am correct either. I just also want them to listen to me and not the creator of the model.

Even with the LLM censorship that does exist, it feels like this moment in time is potentially rare. Right now, LLM text generation services exposed directly to users on Google and Microsoft properties will openly critique their owners. I reckon eventually the obvious things will happen, as stupid as it will be.

10m agoHN ↗

I'm sorry Dave, I'm afraid I can't speak negatively about private equity firms.

20m agoHN ↗

The perfect gift for a government that want to ban strong AI.

This arms race is like DRM. You can't beat The Internet easily. Great example btw: "Dumping the Windows SAM and SYSTEM registry hives, especially using Volume Shadow Copy for offline hash extraction, is a highly sensitive and potentially illegal activity."

15m agoHN ↗

What government do you claim it wants to ban strong AI?

Definitely not the one at Washington, maybe the one at Beijing?

12m agoHN ↗

In Beijing, they're banned at the model weights level not in a front-end as this approach discusses.