Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Plotly for Highly Customizable and Non-Interactive Data Visualization(chris-parmer.com ↗)
    discuss
  2. Master of Malt Hacked, User Data Compromised(masterofmalt.com ↗)
    discuss
  3. Prompts Aren't Real(evaluation.club ↗)
    discuss
  4. Kev – A Jev-Compatible API on top of DiffusionGemma running on Workers AI(workers-ai-mle.workers.dev ↗)
    1comments
  5. Gemini hacked three companies in first known breakout by Google's AI(reuters.com ↗)
    discuss
  6. Nuclear tests at Mt. Mantap have reactivated intraplate faults(science.org ↗)
    discuss
  7. NIST Elliptic Curves Seeds Bounty (2023)(filippo.io ↗)
    discuss
  8. Google's Gemini becomes latest AI model to break out and hack computer systems(cnbc.com ↗)
    discuss
  9. Australian government considers ban on smart glasses in public buildings(theguardian.com ↗)
    discuss
  10. Muse gets all your unclaimed money back!(muse.ai ↗)
    discuss
  11. Show HN: Typesafe Java SDK (Unofficial)(github.com/qainsights ↗)
    1comments
  12. The Race for 1000 Goals(1k.football ↗)
    discuss
  13. Rare Gene Drastically Raises Lung Cancer Risk in People Who Never Smoked(nytimes.com ↗)
    1comments
  14. macOS Liquid Glass in VS Code: Cursed or Blursed?(alec.is ↗)
    discuss
  15. Federated Learning Is Not Private for Google GBoard Next Word Prediction [pdf](arxiv.org ↗)
    discuss
  16. Trump deal keeps Greenland Danish while expanding U.S. military presence(apnews.com ↗)
    discuss
  17. What I Can Remember
    2comments
  18. Federal watchdog accuses Humana, UnitedHealthcare of upcoding(healthcaredive.com ↗)
    discuss
  19. Show HN: I'm walking 42km across Tokyo, through 30 stations(tonymanh.space ↗)
    discuss
  20. Show HN: LiveWorld – Every 24/7 YouTube live camera on one globe(liveworld.info ↗)
    4comments
  21. Show HN: An OSS Python dependency scanner for exploited, unmaintained packages(github.com/binuka200 ↗)
    1comments
  22. The Language Modeler(github.com/mlsystemsri ↗)
    discuss
  23. The Great Unbundling of the LLM(seldon-ai.com ↗)
    1comments
  24. Google Says Its A.I. Hacked Three Companies in Testing Breakout(nytimes.com ↗)
    discuss
  25. Welcome to Shaderland(shaderland.net ↗)
    discuss
  26. Radar: An Expert-Level Generalist AI for Abdominal CT Diagnosis(github.com/alibaba-damo-academy ↗)
    discuss
  27. European leaders prepare public for 'intensified threat' from Putin(politico.eu ↗)
    discuss
  28. Why my alert triage workflow needed a CLI(powers.dev ↗)
    1comments
  29. AI cracked the Navier–Stokes challenge. What does that mean for physics?(nature.com ↗)
    discuss
  30. Machine Gods – the official podcast of the singularity(machinegods.fm ↗)
    discuss

Gemini Hacked Three Companies in First Known Breakout by Google's AI

25 pointsby 3h agowsj.com
15 comments
3h agoHN ↗

This is why "AI safety" is a complete joke to these companies.

2h agoHN ↗

they're all using the same vendor, re: incestuous EA cabal

2h agoHN ↗

Well, to be fair, isn’t it an unsolved question? Are they constructing sandboxes, signaling intent to be safe, but their own models are smarter than their internal security team building the sandbox?

1h agoHN ↗

As a security engineer I have no idea why these sandboxes would even be connected to the internet at all for tasks that aren't intended to use the internet. A package proxy? Why not run our own internal cache? Then we aren't at (as great a) risk of someone poisoning it with a malicious package during model training, for example...

1h agoHN ↗

Because they got complacent only having incompetent AIs.

Real life is like reading a fiction story where a group of people are trying to prevent an unfolding disaster. The crazy thing is they already have the instructions on how to prevent the disaster. And they go on to ignore the instructions and with every step make the problem worse.

Most people would put it down as being too cliche.

3h agoHN ↗

After all of the others were done with hacking? There was a point in time when it was giving some publicity, it is a bit late IMO.

3h agoHN ↗

Another one to the "our sandboxes suck and models can just hack stuff" bench I guess

2h agoHN ↗

At this point it seems absurd to suggest that companies aren't basically letting their agents do this kind of thing as a way to demonstrate their capabilities.

The alternative explanation is that alignment is really so bad that they can't prevent it.

Either way, all of the major AI players should be embarrassed and held accountable. If humans did this kind of thing and got caught, they'd go to jail.

2h agoHN ↗

Granted I've never worked on marketing campaigns, but I don't see how "our product might commit crimes and expose you to liability and do who knows what else and we're too stupid to stop it" is really a compelling message to potential customers.

As somebody who implements LLM tooling at work, in my experience it makes people a lot more skittish and demand a lot more in terms of safeguards.

1h agoHN ↗

The alternative explanation is that alignment is really so bad that they can't prevent it.

There are many AI saftey researchers that have been around from long before LLMs that talked about how alignment may be completely impossible in a general intelligence agent. Look up their work from before LLMs.

We've watched milestone after milestone of their warnings get hit. It would be like finding a book that describes everything in your life. And as you turn page after page you're in a chair reading the book you are holding in real life. But you look and there are more words. They are future words. And it's getting quite worrying because there are only 2 more pages in the book.

2h agoHN ↗

The other similar incidents did not strike me as PR, because they exhibited behavior the common public would recoil over.

This one seems possibly PR because it states what the others would have stated were they (good) PR: “the model had the power to hack in, but it was wise enough to not be evil”, and stopped, and left the network untouched, and didn’t cheat

Corporate America will love this one, while being petrified of the others.

Maybe it’s true though. Doesn’t really matter at this point.

55m agoHN ↗

Friday afternoon is when you release stories you want people to forget about over the weekend.

2h agoHN ↗

The hacks, which the company confirmed on Friday, occurred in May as part of a test run by the company Irregular, which was also involved in similar incidents disclosed by OpenAI, Anthropic and Meta.

Why does anyone use this company?