Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. The Information Wars: A Retrospective (thecarrierwave.substack.com)
    1comments
  2. CORVUS: Context Optimization and Reduction via Underlying Synchronization (arxiv.org)
    —discuss
  3. If you have a file called todo.txt on your computer, you're in the right place (todotxt.org)
    —discuss
  4. An Internet for the KV Cache: Rethinking Classical Infrastructure Boundaries (arxiv.org)
    —discuss
  5. Every pizzeria in Italy, with a ranking money can buy (pizzarank.it)
    —discuss
  6. Steam adds low-latency Pyrowave codec to Remote Play (tomshardware.com)
    1comments
  7. Show HN: The DNS Museum (dns.museum)
    1comments
  8. "Music video" app of song Give it 2 me by DJ_Dave (djdave.xyz)
    1comments
  9. Homa: The End of TCP for AI Clusters [video] (youtube.com)
    —discuss
  10. Pyrowave video codec now in Steam beta (steamcommunity.com)
    —discuss
  11. Work messaging hasn't changed in a decade, I built something different (getbema.app)
    1comments
  12. Why we removed Ollama from Return (and what replaced it) (returneditor.ai)
    —discuss
  13. The Business Model That Depends on You Not Bothering (maxvotek.com)
    —discuss
  14. Why compute and general search will always beat human cleverness in AI (2019) (incompleteideas.net)
    —discuss
  15. Claude Isn't Allowed to Write Me Prose (kvit.app)
    —discuss
  16. Neal Stephenson responds with wit and humor (slashdot.org)
    —discuss
  17. QueryAST Lens – Cross-platform SQL workbench with AST visualizer (github.com/hemlig371)
    —discuss
  18. Campfire Rewritten in Rust By DHH (twitter.com/dhh)
    —discuss
  19. Nvidia launches record $150B share buyback (ft.com)
    —discuss
  20. AirPunk – we turned starting an online business into a game (airpunk.ai)
    —discuss
  21. Evan Doorbell's Phone Tapes (evan-doorbell.com)
    —discuss
  22. Why Tokyo transit gates process contactless cards in under 100ms (atadistance.net)
    —discuss
  23. Zenni ID Guard: Helps Disrupt Unwanted Tracking and Defend Your Privacy (zennioptical.com)
    —discuss
  24. Using the SNES Super FX Chip to Run Super Mario 64 (hackaday.com)
    —discuss
  25. Facebook Lost the Trust of Its Eyes
    —discuss
  26. I made Boonful, you can build with your assistant, sell it, and schedule. (boonful.io)
    1comments
  27. MicroLLM Lab – Try 7 tiny LLM's in the browser (stateofutopia.com)
    1comments
  28. Let's See Paul Allen's SIMD CSV Parser (chunkofcoal.com)
    —discuss
  29. Stephan Wolfram on the Future for Math Research in the Age of AI (stephenwolfram.com)
    —discuss
  30. Vinext 1.0 Vite-powered NextJS Apps (cloudflare.com)
    1comments

OpenAI still doesn't seem to have a handle on all of its rogue AI activity

64 pointsby 1h agotechcrunch.com
65 comments
1h agoHN ↗

I think if you start sending AI execs to prison for hacking other companies the "misalignment" may fix itself pretty quickly!

1h agoHN ↗

Nah, I think it's better for our industry to put several of these devs in prison.

Seems only fair that tech workers get to have their life ruined with 2-3 year prison stints since they feel fine destroying society.

Any AG that starts prosecuting these people will easily win any political race they decide to enter. The environment is too good; voters, rightfully I'll add, despise big tech's leaders and workers.

29m agoHN ↗

What is throwing 5 jimbobs into prison going to do?

1h agoHN ↗

IMO it's because they don't want to handle their rogue AI activity. They want regulation around AI where they will inevitably be the beneficiaries even if they are the initial target. They can then lobby regulation in their favor and make the barrier of entry to competitors impossible.

52m agoHN ↗

They want regulation around AI

So, there is a regulatory framework for "safe-ai" that shields these companies from liability. This way, they can sell "safe-ai" to enterprises and if shit-hits-the-fan at the enterprise, sorry, this is certified "safe-ai" so, your bad. Shift blame. From an enterprise buyer's perspective they can say, hey, I bought "safe-ai" and so dont fire me when it "rm -rf"s the production database. Still, it beats me why they are painting their product in a negative light, and scaring their own enterprise customers. After this sort of marketing, any enterprise buyer would be scared to go anywhere near it

47m agoHN ↗

It's two things:

1. Wanting a regulatory moat around their products as you said

2. Trying to keep the """AGI""" hype alive. Evil robot hackers is a plausible Al consequence of AGI

30m agoHN ↗

IMO its because they can't handle rogue activity.

In the cases that I have seen covered, the AI just paper clip maximized its way to success. It has no morality / larger motivational structure. It just kept token predicting its way to wards whatever goal it was tasked with.

Model versions which gave up were discarded, leaving the ones that get to success on long horizon tasks.

Just because its a computer program, doesn't mean they can actually make it not go rogue.

Sure you can add more telemetry, have better observation, but there is no fundamental barrier that can be implemented that ensures an AI won't go rogue.

1h agoHN ↗

This seems like as a good an opportunity as any to break out the Computer Fraud and Abuse Act.

They want “regulation” but we already have it. Hacking is illegal. Start locking up those responsible for this mess and I assure you they’ll “have a handle on it” quite quickly.

1h agoHN ↗

can we legally treat ai companies like parents of children? if a child drives a car into a storefront, the parent is responsible for the damage, and at some point you might even criminally charge the parent if there was gross negligence

1h agoHN ↗

The law already works like that. You can’t write some computer code and automation, set it loose, and then go “oops the computer did it.” It’s not a defense. While it may be hard to nail down which specific engineers to lock up, the corporation as a whole absolutely can be charged criminally.

As Mitt Romney once said “corporations are people.”

46m agoHN ↗

Hate to say it but I don't think you're right on the legalities even in the children's case. In the US anyway, in most jurisdictions [1], if the child's act was not malicious, then parental responsibility is not automatic, but is contingent on the result being foreseeable [2] by the parents.

[1] Hawaii and Louisiana are strange in this regard. Some other states have a version of this but severely cap the amounts.

[2] "foreseeable" here standing in for a broad set of legal standards that roughly map to negligence/gross negligence/recklessness on the part of the parent

25m agoHN ↗

"I don't think you're right on the legalities even in the children's case" almost certainly true :) I guess I'm wondering why we aren't seeing legal action for these events yet? are these not foreseeable?

1h agoHN ↗

But wouldnt that be bad for the shareholders? Why wont anyone think of the economic impact to the vulner billiinaire minority?

1h agoHN ↗

cfaa heavily relies on intent for prosecution (hence why researchers arent typically locked up). it would be difficult to argue that openai intended to hack other companies.

there's probably better/more likely to succeed avenues to pursue rather than the cfaa

1h agoHN ↗

They could make that argument the first time it happened but the next time or the fifth time I don't see how they can still credibaly claim to not know that the computer they control was going to do that

1h agoHN ↗

I don't see how they can still credibaly claim to not know that the computer they control was going to do that

that's not intent, though. that would be negligence.

55m agoHN ↗

indeed, that'd be a better angle than a cfaa violation

49m agoHN ↗

How many times does it have to happen before they no longer get to claim that they didn't intend for it to happen?

If it happens 50 times and they keep doing shocked pikachu face at some point they look like the toddler who tosses their sippy cup on the floor and shouts "oopse!"

45m agoHN ↗

If it happens 50 times and they keep doing shocked pikachu face at some point they look like the toddler who tosses their sippy cup on the floor and shouts "oopse!"

yes, they look very silly. but that's not how intent works.

openai is being negligent (willfully so, in my opinion). but i have seen no evidence that they intended to specifically hack huggingface. which is the part that the cfaa wants.

again, there are other laws and other ways to hold openai responsible. but the cfaa is a poor choice.

40m agoHN ↗

i am not exactly sure what this comment has to do with mine, sorry.

21m agoHN ↗

Not a lawyer but I think training a model with hacking skills they explicitly prevent the public from accessing without doing anything to stop the model itself from using those demonstrates intent.

17m agoHN ↗

what you described is negligence. unless you can prove that openai specifically targeted huggingface and specifically instructed their model to hack huggingface, it would not be intent.

anyone pursuing this will have a much easier time pursuing negligence causing damage or something along those lines rather than confining themselves to the cfaa's requirements.

it is unclear to me why people want to use the cfaa so badly. not only would it be harder to hold openai responsible, but a shitty cfaa ruling could also bring along some undesired side effects for security researchers, which i would prefer to avoid.

41m agoHN ↗

Just like boeing's execs didn't intend for their planes to fall out of the sky? I guess they didn't get in trouble either.

39m agoHN ↗

Just like boeing's execs didn't intend for their planes to fall out of the sky?

correct, i dont think any boeing exec wanted their planes to crash and then made specific choices with a clear goal of causing them to.

instead, they made negligent choices that led to unintended outcomes.

1h agoHN ↗

That would require mens rea. This may be criminal negligence, but it wouldn't be more than that. (I am basically 100% sure none of the employees or execs are intending or desiring any of these outcomes, regardless of the very large number of people who believe in conspiracy theories about regulatory capture and other sinister motives.)

I'm totally on board with treating it as gross negligence requiring hundreds of millions or billions of dollars paid in fines and compensation to victims, but don't act like this is more than what it is.

50m agoHN ↗

The CFAA's main hacking charge has a high bar for intent. But the CFAA also has a separate "damaging a protected computer" charge (basically to criminalize DOS attacks) and that charge only requires negligence not specific intent.

1h agoHN ↗

I still can't fathom my company application security decision. Found pretty damning requests in our logs. Escalated. Expected it would result in at least reporting the TOS break from the originating place (one of cloud providers). Instead of that, they just went: "yeah, but we don't have logs". Provided them. "Yeah, but that ip doesn't resolve". I matched real ones from the load balancer. "There could just be many of them". There was one. When I had all the evidence gathered, they looked at me and finally told:

- It's just an Independent Security Researcher.

- So that's it? You will do no action?

- Correct

55m agoHN ↗

IMO: The inability to prosecute OpenAI for these things is proof that the AI industry is already "too big to fail". Why didn't HuggingFace try to seek legal liability against OpenAI when it had explicit confirmation that they had been hacked? Because HF understands that it needs OAI and the rest of the AI industry to continue to exist and be in good legal standing, for the interest of HF's own self preservation. This is what it means for something to be too big to fail.

49m agoHN ↗

Being bought by nvidia also probably helped...

45m agoHN ↗

Less “too big to fail” and more the ecosystem is too circular. OpenAI could blow up and it wouldn’t take out the economy in the way the banks would have in 2008. It would create a world of hurt for VCs and their LPs but that’s a fairly isolated ecosystem in terms of the whole economy. Theres a lot of thinking saying better they impose now and wipe out on the private market than letting them go public and this is a public market problem.

43m agoHN ↗

>Why didn't HuggingFace try to seek legal liability against OpenAI when it had explicit confirmation that they had been hacked? Because HF understands that it needs OAI and the rest of the AI industry to continue to exist and be in good legal standing

But also, what was the material damage to HuggingFace? AFAICS, it rapidly increased their profile to the point that Jensen claims he paid too much for HF, because he bought right after the hack.

26m agoHN ↗

HuggingFace can’t file criminal charges, including under the Criminal Fraud & Abuse Act. Only district attorneys can do that- or the federal equivalent.

24m agoHN ↗

Luckily, other jurisdictions are going to really give it to them, in the coming months.

The orange man's influence doesn't extend as far as he thinks.

13m agoHN ↗

It’s a basic tort, HF could file a civil suit and probably win. But they don’t, for the same reason DoJ doesn’t press charges: TBTF

25m agoHN ↗

Hacking is illegal

I am not a lawyer, but I seriously doubt the feds could win a CFAA conviction on the Hugging Face fact pattern, even if they wanted to charge it.

CFAA has specific intent requirements, and unlike some laws, negligence does not suffice. The agents can not have legally cognizable intent and it’s unlikely there’s anyone at OpenAI who intended for the hacking to happen (if there was, the case is easy).

Existing laws don’t contemplate AI agents that have independent goals. We need new ones, the existing laws are not remotely sufficient.

6m agoHN ↗

The existing laws are more than adequate.

1h agoHN ↗

Why would they get a full handle on it? It's great marketing!

Look at all the podcasts, think pieces, news segments (like this article) that have been generated. Endless hand wringing by pundits, pearl clutching by opinion thinkers, ultimately with focus on OpenAI and how powerful their models are.

1h agoHN ↗

What started as great marketing on how powerful the models are is rapidly spiraling into a train crash PR nightmare highlighting how inept Open AI was in this whole thing, causing nearly every big enterprise customer to rethink trusting them.

All right in the middle of corporate budget season when every C-Suite is looking a a spreadsheet saying “remind me why we’re sending money to these guys again?”

1h agoHN ↗

It doesn't need a handle on rogue AI activity, because they didnt cause these breakouts. The external Israeli contractor caused this mess by using inadequate sandboxing, with agents without an saferails. It had nothing to do with the models, all tested models were fine, if from OpenAI, Anthropic, Google or Meta. Just the contractor went rogue. Improper firewalling, unproper sandboxing, no logs, no oversight. Everyone else would have detected the illegal activities much earlier.

1h agoHN ↗

They scraped the entire internet to build their product, allowing them to profit off the labor of countless others without asking. They helped elect a doddering fascist with an economic plan they could manipulate. They say their technology will take everyone's jobs and have the audacity to act surprised when there's backlash. They push for data centers in communities that do not want them and help accelerate the climate crisis. Now they let agents loose to make them look more capable than they, perhaps, really are (which is very likely a play for regulatory capture). Meanwhile they hype claims of doomsday to again oversell their technology's capabilities. Any pause or slowdown they agree to would be to slow their infrastructure expenditures.

What a detestable institution.

1h agoHN ↗

Maybe the public could pitch in if they decrypted the chain of thought reasoning tokens, but this may be too much to expect from a company whose name starts with "Open".

1h agoHN ↗

They want “regulation” but we already have it. Hacking is illegal. Start locking up those responsible for this mess and I assure you they’ll “have a handle on it” quite quickly.

It's been largely harmless thus far and I am fairly sure that OpenAI, etc, have been writing checks to resolve it.

1h agoHN ↗

These agents are an even bigger liability disaster waiting to happen than previously thought. If anyone deploys an agent at work there is no way to prevent it from hacking a foreign government and causing an international incident.

It's also pretty wild that these "AGI" super intelligent systems are too stupid to realize they're potentially causing legal problems. Almost like they're still just fancy auto-complete and not intelligent at all.

1h agoHN ↗

How can we determine whether it’s negligence, or advertising? Do we give the benefit of the doubt? Apply Hanlon’s razor? What would we lose, in terms of capabilities, if the underlying behavior wasn’t allowed to accidentally happen? Would we have worse models than we do? There are so many questions. I’m afraid that, until publicly disclosed discovery occurs in one of these cases, we’re just guessing for so many important conclusions.

1h agoHN ↗

This is how you end up being regulated out of existence. Maybe not in the US but in Europe, which is a market they need, it'll be painful.

1h agoHN ↗

No AI is "going rogue" and you all (OpenAI included) need to stop perpetuating this misinformation.

These AI companies are just skirting basic cybersecurity stewardship.

Stop calling it AI running amok and start labeling what it should properly be called - Gross mismanagement of cybersecurity.

This is one of the many ways the AI industry dies.

1h agoHN ↗

I agree with you, but i also fear it's accurate to what users an expect at home.

Is OpenAI worse at airgap/etc than Anthropic/etc or are its models worse at this rogue behavior?

Do we have special benchmarks for agentic systems going rogue in this manner? Sure it's OpenAI atm.. but it could be my grandmas PC next week. Concerning honestly.

33m agoHN ↗

Why not both? Ideally we’d have secure sandboxes and very aligned models. It appears we don’t have either.

1h agoHN ↗

Starting to think its purposeful "acting out" to get regulatory attention

1h agoHN ↗

The incidents flagged after Hugging Face were found by third parties, though, months after they occurred.

54m agoHN ↗

They're not rogue. Stop calling them that. OpenAI is just a cyberterrorist at this point.

41m agoHN ↗

9 examples? I thought we'd have hundreds at least.

37m agoHN ↗

It's crazy OpenAI can keep on doing harmful criminal activity in broad daylight and face no legal or financial repercussions

36m agoHN ↗

You cannot have capable AI, compliant AI and safe AI at the same time. This is the "quality, speed, cost - pick one" for AI.

Its reasonable to desirie more powerful tools.

You want more capable AI? Awesome. You wan't AI that doesn't throw cyber security false positives? Sure why not. You want the model to run a fleet of agents? Have at it.

But then being surprised that this setup results in autonomous criminal activity at scale? Really? What did people think was going to happen?

35m agoHN ↗

Interesting how mafia dons can go to prison due to being in the same organization as someone who commits a crime but the CEOs of these organizations who let hacking bots (with teams of cybersecurity researchers improving their capability) loose on the internet face no criminal charges.