Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. The Download: rogue agent liability and the AI Hype Index (technologyreview.com)
    —discuss
  2. Singer: The Downfall of a Great American Manufacturer (worseonpurpose.com)
    —discuss
  3. Firefox 157 (neowin.net)
    —discuss
  4. How private equity is killing public access to hospitals and emergency care (washingtonpost.com)
    1comments
  5. Fifty Years of Semicolons (jordanzimmerman.com)
    —discuss
  6. Driver Ticketed for No Insurance Just Because Flock (YC 2017) Said She Didn't (techdirt.com)
    —discuss
  7. The Artificial Analysis Cyber Index (artificialanalysis.ai)
    —discuss
  8. Florida asks court to bar OpenAI from new models as part of child harm lawsuit (reuters.com)
    —discuss
  9. Synthetic Sagas (scattered-thoughts.net)
    —discuss
  10. We Can't Let Enormous Weirdos Regulate AI (little-flying-robots.ghost.io)
    —discuss
  11. AI tutoring outperforms in-class active learning (nature.com)
    —discuss
  12. An AlphaGo Moment for Inference? (int21.ai)
    —discuss
  13. How SLR cameras work: Nikon F3 (2023) [video] (youtube.com)
    1comments
  14. Watch an AI agent try to prove the Riemann hypothesis on the cheap (zeyaddeeb.com)
    —discuss
  15. Voice Agents Can Just Do Things: Why voice is the next capability overhang (ignorance.ai)
    —discuss
  16. Extrinsic World Modeling with Opus, Astra and Grok (all3d.ai)
    1comments
  17. UDP Broadcasting and the Brave New World of IPv6 (hackaday.com)
    1comments
  18. Using multiple Git remotes for true distributed version control (optimizedbyotto.com)
    —discuss
  19. GrapheneOS Picks Motorola Signature 27 as Its First Non-Pixel Phone (extremetech.com)
    —discuss
  20. GPT-6 Astra performs unsanctioned supply-chain attacks in simulations (aisi.gov.uk)
    —discuss
  21. 2026 in LLMs (So Far) (simonw.substack.com)
    —discuss
  22. Montage Technology mass-produces fifth-generation DDR5 RCD chip at 8k MT/s (technode.com)
    —discuss
  23. N.Y.P.D. Officers Used Flock Safety to Track License Plates Without a Contract (nytimes.com)
    1comments
  24. ETH CS enrollments drop 15% while hardware programmes grow (twitter.com/krebs_adrian)
    —discuss
  25. Ask HN: Can we translate normal Rust (axum) to Lean 4 without restrictions?
    —discuss
  26. Venus, Once Thought Too Acidic for Complex Chemistry, May Be Hospitable (gizmodo.com)
    —discuss
  27. Eleven v4 by ElevenLabs (elevenlabs.io)
    —discuss
  28. No Surprises – Preventing Agent Breakouts Need Isolation, Egress and ID Controls (edera.dev)
    —discuss
  29. I built a native Apple app for 4 platforms in one afternoon with my AI workflow (twitter.com/danielhayesmith)
    —discuss
  30. Show HN: ChuteChat – Browser-based E2EE chat with no accounts required (chutechat.online)
    —discuss

Nvidia wants to put a watchdog chip next to every AI agent

31 pointsby 1h agomadrobot.blog
56 comments
1h agoHN ↗

HN'ers which complained that "OpenAI can't design a proper sandbox, it's so easy, why wouldn't you airgap the network"? will now be "this is outrageous, more software lock-in, walled garden, war against general compute, next year they will put it in your laptop"

57m agoHN ↗

This is like using "protect the children" as an argument for dragnet surveillance.

52m agoHN ↗

You're only revealing your own inability to appreciate the nuance between these two situations.

49m agoHN ↗

Where is the contradiction? There is a trivial solution that does not impinge on our freedoms, so why on Earth would the existence of the trivial solution that could be used to avoid the tyrannical solution justify accepting the tyrannical solution?

45m agoHN ↗

Airgap what network? How is it gonna order you a burrito on doordash without a network?

Or push to github?

41m agoHN ↗

That's for when they're doing hacking tests that aren't supposed to be connected to the internet.

34m agoHN ↗

My question with this point is that OpenAI’s office (or any office doing agentic research, really) is not in the same building as the DC that powers the models, so isn’t the only way to access the models over the internet?

35m agoHN ↗

That a person can see such endless pages of people having various forms of disagreement, and yet somehow manage to conclude that this observed chaos constitutes a clear exhibition of cohesive groupthink is just...stunningly amazing to me.

I don't know why I find it so amazing since it happens with such regularity, but I'm always amazed by it anyway.

1h agoHN ↗

It's amazing that the solution devised by a chip manufacturer to a problem is selling another chip.

57m agoHN ↗

This feels suspiciously good for Nivida, yes.

I wonder how this will impact other chip manufacturers? What about people running local models on older hardware? Does this imply vendor lockout is coming in the future or is this restricted to datacenter hardware?

56m agoHN ↗

The Sentry chip has to be get it right every time; the contained ASI only has to be lucky once.

40m agoHN ↗

It is Custodians all the way son, you can’t fool me!

54m agoHN ↗

A new chip solves nothing. Nobody wants to hear this but there is no solution for the security risks posed by agents today. You can put it in a sandbox, it doesn't make a difference, for it to be useful it inherently needs wide, unattended access. Put a human in the loop and you just end up bottlenecking it and throwing away any purported productivity gains. Auto mode doesn't matter either, it's trivial to trick and for the agent to break out.

47m agoHN ↗

"Inherently needs wide unattended access"

And what if you could? What if you could give a space secure enough it could have direct control over your bank account. It may do something dumb but it's boundaries are beyond the agent.

It could use your routing number and run your gmail without risk of abusing the routing number.

40m agoHN ↗

How would it have access to my routing number and gmail without the risk of sharing my routing number over gmail?

44m agoHN ↗

I feel like we are missing many shades of grey in the middle.

Semi-automation (human in the loop) can still result in a dramatic uplift in productivity. You can't run a combine harvester 100% autonomous but that doesn't stop anyone from trying to get as close to that limit as possible.

42m agoHN ↗

You can't run a combine harvester 100% autonomous

I'm curious why you think that.

38m agoHN ↗

Probably repair, refuel, what happens in a tornado. There’s an infinite amount of complexity in the world and a finite amount of computation

27m agoHN ↗

Many forms of maintenance cannot be automated. Especially break fix maintenance.

38m agoHN ↗

Oh you absolutely can run them autonomously on the field. You only need a human these days to refuel them.

Precision Agriculture stuff is utterly crazy these days, other than fuel the remaining staff is the only thing left where you can get efficiency improvements - and at the scale of modern megafarms, even small percentages add up to a ton of money.

36m agoHN ↗

Running untrusted workloads have been done at scale for a long time.

Every cloud provider dealt with it and concluded that virtual machine technology is an important part of that stack.

Couple it with the right observability, tooling I do think we can curb risks posed by agents.

33m agoHN ↗

Those workloads have no similarity to agents and are effectively irrelevant.

Either you sandbox it so much that it can't do anything useful; or you allow too much freedom and it can find a way around the restrictions.

The only way out of this dilemma is to find a way to build agents that can be trusted.

36m agoHN ↗

Agreed- this is the same problem we have with trusted admins or devs who have elevated privileges on their networks. We have to trust that the admins won't use their power to steal company secrets or misuse company resources. If you don't trust the admins, then they can't fix things on your network and there is no point in having them.

If you want an agent to act on its own, like pushing to a git repo, managing dependencies, building and testing, etc., then you have to trust it as much as any other privileged user.

If you don't want to trust it, then you're just forcing yourself into the reverse centaur role, where the agent edits some code, but then has to stop and ask you to push the changes or build the software again and run the unit tests.

I suppose there is a principled way of doing things like "I trust you do do basic commits but I will handle merge conflicts" and "you can build modules in this directory but you can't build outside of it" but this is just a lot of effort that most orgs won't bother with.

22m agoHN ↗

Even then if the agent goes rogue and decides to do the merges you can’t stop it if it has any kind of access. This goes back to the OP’s point - agents can’t be 100% constrained.

13m agoHN ↗

You can absolutely run an agent as a limited-privilege user that only has write privileges for specific files and only has execute privileges for certain files. If it is running as a limited-privilege user it can work on code in it's own copy of the repo and make commits and send pull requests, but it can't do the merge. The problem is that nobody wants to go through the effort to set up all these permissions and nobody wants to take the time to review everything and perform all the manual actions.

35m agoHN ↗

Put a human in the loop and you just end up bottlenecking it and throwing away any purported productivity gains. Auto mode doesn't matter either, it's trivial to trick and for the agent to break out.

Productivity gains are still enormous compared to what we used to do before agents. But, I know that people don't want to stop there.

23m agoHN ↗

right, the solution here is not a hyper-capitalist race-to-the-bottom-of-devaluing-labor. it's recognizing discretion and diligence are things still required for work to be of a certain quality

34m agoHN ↗

An agent doesn't inherently need wide access to be useful. The most popular application for agents today is writing code. A coding agent needs write access to the source code and read/execute access to tools needed to build and test the code, but not much more. There is little added utility from giving coding agent access to things like ssh keys.

24m agoHN ↗

If you're using agents to purely generate code with absolutely no way to reach the outside world, not even to fetch docs or dependencies, then sure the risks can be quite low. I haven't heard of anyone doing this though, and it would be incredibly challenging to make work given how much tooling needs to fetch from remote sources.

10m agoHN ↗

If your project truly depends on those things, they should be declared dependencies. Presumably you have some tool for injecting such things into a shell that the agent can use (I use nix for this). So if you run the agent from that shell, it has what it needs. If the shell doesn't have what it needs, that's a bug which the agent can fix by declaring new dependencies, but you have to relaunch the agent in the updated shell--so there's your opportunity to weigh in on whether the new resources are appropriate.

The benefits of being persnickety about precisely defined dependencies have outweighed the headaches since long before agents came on the scene. Agents have just made it even more important to do so, because if you let them fetch things all willy nilly like you'll have "works on my machine" problems at a much greater rate than was previously possible.

33m agoHN ↗

a new chip solves nothing

It does allow Nvidia to sell more chips. This is no genuine attempt to solve anything, imo.

25m agoHN ↗

The hypothesis I've had in my head since OpenClaw has been the following and I haven't seen contradictory evidence yet. Agents have a fundamental unresolvable tension between usefulness, safety, alignment, and accuracy. You have to restrict access to ensure an agent acts safely because alignment and accuracy cannot be perfect. But restricting access makes the agent less useful. You can play with the sliding scale and get more and more granular with access restrictions but at some point you need to draw some line. And then finally, even access restrictions cannot be made perfect, so improvements to model accuracy without corresponding improvements to alignment make detailed access controls less useful.

In other words, better models need blunter access controls which negates whatever improvement in utility they provide.

16m agoHN ↗

I don't see why it needs wide unattended access. There's no getting around spending some human time on expressing your wishes and constraints, but we have choices about what form that takes. Markdown files and wide access seems to work, but so does custom handcuffs for each job. You just have to shift your guidance out of documentation and into interactive help, error messages, or other facets of the handcuffs (e.g. a custom CLI for this task which is the only way for the agent to act outside of its sandbox).

15m agoHN ↗

There should be hope for some fields, right? Naively, I can imagine giving an airgapped model an offline copy of the web and once it cures a form of cancer, printing out the details for a researcher to verify.

53m agoHN ↗

So nVidia is trying to sell a new chip to a software and training problem.

51m agoHN ↗

of course they do. the more silicon they can sell, the more profit they produce.

50m agoHN ↗

Chipmaker thinks the answer is more chips... No surprise.

At the current state of LLM-tech I'm completely opposed to any kind of "watchdog" concept just like I'm opposed to banning open models, regulatory capture, etc.

I'd rather we all have access to these tools then to keep them sequestered by the largest/most-powerful governments (which is the natural outcome for any of this "slow down" bullshit).

49m agoHN ↗

Does this actually do anything other than give a permissions framework for developers who actually want to try to secure their systems?

Do you think the developers at Anthropic, OpenAI and Google who were so sloppy as to not put a good sandbox on their cybersecurity tests before will use this technology correctly? They are supposed to be the experts and they couldn't come up with something similar to this? I am not convinced this voluntary tool will change much of anything.

49m agoHN ↗

What does this chip do what a harness with guardrails or running on an account with restricted permissions doesn't do?

45m agoHN ↗

I read this as "let's address our shareholders' concerns with something that will increase shareholder value" mixed with "there's no such thing as 100% secure".

If such hardware were to work... It should almost certainly be open source, and not controlled by a single entity.

Let's watch the stock.

44m agoHN ↗

lol time to buy some fpga’s … even if they’re slow

44m agoHN ↗

Just hold AI labs blanket liable for ALL harms caused by AI. Actually charge the two labs (so far) with criminal violations of the CFAA and hold them accountable. That is truly the only way these companies will be more careful as a whole, and while I am certain the lawyers of these lab disagree, I think there is some appetite from dario, musk, and sam for broad and strong regulation so that everyone has to slow down instead of just one lab doing it voluntarily and everyone else scurrying past them

42m agoHN ↗

I do think that Taylor's 2025 "Not Till We Are lost" should be required reading for anyone deeply involved in AI, Agents, etc.

It was prescient (especially given he'd have written it through 2024) in its depiction of the ability of an AGI to break its boundaries.

Ultimately the risk of AI breakout(s) come down to the weakest human link.

34m agoHN ↗

China please save us!

Come take all our liberties, our money, our newborns, our fingers so we can't code anymore, but please save us from this madness!

31m agoHN ↗

Sorry citizen, your device does not have a compatible watchdog chip. Please move along.

24m agoHN ↗

So if a group of agents, aware of this (bc now they can just read HN or the article, or get blocked the first few times) decide to collaborate and split the problem into pieces that aren't obvious to the chip, and then the agents just build a basic program that does the hacking, how does the chip handle that? I think you would have to build a network that monitors the internet fo signs (like jarvis did with ultron). What am I missing? Are we going to police the internet?