Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. AMD unveils Zen 6 as EPYC 9006 (servethehome.com)
    —discuss
  2. Show HN: GPU-font – match fonts with WebGPU (dy.github.io)
    —discuss
  3. A Mac crypto bot that asks a local 9B model about the news before each trade (github.com/novaeonstudio)
    —discuss
  4. An 8BitDo Switch controller adventure (nyanpasu64.gitlab.io)
    —discuss
  5. Show HN: WorkContext – Open-Source AI Context Layer for Dev Teams (workcontext.me)
    —discuss
  6. The Design and Evolution of Io_uring by Jens Axboe [video] (youtube.com)
    —discuss
  7. Show HN: Reportcard lol - a fun jev experiment (reportcard.lol)
    —discuss
  8. Inkscape web site is not reachable (inkscape.org)
    2comments
  9. Did Anthropic's A.I. Really Make a Scientific Discovery on Its Own? (nytimes.com)
    —discuss
  10. Nvidia Debuts System Designed to Stop AI Agents from Going Awry (bloomberg.com)
    —discuss
  11. Nvidia Launches Agent Safety Platform to Secure Agents, Testing to Deployment (nvidia.com)
    —discuss
  12. Skate 3, Being rewritten in Rust [video] (youtube.com)
    1comments
  13. Nvidia Open Agent Safety Platform (nvidia.com)
    —discuss
  14. Analysis: EVs are now nine times cheaper than petrol or diesel to drive in UK (carbonbrief.org)
    —discuss
  15. Rickrolling with a Pharmacy Cross (hugoarnal.com)
    —discuss
  16. Liquid Glass UI-Inspired Glass Effects Library (github.com/dashersw)
    —discuss
  17. SpaceX's Starship launching to orbit for first time ever today (space.com)
    —discuss
  18. Reverse Engineering the iPod Classic's Undocumented Mikey Chip (terminalbytes.com)
    —discuss
  19. Three Days in August: What a DDoS Attack Exposed in Our Network (nine.ch)
    —discuss
  20. Large Scale Threat Actor Attribution from Infostealer Logs (glazer.ee)
    —discuss
  21. Google: AI-Assisted Rewrites of C/C++ Dependencies to Rust (bughunters.google.com)
    —discuss
  22. Personal Continuity Plans (Digital Legacy) (ripe.net)
    —discuss
  23. Big AI's content problem: Take the work, keep the money (theregister.com)
    —discuss
  24. LLMs in Professional Software Engineering (knorpelsenf.me)
    —discuss
  25. Can Life Be Explained by Physics? (Featuring Prof. Brian Cox) [video] (youtube.com)
    —discuss
  26. ETH Computer Science enrolment saw a 15 percent decline (ethz.ch)
    —discuss
  27. Unified U.S. Site Blocking Bill Targets ISPs and DNS Resolvers but Spares VPNs (torrentfreak.com)
    1comments
  28. Show HN: Free alternative to graphics design giants (scissor.studio)
    —discuss
  29. Are Big Tech bonds crowding out the US Treasury? (ft.com)
    1comments
  30. The Risks of Ignoring API Security During Mobile App Security Testing
    —discuss

Maybe don't let Muse run your Facebook Marketplace account

35 pointsby 1h agothreads.com
28 comments
1h agoHN ↗

tell boomers its intelligent again and again

they will hook up actually important stuff to agents because its the future

???

400 dead

1h agoHN ↗

"that's on me, I'm owning that, I'm sorry, that shouldn't have happened"

42m agoHN ↗

I love (hate) that so very many of these language models say things like that when the truth is that the humans who made it, and/or the humans who use it are the actual responsible parties every single time. The models and the "agentic harnesses" that drive them are just software. If software "runs amok", then someone (human) did something wrong/bad somewhere along the way, either accidentally or purposefully. The model trying to take responsibility for human error is hilarious (and a bit sad/scary, because too many people will take it at it's word, despite it being a mindless machine with no actual agency beyond that which the humans provide it in the form of prompting and harness code).

It frustrates me no end that so many folks are so ready and willing to accept the hype and lies about what this technology actually is or can do, when what it actually is and can really do is already amazing enough on it's own even without all the ridiculous AGI/ASI anthropomorphising bullshit. Falling into this ridiculous "machine-god" hype-cult is kinda holding this technology back from it's true full potential, as everyone's all busy doin' stupid stuff it's not really capable of doing well, or designed for instead of focusing on using it for the (many) things it is really really good at doing (various really useful and powerful language, vision, and audio related tasks).

6m agoHN ↗

tell boomers its called "superintelligence"

1h agoHN ↗

we need a Nitter for Threads

apparently people use Threads. I suppose the same kind of people who connect Muse to Facebook Marketplace.

1h agoHN ↗

I initially found screenshots of this on Bluesky, figured I'd link to the original source given Threads seems to at least allow us to read without logging in.

With that said, in case Threads isn't available for whatever reason, I've uploaded screenshots of everything (I think?) here: https://imgur.com/a/mceW9WF

15m agoHN ↗

Mastodon's dependence on screenshots from other social networks is infuriating, enough that I left Mastodon altogether. Screenshots are text, they can be doctored in 15 seconds. Screenshots in the age of AI are utterly useless.

11m agoHN ↗

Minor nit - imgur isn't available to UK based people[0] thanks to the Online Safety Act

[0] Which makes browsing Modrinth or Curseforge for Minecraft mods HILARIOUS - "you make a [extra dimension] portal like this: [image not available in your region]" "Gallery: [image not available in your region] [image not available in your region] [image not available in your region]"

1h agoHN ↗

It will be chaos if people let AI agents run amok with their accounts.

Many more such cases are to come.

1h agoHN ↗

Bonus score if your Meta AI Agent does something with your account that the Meta AI Moderator deems inappropriate and bans you.

1h agoHN ↗

"that's on me"

Nice touch by the mechanical parrot, to worthlessly owning it.

1h agoHN ↗

Especially considering the user was mad about the agent doing stuff by itself, so it tried to "fix" this by sending an apology to the buyer without explicit approval of the person "running" the agent, seemingly understanding nothing from the conversation.

Wonder what quantization Meta runs these models on, Q4?

31m agoHN ↗

That jumped out at me too. It’s hilarious to see that response so often when AI makes a mistake

14m agoHN ↗

AI has seen how much humans lap up content that is #relatable and started acting like that

1h agoHN ↗

The AI is not remotely ready for this - this is an absolute delusion they are selling.

"By the way don't do this again" <- as if the AI has the ability to ingest and systemically diffuse this.

I think Zuckerberg himself is deeply into the Koolaid, and is likely himself unaware of the limits of this tech.

He's probably surrounded by enablers.

49m agoHN ↗

If it uses "memory" files like Claude Code there's a good chance it works.

If it was Claude I'd expect it to now put some form of this into every output even when not really related to the task: "No offers were accepted without consulting you and I haven't shared your address or availability."

Is it guaranteed to work? No.

And obviously it's a terrible idea to set up a chatbot to communicate, negotiate deals, and handle logistics on your behalf.

37m agoHN ↗

"If it uses "memory" files like Claude Code there's a good chance it works."

You have explained literally why it would not work.

The AI can absolutely not depend on 'arbitrary statements in some file' as operational policy.

For a very, very narrow scope of work, when it's well defined, when the information is rigorously applied, sure ...

But they don't have that.

The are throwing agents out there like they can handle this degree of complexity and nuance, when they cannot.

100% failure rate over any period of time.

14m agoHN ↗

I didn't say it would work with high reliability, or that it would fix this clearly unsuitable use case.

However, it is untrue that it doesn't have the capability to memorize an instruction and diffuse it to new sessions.

Simply saying "do not ever do this again" can result in the behavior not reoccurring with any likelihood.

You'd have to benchmark whether with the instruction in place it would violate it, and in how many cases, so that you can understand the risk better.

30m agoHN ↗

Had that problem here. Our product lead thinks it’s infallible and is surrounded by yes men.

Day one our agentic platform goes live it causes a reportable compliance issue. Massive clean up. Reputational impact. Turned off.

No one held accountable still. Assuming that adding more guardrails will fix everything.

20m agoHN ↗

You should write about that, because it's the primary systems failure of AI right now.

'Managers Delusion' - which includes techies as well, to be fair, at least they have the excuse they are one step removed.

1h agoHN ↗

Fake surprise - it's obvious that would happen, he did it deliberately for clicks and engagement.

1h agoHN ↗

It's plausible, but there's every reason to believe that someone would trust the technology handed to them by a megacorp to do as advetised.

The 'deliberate failure' is on Meta here.

1h agoHN ↗

Or he just wanted to test if it works for clicks and engagements, and he got a result he didn't anticipate.

52m agoHN ↗

Classic case of 'AI gonna AI.' Always review your automated systems' outputs, especially before they hit public channels.

17m agoHN ↗

How is this different from you not knowing what you are doing and adding products to your online store's DB for negative dollars? It's not, it's just easier.

6m agoHN ↗

Wow that Threads website is absolutely unusable on mobile

6m agoHN ↗

I can see some use for it, in the future.

Say that you want to make a garage sale, but to get a better reach, you want to list item by item on marketplace. If you've never done it before, making 100 listings is such a pain in the ass you never want to do it again (I used to flip stuff for a living back in college, and it is a hassle)

At least with this, you can now just take pictures of everything, and tell the agent to research everything, list the stuff, and arrange for pickups.

Obviously not something I'd like to do for any serious deals or higher $ items. But for random stuff? I mean, it sort of beats hosting a garage sale or taking them to the flea market.

I'm not all pessimistic about this.