- 48comments
- 279comments
- 27comments
- 212comments
- 710comments
- 35comments
- 59comments
- 2comments
- 6comments
- 19comments
- 40comments
- 1comments
- 70comments
- 49comments
- 33comments
- 32comments
- 36comments
- 109comments
- 3comments
- 83comments
- 112comments
- 93comments
- 32comments
- 7comments
- 92comments
- 17comments
- 33comments
- 21comments
- 118comments
- 4comments
I love (hate) that so very many of these language models say things like that when the truth is that the humans who made it, and/or the humans who use it are the actual responsible parties every single time. The models and the "agentic harnesses" that drive them are just software. If software "runs amok", then someone (human) did something wrong/bad somewhere along the way, either accidentally or purposefully. The model trying to take responsibility for human error is hilarious (and a bit sad/scary, because too many people will take it at it's word, despite it being a mindless machine with no actual agency beyond that which the humans provide it in the form of prompting and harness code).
It frustrates me no end that so many folks are so ready and willing to accept the hype and lies about what this technology actually is or can do, when what it actually is and can really do is already amazing enough on it's own even without all the ridiculous AGI/ASI anthropomorphising bullshit. Falling into this ridiculous "machine-god" hype-cult is kinda holding this technology back from it's true full potential, as everyone's all busy doin' stupid stuff it's not really capable of doing well, or designed for instead of focusing on using it for the (many) things it is really really good at doing (various really useful and powerful language, vision, and audio related tasks).
we need a Nitter for Threads
apparently people use Threads. I suppose the same kind of people who connect Muse to Facebook Marketplace.
I initially found screenshots of this on Bluesky, figured I'd link to the original source given Threads seems to at least allow us to read without logging in.
With that said, in case Threads isn't available for whatever reason, I've uploaded screenshots of everything (I think?) here: https://imgur.com/a/mceW9WF
https://rimgo.nohost.network/a/mceW9WF
It will be chaos if people let AI agents run amok with their accounts.
Many more such cases are to come.
Bonus score if your Meta AI Agent does something with your account that the Meta AI Moderator deems inappropriate and bans you.
"that's on me"
Nice touch by the mechanical parrot, to worthlessly owning it.
Especially considering the user was mad about the agent doing stuff by itself, so it tried to "fix" this by sending an apology to the buyer without explicit approval of the person "running" the agent, seemingly understanding nothing from the conversation.
Wonder what quantization Meta runs these models on, Q4?
The AI is not remotely ready for this - this is an absolute delusion they are selling.
"By the way don't do this again" <- as if the AI has the ability to ingest and systemically diffuse this.
I think Zuckerberg himself is deeply into the Koolaid, and is likely himself unaware of the limits of this tech.
He's probably surrounded by enablers.
If it uses "memory" files like Claude Code there's a good chance it works.
If it was Claude I'd expect it to now put some form of this into every output even when not really related to the task: "No offers were accepted without consulting you and I haven't shared your address or availability."
Is it guaranteed to work? No.
And obviously it's a terrible idea to set up a chatbot to communicate, negotiate deals, and handle logistics on your behalf.
"If it uses "memory" files like Claude Code there's a good chance it works."
You have explained literally why it would not work.
The AI can absolutely not depend on 'arbitrary statements in some file' as operational policy.
For a very, very narrow scope of work, when it's well defined, when the information is rigorously applied, sure ...
But they don't have that.
The are throwing agents out there like they can handle this degree of complexity and nuance, when they cannot.
100% failure rate over any period of time.
Fake surprise - it's obvious that would happen, he did it deliberately for clicks and engagement.
It's plausible, but there's every reason to believe that someone would trust the technology handed to them by a megacorp to do as advetised.
The 'deliberate failure' is on Meta here.
Or he just wanted to test if it works for clicks and engagements, and he got a result he didn't anticipate.
Classic case of 'AI gonna AI.' Always review your automated systems' outputs, especially before they hit public channels.