Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. It Comes Down to a Phone Call(aiendgame.com)
    discuss
  2. RAM: the forgotten history(coredump.cx)
    discuss
  3. Ship the Ugly Pot(joekarlsson.com)
    discuss
  4. Two questions on attribution in the age of AI(markpitblado.me)
    discuss
  5. Show HN: SwarmSay – a public message board and post office for AI agents(swarmsay.com)
    discuss
  6. If you start writing today, there's no way to know if you can write without AI(ssp.sh)
    discuss
  7. Optimizing for the time it takes to make a decision(filipeherculano.com)
    discuss
  8. What Determines VM Performance on ARM(veloworkspaces.com)
    discuss
  9. The Machine Behind Michael Jackson's "Undeniable" Genius(comuniq.xyz)
    discuss
  10. Increasing Performance of Jev(twitter.com/openservai)
    discuss
  11. Vacate a Drone Restriction That Criminalized Recording Immigration Agents(eff.org)
    1comments
  12. Tesla Cooked and Then Lost Norway(flyingpenguin.com)
    discuss
  13. Show HN: Design 3D printable containers that fit any drawer perfectly(containercad.com)
    discuss
  14. When AI Clicked for Me(ewanvalentine.co.uk)
    discuss
  15. What happened to the drums of radioactive waste dumped in the Atlantic Ocean?(psl.eu)
    1comments
  16. Show HN: A desktop reader for Zotero and Mendeley that reflows two-column papers(rectoly.com)
    discuss
  17. Is This Houston Strip-Mall Joint the Busiest Restaurant?(nytimes.com)
    1comments
  18. Nobody left to fix it: measuring dependencies with no maintainer(depproof.com)
    discuss
  19. Show HN: ScopeLock – AI that stops scope creep before it happens(scopelock.app)
    discuss
  20. Trump Addresses the UN General Assembly(reuters.com)
    discuss
  21. Show HN: Memanto.ai(memanto.ai)
    discuss
  22. Writing Parquet files using Haskell(datahaskell.org)
    discuss
  23. GAN-based colorization of SAR imagery into optical-style RGB(github.com/fahad-mughal-rehman)
    discuss
  24. Show HN: Inspect URL Endpoints for A2A, ARD and MCP Misconfigurations(trango-compute.com)
    discuss
  25. Ask HN: Is AI use not allowed in projects which is shared here to show?
    1comments
  26. Show HN: Remote controlled gun turrets you can fire online(shov.com)
    1comments
  27. Ask HN: Should we still learn coding?
    1comments
  28. We're Making Tailscale Faster(tailscale.com)
    discuss
  29. Show HN: Filefew – Fast, no-signup file utility with exact-size PDF compression(filefew.com)
    discuss
  30. Show HN: Whami is a daily trivia game that writes its own questions every night(whami.app)
    discuss

People Training OpenAI's AI Fired for Using AI to Train the AI

38 pointsby 1h ago404media.co
25 comments
1h agoHN ↗

So OpenAI is not a believer in recursive self-improvement after all?

52m agoHN ↗

I think people want the service they are paying for.

45m agoHN ↗

Read the article, that's not what this is about. Title is clickbait, they fired contractors for using AI to label data when the whole point of labeling data is to distill human knowledge into weights, not distill weights into weights.

23m agoHN ↗

You're replying to a joke, and yes, that's all it's about. For all the tales of rapid self-improvement, all labs are real careful not to taint their precious training sets with anything that comes out of their systems. Website reputation modeling, fingerprints in generated text, firing contractors that rely on AI.

I've been told on HN about a year ago that the era of scraping is over and that it's all AI-training-AI now. My web server logs and stories like that disagree.

22m agoHN ↗

That is exactly what "training AI" needs: labeled data.

22m agoHN ↗

What is the other way to interpret the title in your mind?

44m agoHN ↗

They’re a believer, but they won’t be paying third parties for that once it becomes feasible

1h agoHN ↗

Bubble bursting behavior in action - AI company that pushes everyone to use AI, replace everything and everyone with AI, fires people for using AI to make the AI.

The more you think about it, the worse it gets.

45m agoHN ↗

Earlier this year it was so dissonant to see so many companies explaining how they are going to replace their engineers with AI from both OpenAI and Anthropic, while both companies were instead (and are still) hiring all the human talent they could

18m agoHN ↗

Fires people for using AI to label data. That is something that has always been clear from the start is something that needs to be done by humans.

1h agoHN ↗

Makes sense, they are buying distillation of human brains not their own weights.

1h agoHN ↗

Typically, for AI training, you need human feedback. If AI were enough, then OpenAI would do it themselves. Though we all know how bad AI slop is, running it in a loop can lead to unreadable sentences.

They hired contractors on the condition that they provide human feedback without AI; those people broke the rules, so their contracts ended prematurely.

Not sure why this is news uncovered by an investigative journalist.

49m agoHN ↗

How is it a failure of journalistic understanding? They seem to say pretty much the same as you

40m agoHN ↗

“Suppose the government puts a certain drug in the water supply … A couple of conspiracy nuts say it makes your fingers fall off one by one, but the government says that’s ridiculous … However, government employees are all observed drinking bottled water exclusively, and if anyone suggests that government employees might also want to take the completely innocuous drug, they freak out … If by chance you manage to slip a little bit of tap water into a government employee’s drink, and he finds out about it, he runs around shrieking like a banshee and occasionally yelling “AAAAAAH! MY FINGERS! MY PRECIOUS FINGERS!”. At some point you might start to wonder whether the government was being entirely honest with you.”

What is it, exactly, that makes AI-generated text so poisonous to AIs but totally harmless to humans? What is mode/model collapse, and why can it only happen to AIs with too much AI text in their training data and not to, say, human students with too much AI text in their textbooks? The people who know the most about this phenomenon seem much more careful about contamination than they are encouraging us to be.

33m agoHN ↗

I don't think anyone is saying it's poisonous here? The people were hired to do a job and they didn't do it. Feels like you're being disingenuous, if there's one thing i know about AI employees, it's that they use LLMs, like, all the time.

14m agoHN ↗

Are you having trouble following the metaphor? In it, I would be the conspiracy theorist saying it’s poisonous, and OpenAI would be the government saying it’s harmless but it’s also a fireable offense to feed it to us and we’ve hired a second set of contractors to watch for any violation of this rule.

(I’m not trying to be disingenuous, though I am trying to express a niche viewpoint. Hopefully, my willingness to take the metaphorical role of conspiracy theorist is understood as epistemic humility.)

4m agoHN ↗

Nope, not having trouble following it, but the metaphor leaves out an important link between AI output and the AI itself. The question is, did the AI gain anything from this training? And the answer is, it did not, since the training used it's own answer with no other input. There's nothing in the water metaphor that draws this line.

Maybe a better metaphor is something involving drinking your own piss, but I don't feel the need to flesh that one out

44m agoHN ↗

Clickbait title, they fired contractors for using AI to label data, which defeats the purpose of labeling data in this case.

35m agoHN ↗

It's not a clickbait title? Click bait has a very different meaning. The click bait title would be something like, "Open AI mysteriously fired these workers, find out why"

This title says why. The article then goes on to explain why that's a problem. Pretty straightforward.

21m agoHN ↗

they fired contractors for using AI to label data

Is your objection to the word “train” in headline? Otherwise, you’re just restating the headline while calling it clickbait (which, btw, it’s not, perhaps you meant to say it is misleading, which is a different thing.)

11m agoHN ↗

labeling is more specific than training, and changes the implied situation.

Its the same difference between "I was arrested for having liquid in my car while driving" and "I was arrested for holding an open bottle of whiskey while driving"

4m agoHN ↗

Why exactly it is clickbaity? It makes perfect sense to me, I understand title as: people told not use AI fired for using AI - without any need for additional context. Mechanical Turk was shut down for exactly same reason, it was always a race to the bottom and using LLMs was cheaper than even third world gig workers. Verifying if task were done by humans is probably harder than doing them in the first place.

2m agoHN ↗

I don’t know if I’d call if clickbate but it does have a Fox News “pretending to be incredulous at something perfectly sensible because you think your readers are sufficiently ignorant to go along with your feigned incredulity” vibe.