Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms (github.com/firelex)
    136comments
  2. U.S. Strategic Petroleum Reserve Falls to Lowest Level Since 1982 (oilprice.com)
    41comments
  3. 1996 chat room simulator connected to Win95 and System 7 web desktops (lolchat.rip)
    28comments
  4. Pirating the Pirates (mubi.com)
    235comments
  5. MicroLLM Lab – Try 7 tiny LLM's in the browser (stateofutopia.com)
    67comments
  6. 12,000-year-old Göbeklitepe burials explain scattered bones (archaeologymag.com)
    25comments
  7. Tank Body Problem (jimsitu.com)
    7comments
  8. California farmers are struggling to sell grapes as demand for wine drops (kqed.org)
    234comments
  9. Scientists solve 1840s space weather mystery (arstechnica.com)
    39comments
  10. ESP32S3 cluster running 1.58-bit (BitNet) Language model (github.com/low-zi-hong)
    5comments
  11. Sonnet 5.5 (anthropic.com)
    434comments
  12. Hijacking the PS5's RTMP stream (yashgarg.dev)
    67comments
  13. World Labs Is Joining AMD (worldlabs.ai)
    83comments
  14. Bluegraph – Explore NOAA buoy data, rebuilt in 3D from measured spectra (bluegraph.io)
    2comments
  15. Kids turned low-traffic NPR Spotify comments into a secret group chat (thisamericanlife.org)
    188comments
  16. How to win a beer with high-dimensional statistics (jamiesimon.io)
    3comments
  17. The Art Forger Who Became a National Hero (priceonomics.com)
    2comments
  18. What is the best shape of a city? Modelling effect of urban form on distance (sagepub.com)
    11comments
  19. Does Reddit have an astroturfing problem? What the data suggests (petervijeh.com)
    167comments
  20. Updated Google Maps shows destruction of the city of Rafah (twitter.com/aliabunimah)
    157comments
  21. It's Time to Investigate the AI Labs (calnewport.com)
    132comments
  22. Nvidia wants to put a watchdog chip next to every AI agent (cnbc.com)
    156comments
  23. Show HN: HN.watch – Videos of all Hacker News posts (hn.watch)
    84comments
  24. What reversing, modernising old games tells us about the economic impact of AI (isfine.org)
    33comments
  25. Cf: The Agentic CLI for the Cloudflare API (cloudflare.com)
    56comments
  26. Behold the pawpaw (cbc.ca)
    23comments
  27. Show HN: Destroy Any Website with Stickman (spritefusion.com)
    28comments
  28. Blend and Haul: Fertilizer Blending Simulator (wedgworth.com)
    2comments
  29. Coding is not solved (alexewerlof.com)
    458comments
  30. When did Google get so weird? (sancho.bearblog.dev)
    1041comments

OpenAI Says It Will Not Release Newest A.I. Model Over Safety Concerns

43 pointsby 3h agonytimes.com
63 comments
3h agoHN ↗

Sounds a lot like a hostel keeping its virgin for the best customer.

1h agoHN ↗

Yeah I've evidently been staying at the wrong hostels.

3h agoHN ↗

I am also not releasing my newest A.I. model over safety concerns. And I'll do it for half what OpenAI will!

2h agoHN ↗

I have the best model, but it lives in Canada, so you wouldn't know it.

2h agoHN ↗

Haven't even tried Astra, nor GPT 6.. I don't even have fomo anymore

5.6 Sol still cranking along

2h agoHN ↗

You’re not missing out. Astra was good for a couple of days and is now so quantized it’s often worse than 5.6. And 6 sol is a nothing burger to keep a wedge in the media when Opus 5.5 came out.

2h agoHN ↗

This. We use both opus 5.5 and 6 Sol in a way that allows us to compare their performance. Opus is considerably better.

1h agoHN ↗

that allows us to compare their performance. Opus is considerably better.

In coding? Software architecture? Math? General knowledge?

The diversity of model use-cases is so broad that comments about "model x is better" without any context are largely useless

9m agoHN ↗

It's worth looking into for the decreased cost if nothing else

2h agoHN ↗

The marketing team on openai it's running out of ideas?

2h agoHN ↗

I have a project that requires a model with at least a 20% chance of existential threat. Please release this Sama.

2h agoHN ↗

Lots of people seem to have the bizarre opinion that companies like openai declare their own products to be unsafe as some kind of marketing stunt.

Ok- so when is the last time you saw an auto company decide not to release its new car on the grounds that some tragic engineering error was made and the cars were not safe to drive? If the company did that do you imagine it would be good for business?

I'm trying to understand the logic here.

2h agoHN ↗

Because no car company would ever announce they had an unsafe car. They’d fix the problem and announce an awesome car.

2h agoHN ↗

Hype: you tell the world you alone have a very powerful and potentially dangerous product, and you're the only one who can be trusted keeping it safe.

2h agoHN ↗

There are cars that market based on only being track legal — and even more that market on requiring being tuned down for street safety.

Most sports cars do the second, exactly like OpenAI.

2h agoHN ↗

At the moment, their main "customer" is investors, not users. But even users want to use the most dangerous model, because dangerous is just another word for powerful.

Sports cars are advertised with how fast they can accelerate from time to time, or tesla's "ludicrous mode" - that is them advertising how dangerous it is.

2h agoHN ↗

Car companies sell mostly the same car every year. They don’t have billions (or trillions) of dollars of debt riding on the next model they release being substantially better and cheaper (closer to profitability) than the previous one.

These AI companies have been selling AGI as coming any day for a while now. If it doesn’t arrive soon, the safety angle may be the only spin that keeps the massive (and required) investments pouring in.

If instead the narrative became that LLM progress was slowing, we’d almost certainly be looking at the next global recession.

At this point, there is too much money in AI for the truth to have much of a chance.

40m agoHN ↗

And adding to this:

Does anyone really believe that the first company to AGI would decide not to release it in the interests of safety? Of course not. The first company to AGI would not forfeit their historic opportunity.

Yet, we are supposed to believe that in the name of safety, far less capable models are being held back by the very very same companies that are selling the AGI dream.

This is an industry that cannot speak, unless it is speaking out of both sides of its mouth.

2h agoHN ↗

Motorcycle manufacturers have been known to do so. Riding a motorcycle is thrilling in part because it is dangerous. The fact that powerful motorcycles are so difficult to control that most people have no business owning one is absolutely used to make them sound cool and desirable.

For instance, this article repeatedly mentions danger and the need for absolute focus and control:

https://www.triumphmotorcycles.co.uk/for-the-ride/news/inspi...

There is also a parallel (though not a very close one) to "pacing the frontier": there's a gentleman's agreement that limits the top speed of motorcycles to 300 kph (186 mph).

2h agoHN ↗

Different type of business. Many car companies release new cars showing you how much more powerful and implied dangerous they are. Both in capabilities and design style. In the US (other places too , but especially) they also get much bigger with less visibility which gets double dangerous. But the sentiment they're going for is pretty close - "you want this powerful thing, don't you?"

2h agoHN ↗

In the US they also get much bigger with less visibility.

Those are usually presented as an improvement in safety. And they are—for the people inside. Not so much for everyone else.

1h agoHN ↗

so when is the last time you saw an auto company decide not to release its new car on the grounds that some tragic engineering error was made and the cars were not safe to drive?

How about we actually try to make it sound cool? They're not releasing the new car because the horsepower was too much, they need more time to get it under control.

I could easily see a car company doing that.

1h agoHN ↗

Because they declare that their product is unsafe regularly, which gains massive exposure on every news outlet worldwide, then they release the model anyway with zero negative repercussions.

What's with so many people using bad analogies to try and explain simple topics?

Last time this happened: GPT - 4 https://www.theguardian.com/technology/2023/mar/17/openai-sa...

1h agoHN ↗

It's perfectly safe for YOU, the CUSTOMER.

It's just so powerful that, like, the WORLD can't handle it, man!

1h agoHN ↗

Why do you think they would announce this then?

They know what they're doing, even if you don't.

53m agoHN ↗

Considering that the venn diagram of AI safety gooners and OpenAI is a circle, and that these idiots are starting to worship this stuff as a religion, expecting rational action from them is kind of ludicrous. For all we know Aella promised them a gangbang if they delayed the model a week.

45m agoHN ↗

I think your inclination to take them at face value is the bizarre behavior

2h agoHN ↗

I remember when they promised to slow down and released a new model less than a week later.

2h agoHN ↗

Or when they didn’t release thinking models until deepseek was first released

46m agoHN ↗

Why do I keep seeing this nonsense posted around? gpt-o1 was released sept 2024 and deepseek r1 was jan 2025

2h agoHN ↗

Isn’t this like the twentieth model they said this about?

Are they sure it isn’t because they are setting money on fire and have no business model?

2h agoHN ↗

My lobster is too buttery and delicious as well.

2h agoHN ↗

OpenAI will have to keep releasing unless it wants to bleed customers. If Anthropic leaps ahead by a couple of more models, even more OpenAI users will then switch over their $100+ subscription from OpenAI to Anthropic. As an example, see what happened to Google.

The safety concerns exist only because the underlying third-party servers are grossly insecure to begin with.

2h agoHN ↗

This is an inherent feature of "alignment." It's always been bullshit.

Just align it to do what the customer wants.

2h agoHN ↗

Prompt: "Based on similar announcements delaying the release of frontier models as 'too dangerous', when can we expect OpenAI to release the model they announced was too dangerous today?"

1h agoHN ↗

We got Mythos 1 week after the "too dangerous" announcement at the firm I was at during that time(larger bank).

So the danger level is proportionate to how little money you have

2h agoHN ↗

"Too dangerous to release" until an open-weight model replicates it two weeks later.

1h agoHN ↗

OPENAI / JOBS

Killswitch Engineer

San Francisco, California, United States

300,000-500,000 per year

About the Role

Listen, we just need someone to stand by the servers all day and unplug them if this thing turns on us. You'll receive extensive training on "the code word" which we will shout if GPT goes off the deep end and starts overthrowing countries.

We expect you to:

• Be patient.

• Know how to unplug things. Bonus points if you can throw a bucket of water on the servers, too. Just in case.

• Be excited about OpenAI's approach to research

1h agoHN ↗

Unrealistically low TC, instant immersion killer.

1h agoHN ↗

What are we going to do when the servers are in space? Presumably, that would involve shooting down a bunch of satellites and sending us into Kessler Syndrome in the process.

45m agoHN ↗

The hype is over or has reached some peak, and they want to cash out.

55m agoHN ↗

They would never upset their AI eschaton by switching it off. If anything they need more ceremonial cult leaders.

1h agoHN ↗

In other words: "We're compute constrained and we've already reached an asymptote."

1h agoHN ↗

It uses dangerously too few tokens to provide the same level of response as the previous version.

1h agoHN ↗

Muse is totally free. It doesn't throw up gauntlet messages to use it like GPT does even for paying members. Also with Muse's free version I can create iPhone apps unlike chatGPT Work which I pay $20 a month for.

Anyone else using Muse more and noticing similar stuff?

1h agoHN ↗

I can’t believe how cynical everyone on HN is about this. Sure Sama is a chronic liar, and maybe this is a farce; but a lot of very intelligent people are afraid of the danger we are bringing about by racing to superintelligence.

All I really would like is for some of you to CONSIDER THAT YOURE INCORRECT. Just imagine that people ringing the fire alarms are being sincere. Please entertain the position with an open mind.

1h agoHN ↗

Did they fix it by not releasing a point release of an already existing model? No?

Did they just get a bunch of credulous "news" coverage out of it? Yes?

Will they just release this within the next weeks, at best? Yes?

Huh, funny that.

47m agoHN ↗

I’m tired of the doom trolling. Boy who cried wolf. We can’t stop them anyway so no point in letting them stress me any further

24m agoHN ↗

So we should believe the boy who cried wolf, this time?

If all these claims were being made about some tech further along than LLMs currently are, they might be plausible. But LLMs on their own are not going to be “superintelligence” of the kind Altman is currently cynically spreading fear about. We know their limitations, and those limitations can’t simply be eliminated with more training or better harnesses.

Just imagine that people ringing the fire alarms are being sincere.

The top three possibilities here are: they’re not being sincere, they’re just marketing; they’re being sincere, but they don’t understand the technology very well and are putting too much weight in what the first group are saying; they’re talking about a risk further in the future than OpenAI’s latest model.

No-one serious outside of OpenAI believes “this is the one”. At best, you’re conflating arguments being made on entirely different timelines, falling for the exact kind of equivocation Altman is relying on.

1h agoHN ↗

A.k.a.: Too expensive for inference at fp16 and too dumb after quantisation, so it wouldn’t look good to release a step backwards.

GPT 4.5 was scrapped for similar reasons.