Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. U.S. Media Freedom: Perceptions Hit New Low (gallup.com)
    —discuss
  2. Learning How to Forget: Fine-Tuning for Long-Context Sparse Attention (arxiv.org)
    —discuss
  3. OpenAI Answers TypeSafe's Jev with a Decision API Built on Luna (thenewstack.io)
    —discuss
  4. ChatGPT Pro 500 (help.openai.com)
    —discuss
  5. Performance Improvements in JDK 27 (inside.java)
    —discuss
  6. Show HN: Greek and Cypriot Greek TTS that runs faster than real time on CPU (huggingface.co)
    —discuss
  7. The worst problem with Claude and ChatGPT is how we communicate with them (caseorganic.substack.com)
    —discuss
  8. Infinite Jungle (mo.town)
    —discuss
  9. GPT-6 UltraFast (developers.openai.com)
    —discuss
  10. AI Models Fail at Physics: Coding Harnesses Are to Blame (juliahub.com)
    —discuss
  11. Isotropic Superconductivity in Room-Temperature Superconductor LaSc2H24 (arxiv.org)
    —discuss
  12. New Features. Insight Frequency and Run the Numbers (ichingportal.com)
    —discuss
  13. The Next 3x in Inference Won't Come from Faster Kernels (twitter.com/usemuna)
    1comments
  14. Show HN: A working 3D model of an Enigma machine (enigma.design)
    —discuss
  15. Is there a lanscaping tool which can take photos of my garden and make 3D model?
    —discuss
  16. What Was Hidden Under Double D Hat? (noxluneworld.com)
    —discuss
  17. WebMCP: Stop scraping the Web, start calling it [video] (youtube.com)
    —discuss
  18. Show HN: Orderbook – a matching engine in Go, tested against a second one (github.com/intrepidkarthi)
    —discuss
  19. Tcl/Tk 9.1 Released (tcl-lang.org)
    —discuss
  20. Force Yourself to Struggle Even If Using AI Is Easier (nathanlangley.dev)
    —discuss
  21. Dontkyc.me – directory of no-KYC services, tracked incidents and a trust score (dontkyc.me)
    —discuss
  22. IRGC writes an open letter to American people (mehrnews.com)
    —discuss
  23. Show HN: ChooseBrowser – route Mac links to a Chrome profile by URL rules (leeguoo.com)
    1comments
  24. Show HN: Iiveg.app – Find vegan food, restaurants and cosmetics around the world (iiveg.app)
    —discuss
  25. Hofstadter's Tangled Hierarchy with Video Feedback (youtube.com)
    —discuss
  26. Our Year of Saturday Dinners (amandalitman.substack.com)
    —discuss
  27. Apple patches CoreGraphics zero-day already exploited in targeted attacks (theregister.com)
    —discuss
  28. Kompyla – a self-hosted research monitor that builds a cited wiki (github.com/damien220)
    —discuss
  29. The Texas courtroom where toddlers face immigration judges (cnn.com)
    —discuss
  30. Dots (openai.com)
    45comments

GPT 6.1 Sol

158 pointsby 23m agoopenai.com
77 comments
20m agoHN ↗

Weren't there headlines just yesterday that they weren't releasing this due to safety concerns?

19m agoHN ↗

GPT-6.1 Astra is what those headlines referred to. This is GPT-6.1 Sol.

18m agoHN ↗

That was 6.1 Astra. And I'm assuming it's being tabled because it still doesn't match Opus 5.5.

This is a decent win though, if it really is better. 6-sol was really no good, at least in my work.

17m agoHN ↗

6 Sol was worse than 5.6 Sol from my own experiences. Far worse.

Will see if this remedies things.

14m agoHN ↗

Yup, same experience here. I used it for one day, spent the next day fixing its lousy code, then went back to 5.6.

18m agoHN ↗

Shots fired, half the price of Opus 5.5.

18m agoHN ↗

Wasn't 6 released like last week? I can't keep up anymore.

15m agoHN ↗

Yes but it was underwhelming, so they seem to have rushed 6.1 Sol out. Also Opus 5.5 may have spooked them too.

15m agoHN ↗

Do you need to? Do you always keep up with all the version bumps on the software you use?

11m agoHN ↗

It was so underwhelming that it didn't even make it to chatgpt chat interface

18m agoHN ↗

Ominous for the industry and investors that token price is becoming the main battleground. Could be Anthropic's rationale for IPOing this year.

16m agoHN ↗

Great for the consumer.

I remember when bandwidth was super expensive and now it’s dirt cheap.

8m agoHN ↗

China will do to llms what they did to german cars

3m agoHN ↗

Why make a new account to post this comment?

It's not even anything controversial..

4m agoHN ↗

Another piece of evidence on the pile that the sudden panic and desire to "slow down" is because they're hitting the plateau on capability

Which, honestly, is fine. A lot of juice to squeeze in efficiency and even if models got zero more capable, making the capability that is already here cheaper is a huge win for everyone (except Nvidia)

18m agoHN ↗

If these models are so smart, can't _they_ select the right model for each task?

16m agoHN ↗

the right model for the task is the one that transfers the maximum amount of USD from your pocket to the provider's bank account.

13m agoHN ↗

Why don't you simply ask the respective model which model is best for a specific task? :-)

5m agoHN ↗

Switching models is _very_ expensive in compute (you have to rerun everything from the beginning), and highly variable in cost. Cursor tried doing this for awhile, but inconsistent performance/usage means most users turned it off and pick models specifically.

5m agoHN ↗

I guess the question is, does the Dunning Krueger effect apply to models? The dumb ones might think they're up to the task.

18m agoHN ↗

What's driving the increase in release cadence here? We seem to get new models every week or so now, is this RSI?

13m agoHN ↗

Wanting to have the newer model than the competitor, presumably.

3m agoHN ↗

The old "the bigger number is better", GPT announces model 6.1, the obvious thing to do is to announce Gemini 27, and after that Claude 3000, then a flute album.

12m agoHN ↗

Response to DeepSeek’s technical paper and competition.

12m agoHN ↗

No, we're pacing ourselves to have the time to evaluate the impact each new model could have, obviously.

11m agoHN ↗

Chinese model pressure. Many of my SWE friends switched to Chinese models. I also use QWEN and GLM for many of the api requiring projects and dropped OpenAI and Anthropic. The only reason was the cost.

10m agoHN ↗

Its a news cycle more than anything, and its ONLY going to get much, much worse. Daily releases, or multiple daily, 30-45, by EOY. Welcome to RSI!

8m agoHN ↗

What I don't understand is how much people have to say about every single one. Aren't we at the diminishing returns stage yet? Is there really that much to discuss?

7m agoHN ↗

I do wonder if people switch back and forth between primary models (GPTvsClaude) that it may be a better idea to simply keep releasing updates as soon as possible in order to keep users from bouncing back and forth.

4m agoHN ↗

New models are distill from the actual unrelease frontier models. They are just giving us better checkpoints.

3m agoHN ↗

The initial response to 6 Sol was bad, and Opus 5.5 was definitely winning the public vibes war. Makes sense to rush something out

18m agoHN ↗

Aka "we made an oopsie last week and released what should have been called GPT 6 Terra with the name GPT 6 Sol"

17m agoHN ↗

GPT‑6.1 Sol matches GPT‑6 Astra at roughly one-fifth of the cost

Astra is a pretty impressive model. Excited to try this.

17m agoHN ↗

I love free market competition. We're getting insane advancements every day. I remember when llms used to cost an arm and a leg for decent intelligence

16m agoHN ↗

This is great. But maybe part of the motivation is that 6-Sol wasn't as good as initially advertised so they needed to tweak it. I felt a clear degradation in quality in some simple refactoring tasks vs 5.6-Sol.

16m agoHN ↗

I wish they'd list the environmental cost. My employer has an unlimited AI budget so I don't care about using Astra if it's just more profit for OpenAI. I care more if it actually uses 5x more energy.

11m agoHN ↗

I don't understand the point of this, why just now when it comes to llms. Why wasn't anyone enraged with the environmental costs of kids playing video games. I would not be surprised the environmental cost of that is an order of magnitude bigger than what llms have.

8m agoHN ↗

Given that the number one cost of inference is memory and compute, and the incremental cost of each is energy, cost per inference is roughly proportional to energy consumption.

16m agoHN ↗

Okay, now price cut 6 Sol (and rename it to Terra again).

15m agoHN ↗

GPT-6 Sol released a week ago. Shortest model life ever?

12m agoHN ↗

Taking GPT-6 "Sol" outside behind the shed and giving it a merciful end is about the best outcome possible.

Huge misstep releasing it.

8m agoHN ↗

Sorry, I'm GenX. Growing up they showed us "Old Yeller" in the school gym every year like that was some kind of treat.

15m agoHN ↗

I wonder if releasing this soon sort of validates the rumor that Sol 6 was just the Terra model they bumped up and slashed the price.

Then Opus 5.5 caught them off guard and now they're actually releasing the correct sized model.

15m agoHN ↗

Looking at the token prices, if this is half as good as 6-Astra for 3D model creation in Blender, it's going to be an absolute game changer.

Opus 5.5 is definitely better at coding, but nothing even comes close to 6-Astra for work in 3D graphics...

10m agoHN ↗

How is it with animations?

I have played around a little bit with fixing some rigging problems and was impressed, but Opus even warned me it was bad at animations cause it can only really grab screenshots to process static content.

13m agoHN ↗

Cached input costs just $0.10 per million tokens—95% less than standard input pricing and 50% less than GPT‑6 Sol’s cached input pricing

This is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.

8m agoHN ↗

Exactly half as expensive as Opus 5.5 in every API pricing metric

13m agoHN ↗

So yesterday we were consumed with how this was being delayed because of safety, yada yada.

Guess not?

9m agoHN ↗

That model was implied to be GPT 6.1 Astra, not Sol.

8m agoHN ↗

I see where you are coming from. But 6.1 Sol seems like a new frontier in pricing, not intelligence. I do think the deceleration stuff was mostly bluster, but I don't think this release in particular contradicts it too much.

13m agoHN ↗

I can blow through my weekly on astra in a few hours; hopefully this really is as good.

12m agoHN ↗

Cache is priced at $0.1/M, 50% as sol 6 and sonnet 5.5.

12m agoHN ↗

$2/10 is pretty cheap for a frontier model...

9m agoHN ↗

I stopped using LLMs. I shit you not. My life got better.

9m agoHN ↗

So when does Anthropic answer? Tomorrow?

8m agoHN ↗

Why they are not even benchmark model against Anthropic or anybody ?

8m agoHN ↗

GPT 6.0 Sol was so terrible—I wonder if 6.1 Sol will be good?

5m agoHN ↗

Let's all boycott and move to Claude until they release 6.1 Astra. I don't like to be teased.

When is the alleged "safety" concern satisfied? Does this mean releasing new capability to consumers is going to get a lot slower? Lower price for 6 Astra capability via this 6.1 Sol is exciting, but that is because of Astra capability not merely the low price point.

When do we get the next jump in capability? When is 6.1 Astra released?

5m agoHN ↗

The real announcement is the ultra fast mode ... Astra at 300t/s is insane!

4m agoHN ↗

Are you guys in dev day? did they start?

3m agoHN ↗

didn't 6 sol just come out a couple weeks ago?