Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Chinese AI models surge in global popularity – and Washington is worried (cnbc.com)
    —discuss
  2. The Americans Having the Most Sex (theatlantic.com)
    1comments
  3. Show HN: Asụsụ Ọma, a free app for learning Igbo with native-speaker audio (asusuoma.com)
    —discuss
  4. The Cost of Blue Light: Multiorgan Pathology Induced by LED Exposure (biorxiv.org)
    —discuss
  5. OK, so where are all these data centres? (ft.com)
    1comments
  6. What and Whose Knowledge? Measuring Epistemic Diversity in Large Language Models (arxiv.org)
    —discuss
  7. The Aura Premium (twitter.com/netcapgirl)
    —discuss
  8. Show HN: Ollaya Web UI (github.com/samarkandiy)
    —discuss
  9. Software Renderer in Odin from Scratch (marianpekar.com)
    —discuss
  10. Zeta(5) has been proven irrational (also Lean-verified) (reddit.com)
    1comments
  11. Claude Opus 5.5 Should Raise Your Ambitions (thezvi.substack.com)
    —discuss
  12. Beyond Tokenmaxxing (isaca.nl)
    1comments
  13. Screw Is Needed for Missiles and Planes. Buying One Can Take Years (nytimes.com)
    —discuss
  14. A Slice of Europe in the Heart of Lucknow – Oranje Castle
    —discuss
  15. Tuning a Server for Benchmarking (alvarezrosa.com)
    2comments
  16. Is the AI Bubble About to Be Tested? (youtube.com)
    —discuss
  17. Every Package Is Installed (fzakaria.com)
    —discuss
  18. Restricting AI Agents When No Human Is Watching (a16y.ai)
    —discuss
  19. Schoolchildren without smartphones penalised with higher bus fares (theguardian.com)
    —discuss
  20. Stress-related sick days in the Netherlands rose 12% in a year, 49% in 5 years (nltimes.nl)
    —discuss
  21. Focal Prompt: tools for studying how AI systems allocate attention (focalprompt.com)
    —discuss
  22. Two Writes, One Row: Who Wins? (itsnas.me)
    1comments
  23. The Bayeux Tapestry Makes Landfall at the British Museum (wsj.com)
    —discuss
  24. Christophe Laudamiel (worldsensorium.com)
    —discuss
  25. Time, Layers, and Climate Futures (blue-continuum.com)
    —discuss
  26. Show HN: I built a tool that gives any website an API and MCP
    —discuss
  27. An API that beats GPT-4V, Gemini and Claude at detecting photo rotation (blueveta.com)
    —discuss
  28. The Greatest Pun in JavaScript (shukla.io)
    —discuss
  29. Book Review: Lee Kuan Yew's Memoirs (astralcodexten.com)
    —discuss
  30. NASA's Roman Team Confirms Ground Stations Receiving Data (nasa.gov)
    —discuss

Tell HN: OpenAI $500 ProMax plan listed in API

24 pointsby 1d ago
36 comments
1d agoHN ↗

the bottom line is that a.i is getting more and more expensive and they have excuses for that. just last week trae.ai reduced its plan credits from 400$ to 100$

21h agoHN ↗

Inference for the same level of intelligence is getting cheaper. Frontier intelligence keeps finding ways to get more expensive, yes.

But gpt-6-sol, for instance, dropped price for the same level of intelligence. That will keep happening, and someday we'll have basically-free sol. And hopefully by then, very cheap astra.

But some new wild thing will be very expensive.

I'm sure I'll figure out how to use the new one, but I would be extremely happy with near-free Astra.

13h agoHN ↗

Sol6 is clearly worse than 5.6. Have you tried using it? And Astra is more expensive.

11h agoHN ↗

Claude subs are great value again ever since Astra release moved lots of people to OpenAI subs. Even moreso since Opus 5.5 release.

2h agoHN ↗

We should measure cost per standard task. This is falling fast.

1d agoHN ↗

We stole all your work, just subscribe to get it back... now at $500/month.

12h agoHN ↗

I'd be tired too if my industry was being destroyed and scavenged by millionaires

5h agoHN ↗

your industry? lol

it didn't exist 100 years ago

adapt

3h agoHN ↗

when people say "if I was in situation x" there is an implication that they are not currently in situation x. it looks a bit odd when you go for the attack from uncertain footing man, I'd feel bad for attacking back knowing that you've interpreted my words differently to how others would.

1d agoHN ↗

I wonder what's the ceiling for these plans' prices. At some point, it becomes a non-negligible fraction of the salary for the employee, especially in non-US employment markets.

13h agoHN ↗

Its cute you tgink youll be the one leveraging AI.

13h agoHN ↗

A lot depends on how good open source models do. I know I can buy a 20k rig or whatever to get model X. If someone wants to charge me 3k a month to write code I just buy the hardware.

1h agoHN ↗

I’m interested: how would you spend $20k to get close to performance parity (speed and ability) with a frontier model?

13h agoHN ↗

Once people become fully dependent the ceiling will be whatever they want it to be.

8h agoHN ↗

I think there will be an employer backlash at some point. Paying an extra couple thousand in subscriptions so a junior can write copy-paste emails and Slack messages? Nah fam, write those yourself.

5h agoHN ↗

Or just use AI mode on Google's homepage. I asked it to find an article that compares a certain thing to another, which doesn't exist. So Google automatically wrote a 5 page article comparing the two things and built a proper article with cited references. Most people don't need a subscription to anything.

31m agoHN ↗

The google AI thing on their website gives me terrible results when I use it at work. Basically it's so unreliable I don't trust it. FWIW I don't log into a google account on my work devices (because we don't have Google enterprise services, and I'm not going to log into my personal Gmail account on my work computer) so that could be making a difference.

12m agoHN ↗

In my observation employers are willing to throw practically unlimited cash, probably because it’s directly funding a future where they don’t have to pay pesky engineers with their holidays and their pensions.

1d agoHN ↗

to be honest wasn't much surprised cause of the current situation like frontier models being so expensive and the models like opus consuming way much more tokens you just can't justify any of that. this token math is just so damned right now :(

1d agoHN ↗

If it's unlimited that would be tempting. But that will never happen and if it does "unlimited" would not be

21h agoHN ↗

Particularly in this situation, it cannot be unlimited. A command to one agent could scale into a whole datacenter very quickly.

It can feel unlimited for many use cases, probably. But I doubt it can, let's say, 3D model things in blender on 3 different computers for 24/7. And hey, why not 30 things? 300? If you're just watching them, it's easy to get ridiculous.

20h agoHN ↗

I understand your point; but I disagree - Unlimited is possible.

It gets into a "define unlimited" discussion. You could give "unlimited" usage, but still "limit" both concurrency and throttle speed. Even if you run 24/7, they still have knobs they can twist.

19h agoHN ↗

Fair enough, but you're kind of changing the parameters.

The limiting thing now is total tokens, so unlimited would mean infinite tokens.

And it's not hard to spin up to infinite tokens, one agent and one command can orchestrate it.

12h agoHN ↗

Is there really no limit on number of concurrent agents on either OpenAIs or Anthropic's plans?

9h agoHN ↗

There is always a limit unless you start scaling horizontally in the sense that you have multiple accounts via subscription or API. There is always some limit to any single account, even via API.

I'm thinking the persistent agent is something that can't scale to multiple agents. Imagine only one agent you can invoke as many times as you want on a single account, but it cannot invoke subagents or scale itself. It's just always there ready to do work when you need it. But it behaves as "one".

This would be trivial to implement. Of course, if you have the money, you can have multiple $500 subscriptions and have multiple persistent agents.

19h agoHN ↗

Cloud agents in Codex are quite poor solution compared to Claude Code, you are forced to use your own computer, which is not the desired setup for anyone with a standard laptop not suitable for doing lots of work in parallel.

So I would still prefer to spend some dollars in Anthropic or even Cursor solutions, which work much better as cloud dev envs.