Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Claude Code now reads AGENTS.md if there is no Claude.md(claude.com ↗)
    144comments
  2. Android 17 is the first since 3.x to add new APIs without releasing to the AOSP(grapheneos.social ↗)
    200comments
  3. Saving another 100TB of RAM(cloudflare.com ↗)
    36comments
  4. How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip(ieee.org ↗)
    9comments
  5. Cloudflare Quick Tunnels(cloudflare.com ↗)
    232comments
  6. Xcode 27.1 Beta Release Notes(developer.apple.com ↗)
    62comments
  7. Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)(arxiv.org ↗)
    11comments
  8. Photon-Emission-Guided Laser Fault Injection Enables RP2350 Secure Debug(ledger.com ↗)
    46comments
  9. How to Write with an LLM(sockpuppet.org ↗)
    253comments
  10. Show HN: Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash(cactuscompute.com ↗)
    72comments
  11. US troop deaths during Iran war exceed Pentagon count by at least four(reuters.com ↗)
    46comments
  12. The first new cat species discovered in 100 years(nationalgeographic.com ↗)
    37comments
  13. OpenJev(openjev.com ↗)
    238comments
  14. Cyclomatic Complexity in C#(ndepend.com ↗)
    10comments
  15. Two parallel neural ectoderm progenitors contribute to the developing brain(newscientist.com ↗)
    51comments
  16. The Implications of Linguistic Illegibility for LLM Security(arxiv.org ↗)
    17comments
  17. A 1542 papal cipher cracked with simulated annealing(simonklee.dk ↗)
    discuss
  18. C++26: Trivial infinite loops are no longer undefined behaviour(sandordargo.com ↗)
    168comments
  19. Warez: The Infrastructure and Aesthetics of Piracy (2021)(archive.org ↗)
    13comments
  20. Minimal Phone 2(minimalcompany.com ↗)
    155comments
  21. How SpaceX streamlined the Raptor engine(construction-physics.com ↗)
    25comments
  22. Size-Specialized Memory Allocation(go.dev ↗)
    3comments
  23. Inside ZCode: Silently uploading your Git history to the cloud(ferstar.org ↗)
    89comments
  24. A search-and-inference database from scratch in pure Zig(antfly.io ↗)
    16comments
  25. I vibed a proof of Conway's conjecture(overreacted.io ↗)
    177comments
  26. From Geometry to Algebra and Back Again: 4000 Years of Papers (2023) [video](youtube.com ↗)
    discuss
  27. Korea raises data breach fines to 10% of revenue(koreajoongangdaily.com ↗)
    74comments
  28. North Korean nuclear test sets off years of earthquakes(science.org ↗)
    149comments
  29. Cekura (YC F24) Is Hiring(ycombinator.com ↗)
    discuss
  30. US Military had close call after using AI for hallucinated intelligence report(cnn.com ↗)
    283comments

Claude Code Removed from $20-a-Month "Pro" Subscription for New Users

79 pointsby 5mo agowheresyoured.at
38 comments
5mo agoHN ↗

Maybe those already on $20 a month plans won't be nerfed much more?

It's yet another austerity move, pretty much in line with the recent ones.

4mo agoHN ↗

Maybe you should adapt your expectations for your 20$ monthly subscription. That's by far not what it's worth

I burned already through over 100$ in a single day with per token payment for heavy usage. I don't have any issues with Claude not working or being "nerfed". It just works

5mo agoHN ↗

This should be a warning to those who feel that it's ok to offload your creativity to a subscription service. Always need a local model in some form.

4mo agoHN ↗

I’m sorry but that’s just dumb. An LLM is a tool. Your brain is not a substitute for an LLM in the same way your fingers are not a substitute for a wrench.

The year is 2026 and if you are using your brain on chore work like one-off scripts, refactoring, boilerplate test code, then you are wasting time and money and I don’t want to work with you.

Local models are fine for this and can do it in a fraction of the time your brain will take to even get bootstrapped

4mo agoHN ↗

The year is 2026 the average RAM for the most common type of developer’s (web) machine is 16GB. 8 will be the lower end. Tell me which model can one run on this machine locally?

4mo agoHN ↗

You can use free models on opencode. Minimax whatever is just free. It's more than enough for these tasks.

4mo agoHN ↗

You could judge the costs of the AI products you're using by the standard API pricing, not promotional subscription offers.

4mo agoHN ↗

Not even that way, given that the price is still highly subsidized by investors and circular deals.

4mo agoHN ↗

For me, it's not even cost necessarily. If they decide to change the product they offer, the old one is gone. I refuse to use anything for personal use that's not at least _available_ as model weights.

4mo agoHN ↗

Bingo. This same scenario with IOT hardware/software requirements and the ever changing software updates where features get added/removed (and firmware), etc would have so many here up in arms!

Oddly, less (vocal) arms on this specific case.

4mo agoHN ↗

Are there local models that are anywhere near as good at coding as opus 4.6?

4mo agoHN ↗

People will insist otherwise, but I haven't seen anything close to sonnet 4.6 that can be run locally.

4mo agoHN ↗

I don't think anyone can honestly say a huge frontier model is actually going to be matched by something running on 64gb locally?

4mo agoHN ↗

I have read many comments saying Qwen3.5 various ~30B models, Gemma 4 ~30B models and now Qwen3.6 "better than sonnet".

I don't know how large sonnet and opus are but the rumor is 1T and 5T respectively.

4mo agoHN ↗

You don't have to use the most recent bleeding edge model to succeed. A local FOSS coding agent coupled with a reasonably priced LLM could yield the optimal ROI.

4mo agoHN ↗

Not really. Qwen 3.5 and Gemma and a couple of others are quite good though, and the quants are _very_ runnable on a good gpu.

4mo agoHN ↗

I keep telling them and they still want to spend money on tokens at the Anthropic casino, even though they are egregiously price gouging and applying upper limits so you spend more on tokens.

Sometimes you can't help gamblers who want to gamble on tokens to hit the jackpot on fixing a typical issue which can be done by local models or even reading the documentation.

4mo agoHN ↗

There is very little vendor lock. We can keep using subsidized model until it’s not. Then switch to next subsidized model.

4mo agoHN ↗

This doesn’t affect existing users.

This is a simple supply and demand curve.

Higher demand means the price goes up .. this has been true of things since before SaaS and before computers

4mo agoHN ↗

Thanks for all the logical fallacies in one comment.

4mo agoHN ↗

No. It means eventually existing users will be affected too. These companies are deeply in the red.

4mo agoHN ↗

Local models are not comparable to the FOTA models at all. I know what I'm saying because I do have 4 local H100's in my server, and could run the very best local models. It's night and day. They are unusable and stupid.

4mo agoHN ↗

For what do you use the 4 local H100s then?

4mo agoHN ↗

For training our AI model of course. Inference is for the cheaper machines.

4mo agoHN ↗

I get perfectly acceptable results from a Strix Halo PC the size of a shoebox, man. An APU that uses ~150w, has 0 discrete GPUs, and a bill of $0/m. What's more, it doesn't go down every week, limit use, or change the terms at a whim.

I'll burn/discard 'frontier' tokens (at work) only because they're mandated and they foot the bill. I'd rather resell them; meet the asinine requirement from $EMPLOYER, provide cover for outsourcing to my equipment, and get a return for the hassle.

TLDR: perhaps you're holding it wrong or haven't tried the latest, as we so often hear. That's a lot of GPU for not much utility.

4mo agoHN ↗

Well, my python and typescript folks are also happy with the simplier local models. But I'm using more advanced stuff, C/C++ embedded real-time, vision AI, and compilers.

4mo agoHN ↗

Fair point. I treat LLMs like the forgetful junior we often hear about. The things I don't care to do, they (both local and hosted/'frontier') can. Boilerplate, very-well-described edits, some research/report, etc; a lot is riding on 'acceptable'.

Easier to spawn another terminal pane/browser tab than hire a contractor, I just don't find the 'frontier' services/terms compelling.

4mo agoHN ↗

That's my auto-correction, because I'm doing too much embedded (Firmware-over-the-air updates). Frontier it should be called.

4mo agoHN ↗

I’ve been trying to bring this up at my work, you’re putting all your intelligence into a service you don’t own. What do you do when it’s down or they quadruple the price ?

4mo agoHN ↗

He mentions Max as another place that they didn't properly predict plan and pricing relative to usage. I'd bet the farm that it's the next to be 'A/B' tested.

4mo agoHN ↗

They've been folding under government pressure for months. I think they lost control of their own company. Still they have a nice writing voice, but I think Google will be the last man standing when this is all over.

4mo agoHN ↗

This is to curb users like me who are incrementally adding claude code capacity as the project ramps up by opening new accounts on $20 plans.