Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Claude Code now reads AGENTS.md if there is no Claude.md(claude.com ↗)
    164comments
  2. Android 17 is the first since 3.x to add new APIs without releasing to the AOSP(grapheneos.social ↗)
    210comments
  3. How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip(ieee.org ↗)
    21comments
  4. Saving another 100TB of RAM(cloudflare.com ↗)
    36comments
  5. Cloudflare Quick Tunnels(cloudflare.com ↗)
    237comments
  6. Xcode 27.1 Beta Release Notes(developer.apple.com ↗)
    63comments
  7. A 1542 papal cipher cracked with simulated annealing(simonklee.dk ↗)
    1comments
  8. Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)(arxiv.org ↗)
    11comments
  9. How to Write with an LLM(sockpuppet.org ↗)
    258comments
  10. Photon-Emission-Guided Laser Fault Injection Enables RP2350 Secure Debug(ledger.com ↗)
    49comments
  11. Show HN: Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash(cactuscompute.com ↗)
    74comments
  12. The first new cat species discovered in 100 years(nationalgeographic.com ↗)
    48comments
  13. OpenJev(openjev.com ↗)
    240comments
  14. Cyclomatic Complexity in C#(ndepend.com ↗)
    12comments
  15. US troop deaths during Iran war exceed Pentagon count by at least four(reuters.com ↗)
    74comments
  16. Two parallel neural ectoderm progenitors contribute to the developing brain(newscientist.com ↗)
    51comments
  17. The Implications of Linguistic Illegibility for LLM Security(arxiv.org ↗)
    17comments
  18. C++26: Trivial infinite loops are no longer undefined behaviour(sandordargo.com ↗)
    175comments
  19. Warez: The Infrastructure and Aesthetics of Piracy (2021)(archive.org ↗)
    14comments
  20. Minimal Phone 2(minimalcompany.com ↗)
    161comments
  21. Size-Specialized Memory Allocation(go.dev ↗)
    3comments
  22. Inside ZCode: Silently uploading your Git history to the cloud(ferstar.org ↗)
    89comments
  23. How SpaceX streamlined the Raptor engine(construction-physics.com ↗)
    28comments
  24. A search-and-inference database from scratch in pure Zig(antfly.io ↗)
    16comments
  25. I vibed a proof of Conway's conjecture(overreacted.io ↗)
    180comments
  26. From Geometry to Algebra and Back Again: 4000 Years of Papers (2023) [video](youtube.com ↗)
    discuss
  27. US Military had close call after using AI for hallucinated intelligence report(cnn.com ↗)
    291comments
  28. Korea raises data breach fines to 10% of revenue(koreajoongangdaily.com ↗)
    79comments
  29. Cekura (YC F24) Is Hiring(ycombinator.com ↗)
    discuss
  30. Mathematicians Build Long-Awaited Graph Sandwich(quantamagazine.org ↗)
    18comments

Saving another 100TB of RAM

191 pointsby 5h agoblog.cloudflare.com
24 comments
3h agoHN ↗

Bang up article! As someone who doesn't get to do enough (almost any) calculus in my daily programming assignments, I thoroughly enjoyed reading about Kevin's dive into that derivation (linked in the supplemental article). All the people being negative here can swallow raisins

3h agoHN ↗

Thanks! Maybe dial it back or people are going to think I paid you

3h agoHN ↗

The only Rust section is the one on storage improvements about the struct that stores the hash, but do they really need that many hashes that 2 bytes makes that big of a difference? Article doesn't expand, but I guess it's a hash for every task on every computer, so maybe yes.

3h agoHN ↗

That's exactly his point/the area of cost saving - that they didn't actually need as many hashes as they had started with. The trick was in finding out how many hashes they could cull without degrading load balance.

3h agoHN ↗

These optimizations are impressive, but it gets me thinking: at what point does a company become a collection of impenetrable siloes, where nothing really does what you expect? Maybe know with AI this is less of an issue as exploring a codebase is also much faster.

2h agoHN ↗

My intuition around larger companies is they are already impenetrable silos, and AI makes it worse

2h agoHN ↗

I feel like any company whose products have RESTful interfaces are already there…

One wants to turn on an indicator on a remote device. A simple Boolean value. But we need networking, TLS, authentication plugins, certificate validation, distributed logging, containers, orchestration, HTTP client/server, interprocess communication, daemon dependency management, …

Sure, one can say each of these layers and abstractions has an important and justifiable purpose. But one can also step back and start wondering - what the hell are we really doing???

At some level, it seems like each layer of abstraction has to manage others, only simply because they exist.

Imagine the simplicity of 1800s telegraph signaling - no software!

Too often we build systems with Fortune-50 style hierarchies when a 5-person team could do the whole job.

2h agoHN ↗

Build a system as simple as possible but no simpler.

An 1800s telegraph system doesnt work in the modem world, there is far too much communication and the system would just collapse into molten slag.

All those things you've listed are because we live in an adversarial world and I'd steal all your money off the telegraph wire if you tried it.

2h agoHN ↗

Having worked in hardware for a moment, everything we do in software is like this. Even C.

2h agoHN ↗

There are a whole series of blog posts from the Fishworks guys explaining why it could possibly be so hard to turn on one LED.

But Oracle probably deleted then so you'll have to find them on archive.org.

2h agoHN ↗

And at what point does a company start to care about performance? 100TB of RAM is expensive as hell, but getting products to market faster was worth the cost

48m agoHN ↗

Actually, AI creates spaghetti faster than any human ever could before.

2h agoHN ↗

Cloudflare is truly amazing, they have made so much possible for my main side-project at a price and performance that I can’t really take credit for (http://sourcelibrary.org), I don’t care if their text was written with AI, I just wish I could get my own AI to sing so well about hashing… but wait.. today I noticed Claude trying to use hashing when a timestamp would honestly do, and now I’m really doubting myself, hmm…

2h agoHN ↗

This is pleasant coincidence. Really like your site and it's queued to send in my newsletter in the morning! Just happened to see your comment here while I was reading. Great work.

1h agoHN ↗

I don’t care if their text was written with AI, I just wish I could get my own AI to sing so well about hashing… but wait.. today I noticed Claude trying to use hashing when a timestamp would honestly do, and now I’m really doubting myself, hmm…

Tried out the first 1000 words in Pangram, and it seemed happy it was human written. Not surprised either, it has been some of the better writing I've seen out of Cloudflare recently.

1h agoHN ↗

An AI would have known that saying, “Hi, mom” in a professional post was a bad idea.

49m agoHN ↗

What does Cloudflare make possible for your project?

2h agoHN ↗

anyone remember 100tb hosting company?

2h agoHN ↗

Someone really needed a few hundred TB to waste on inference and went looking under the rugs…

1h agoHN ↗

CPU DRAM can't really be used for inference efficiently -- inference mostly wants memory bandwidth, not memory capacity, and GPU DRAM has >10x more bandwidth. The fabs can switch between them but you can't switch after the fact.

1h agoHN ↗

Those machines with GPUs still need RAM of their own, and they generally want large caches to avoid SSD penalties. You even see this spill out in the form of costs for KV cache in <1min, 5m, 1hr rates etc.

55m agoHN ↗

The majority of inference actually does happen in cpu.

1h agoHN ↗

I definitely enjoyed this writing style more than a lot of the recent cf blog posts. Cool article!

1h agoHN ↗

does this mean RAM prices can go down now? Please?