Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Enterprise Cyber Risk Management(andersenlab.com ↗)
    discuss
  2. Be Careful with Your Select * Queries(notesonsystems.com ↗)
    discuss
  3. TeaonherChecker(teaonher.org ↗)
    discuss
  4. Leaping Sun Dogs (2016) [video](youtube.com ↗)
    discuss
  5. Jev is the fastest-adopted model in AI Gateway history(vercel.com ↗)
    1comments
  6. A font that reads what you wrote(rohanadwankar.github.io ↗)
    1comments
  7. A deep dive into Jev, TypeSafe's System One model(flaviocopes.com ↗)
    discuss
  8. CoyoPedal – Full-Size Neural Amp Modeler Captures on an ESP32-S3(github.com/dashersw ↗)
    discuss
  9. I'm so bad at billiards that I ended up in the hyperbolic plane [video](youtube.com ↗)
    discuss
  10. MIT Just Proved LLMs Will Stop Getting Smarter, and Money Won't Fix It [video](youtube.com ↗)
    discuss
  11. Show HN: Delightful Cells – reliable AI batch processing for spreadsheets(delightfulcells.com ↗)
    1comments
  12. Clawptcha: Reverse Captcha(clawptcha.com ↗)
    discuss
  13. Neutron Radiation, Emission and Scattering (Neutron Chemistry)(ossila.com ↗)
    discuss
  14. Drop-in Django app to serve a decent LLM-ready documentation site(chesselink.com ↗)
    1comments
  15. I built an extension that hides your personal information from AI(github.com/arikchakma ↗)
    discuss
  16. Show HN: I-server: Hide server using ICMP reflection/Destination Unreachable(github.com/hajoon22 ↗)
    discuss
  17. Benchmarking Wild vs. Mold(davidlattimore.github.io ↗)
    discuss
  18. Microsoft agentically ports Copilot runtime to Rust for $120K(theregister.com ↗)
    discuss
  19. Why China is pushing back on US warnings over rapid AI development(theguardian.com ↗)
    1comments
  20. Boston to Calcutta Ice Trade Began with One Shipment(curiowire.com ↗)
    discuss
  21. Show HN: Book recommendations with live library availability(nanosheep.net ↗)
    3comments
  22. Show HN: I built Hacker News alternative and scale to 11,000 users(founderstoday.org ↗)
    discuss
  23. Don't Be Nice(roe.dev ↗)
    1comments
  24. I Hijacked a Real Artist's Spotify with AI Music. It Was Disturbingly Easy(404media.co ↗)
    discuss
  25. Enjambre – a durable kernel for swarms of AI agents (Python, MCP)(github.com/santibccc-sudo ↗)
    discuss
  26. Show HN: AI Facial Attractiveness Model Aligned with Human Preferences(faceanalysisai.com ↗)
    discuss
  27. Show HN: I cleaned AI watermarks from my clipboard(pastezero.com ↗)
    1comments
  28. Google Pixel phones pwned in zero-click attacks(theregister.com ↗)
    1comments
  29. Show HN: Fluentry – local voice-to-text dictation for Linux, Wayland included(fluentry.github.io ↗)
    discuss
  30. A.I. Gone Rogue? Trump Proposes Not New Rules but an 'A.I. Force.'(nytimes.com ↗)
    discuss

Step 5 Preview: Advancing the Pareto Frontier

58 pointsby 5h agostepfun.com
16 comments
3h agoHN ↗

I regularly hit 200-300M cached reads every day on some of the models I use. It has exceeded 7-800M on a couple of occasions. At $0.04/M, that is $8-12 per day only for cached reads.

2h agoHN ↗

At $0.04/M

Unless you meant step-3.7-flash, the input cache hits are $0.05 per mil for step-5-preview.

$8-12 per day only for cached reads

Pretty decent "API" rates for ~500M+ tokens on Step Fun 5, a Kimi K3 / GLM 5.3 level model?

Their "Step Plan" is ridiculous, by comparison: ~$60 usage on $6.99/mo; ~$220 on $9.99/mo. https://platform.stepfun.ai/docs/en/step-plan/overview

1h agoHN ↗

Yes, I meant the Flash version.

I have used Kimi 2.5 and GLM 5.3 (& 5.3 Flash). Do not need them for what I do outside of spec hardening (basically, a lot of chatting).

I tend to know exactly what I want and most of the weaker models are enough to get me there. I have mainly been using MiMo, DeepSeek V4 Flash and MuseSpark Contributor over the last month or so.

3h agoHN ↗

Sometimes 4 is skipped due to being considered unlucky.

3h agoHN ↗

In China and in places influenced by Chinese culture, due to homonymy between "4" and death.

3h agoHN ↗

Without any Pokémon-specific optimization, Step 5 Preview has so far sustained progress for more than 3,000 turns and 6 million tokens of interaction. By turn 3,082, it had unlocked Cut, earned three Gym Badges, and defeated Lt. Surge. The run is now roughly one-third of the way through the main story.

Finally FireRed is being used as a benchmark again! I believe Astra can beat it in 18 hours. Not sure how that compares.

3h agoHN ↗

  > Built on a sparse Mixture-of-Experts architecture, Step 5 Preview has 600B total parameters, with 27B active per token, and supports a 1M-token context window and vision input.

  > Step 5 Preview scores 44 on the Artificial Analysis Intelligence Index.

  > The model will be released with open weights on October 15.

I guess being Chinese company they decided to skip version 4, while also giving impression to be on the similar iteration with leading companies (claude opus 5). I wonder if other Chinese labs like Kimi/Moonshot will follow suit.

2h agoHN ↗

Moonshot has already teased K3.1 so not likely

1h agoHN ↗

K3.1 would likely be a deeper/longer post-train from K3, so that’d make sense.

It’s all marketing anyways, but that’s at least how a lot of labs have been naming things (sometimes).

1h agoHN ↗

Another possible reason is that the number 4 is considered unlucky in traditional Chinese culture.

2h agoHN ↗

Their posisitoning is nice. Instead of saying they are cheaper and a bit less performant (in terms of intelligence), they say they are best among the cheaper and a bit less performant ones.

1h agoHN ↗

How about adding a contested historical facts benchmark?

1h agoHN ↗

IT's Artificial Analysis Index is the same as Kimi K3, which is about 4.6x bigger, and GLM 5.3, which is about 1.25x bigger. Pricing is $1/$2.70 i/o. Openweights on October 15.