Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Don't Be Nice(roe.dev ↗)
    discuss
  2. I Hijacked a Real Artist's Spotify with AI Music. It Was Disturbingly Easy(404media.co ↗)
    discuss
  3. Enjambre – a durable kernel for swarms of AI agents (Python, MCP)(github.com/santibccc-sudo ↗)
    discuss
  4. Show HN: AI Facial Attractiveness Model Aligned with Human Preferences(faceanalysisai.com ↗)
    discuss
  5. Show HN: I cleaned AI watermarks from my clipboard(pastezero.com ↗)
    1comments
  6. Google Pixel phones pwned in zero-click attacks(theregister.com ↗)
    1comments
  7. Show HN: Fluentry – local voice-to-text dictation for Linux, Wayland included(fluentry.github.io ↗)
    discuss
  8. A.I. Gone Rogue? Trump Proposes Not New Rules but an 'A.I. Force.'(nytimes.com ↗)
    discuss
  9. Show HN: Vyne, a 205MB on-device decision model with typed, calibrated outputs(toolchain.studio ↗)
    discuss
  10. Show HN: Will Jev pull the lever in the trolley problem?(gpu.studio ↗)
    discuss
  11. AI CLI Insights for popular CLI products(libraries.io ↗)
    1comments
  12. Lollipop camera to HomeKit: RTSP extraction and setup(changologs.site ↗)
    discuss
  13. Tapa Shotor (Camel Hill), Afghanistan(wikipedia.org ↗)
    discuss
  14. Pluto AI is top ranked in StarCraft BroodWar competitive scene(youtube.com ↗)
    1comments
  15. Navier–Stokes Priority Controversy(wikipedia.org ↗)
    discuss
  16. It's finally here: Porsche puts wireless EV charging into production(electrek.co ↗)
    discuss
  17. Electrif-AI Everything(energyandstuff.substack.com ↗)
    discuss
  18. OpenAI and Microsoft knew they were starting a 'doom loop' for the web(theverge.com ↗)
    1comments
  19. Show HN: ManifestGo – AI tuned for MV3 extension boilerplate and edge case(manifestgo.app ↗)
    discuss
  20. Teleoperated Humans(jefftk.com ↗)
    discuss
  21. Japan Opens the Door to Acquiring Nuclear-Powered Submarines(twz.com ↗)
    discuss
  22. Analysis: Global fossil-fuel emissions set to fall in 2026 amid Hormuz crisis(carbonbrief.org ↗)
    discuss
  23. Show HN: Hush – GitHub issue triage that abstains when it isn't confident(github.com/emreozyoruk ↗)
    discuss
  24. Show HN: How to Vibe Code Hardware(charlesazam.com ↗)
    discuss
  25. Van Der Waals Force(wikipedia.org ↗)
    discuss
  26. AI on Top of a Dysfunctional System: The Product Backlog(age-of-product.com ↗)
    discuss
  27. Every Nvidia GPU has 10 to 30 RISC-V cores inside it(xda-developers.com ↗)
    discuss
  28. Are global stock markets heading for a crash?(theguardian.com ↗)
    2comments
  29. KDE turns 30 and someone's brought an AI-native desktop proposal(theregister.com ↗)
    15comments
  30. Parth Anand – Projectsadak.in Hacked
    discuss

Step 5 Preview: Advancing the Pareto Frontier

54 pointsby 4h agostepfun.com
16 comments
3h agoHN ↗

I regularly hit 200-300M cached reads every day on some of the models I use. It has exceeded 7-800M on a couple of occasions. At $0.04/M, that is $8-12 per day only for cached reads.

1h agoHN ↗

At $0.04/M

Unless you meant step-3.7-flash, the input cache hits are $0.05 per mil for step-5-preview.

$8-12 per day only for cached reads

Pretty decent "API" rates for ~500M+ tokens on Step Fun 5, a Kimi K3 / GLM 5.3 level model?

Their "Step Plan" is ridiculous, by comparison: ~$60 usage on $6.99/mo; ~$220 on $9.99/mo. https://platform.stepfun.ai/docs/en/step-plan/overview

26m agoHN ↗

Yes, I meant the Flash version.

I have used Kimi 2.5 and GLM 5.3 (& 5.3 Flash). Do not need them for what I do outside of spec hardening (basically, a lot of chatting).

I tend to know exactly what I want and most of the weaker models are enough to get me there. I have mainly been using MiMo, DeepSeek V4 Flash and MuseSpark Contributor over the last month or so.

2h agoHN ↗

Sometimes 4 is skipped due to being considered unlucky.

2h agoHN ↗

In China and in places influenced by Chinese culture, due to homonymy between "4" and death.

2h agoHN ↗

Without any Pokémon-specific optimization, Step 5 Preview has so far sustained progress for more than 3,000 turns and 6 million tokens of interaction. By turn 3,082, it had unlocked Cut, earned three Gym Badges, and defeated Lt. Surge. The run is now roughly one-third of the way through the main story.

Finally FireRed is being used as a benchmark again! I believe Astra can beat it in 18 hours. Not sure how that compares.

2h agoHN ↗

  > Built on a sparse Mixture-of-Experts architecture, Step 5 Preview has 600B total parameters, with 27B active per token, and supports a 1M-token context window and vision input.

  > Step 5 Preview scores 44 on the Artificial Analysis Intelligence Index.

  > The model will be released with open weights on October 15.

I guess being Chinese company they decided to skip version 4, while also giving impression to be on the similar iteration with leading companies (claude opus 5). I wonder if other Chinese labs like Kimi/Moonshot will follow suit.

1h agoHN ↗

Moonshot has already teased K3.1 so not likely

1h agoHN ↗

K3.1 would likely be a deeper/longer post-train from K3, so that’d make sense.

It’s all marketing anyways, but that’s at least how a lot of labs have been naming things (sometimes).

24m agoHN ↗

Another possible reason is that the number 4 is considered unlucky in traditional Chinese culture.

1h agoHN ↗

Their posisitoning is nice. Instead of saying they are cheaper and a bit less performant (in terms of intelligence), they say they are best among the cheaper and a bit less performant ones.

33m agoHN ↗

How about adding a contested historical facts benchmark?

22m agoHN ↗

IT's Artificial Analysis Index is the same as Kimi K3, which is about 4.6x bigger, and GLM 5.3, which is about 1.25x bigger. Pricing is $1/$2.70 i/o. Openweights on October 15.