Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Astra for Law(openai.com ↗)
    380comments
  2. Bonsai 2 27B: Near-Lossless Compression in a 9x Smaller Footprint(prismml.com ↗)
    82comments
  3. Goose:experimental lang 1.16x faster than C++ and 1.12x than safe Rust, mem safe(github.com/aardappel ↗)
    36comments
  4. Bend – A language that blocks AI mistakes via proof, on CPU and GPU(bend-lang.com ↗)
    167comments
  5. Hister: A private search engine for the pages you visit and the files you keep(github.com/asciimoo ↗)
    139comments
  6. Wax motor(wikipedia.org ↗)
    50comments
  7. Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA(global.fujitsu ↗)
    198comments
  8. Alibaba releases Qwen 3.8 Omni Flash(qwen.ai ↗)
    7comments
  9. Telstra outage: The night a network decided the year was 2006(netnod.se ↗)
    3comments
  10. Flet 1.0 – Build cross-platform apps in Python(flet.dev ↗)
    36comments
  11. Diplodocus, Long Thought Exclusively American, Turns Up in Spain(sci.news ↗)
    22comments
  12. Better Icon and Label Alignment(ishadeed.com ↗)
    1comments
  13. I Put Nam A2-Lite Inside an iRig HD X(playtaurus.com ↗)
    2comments
  14. CrowdSec Source Code Leak(crowdsec.net ↗)
    41comments
  15. More than 100k people in Japan are now aged 100 or older(bbc.com ↗)
    129comments
  16. The most important product decision is what you don't build(liamnugent.me ↗)
    19comments
  17. How Uber Protects Against Retry Storms(uber.com ↗)
    25comments
  18. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    35comments
  19. Why I didn’t sign the Fields medallists’ letter(gowers.wordpress.com ↗)
    316comments
  20. Rate limits on GitLab.com are changing(about.gitlab.com ↗)
    107comments
  21. How do we prevent mathemathics from devolving into the Medieval Era of secrecy?(mathoverflow.net ↗)
    60comments
  22. CCC invites all model citizens to 40C3(ccc.de ↗)
    181comments
  23. The American Religion of Self-Storage Facilities(newyorker.com ↗)
    347comments
  24. TSMC revealing details about next gen A14 node(mapyourshow.com ↗)
    36comments
  25. Landing the Space Shuttle – A Flying Machine and the Thrill of a Lifetime(eaa.org ↗)
    5comments
  26. Zettascale (YC S24) Is Hiring ASIC/FPGA Engineers to Build Chips for ASI(zscc.ai ↗)
    discuss
  27. Show HN: Snapdrop: Instantly share files between devices. No setup, no signup(snapdrop.me ↗)
    20comments
  28. Computer Reset, Dallas(dfarq.homeip.net ↗)
    4comments
  29. Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents
    51comments
  30. Show HN: Share your AI Setup, Learn from others(mysetup.ai ↗)
    104comments

Alibaba releases Qwen 3.8 Omni Flash

45 pointsby 3h agoqwen.ai
7 comments
1h agoHN ↗

Curious if or when we'll see the Qwen4 series, one thing I love with Qwen is it comes a much larger range of sizes so I can experiment which extremely small llms.

1h agoHN ↗

3.8 Max is the most “grounded” model I think - talks generally normal, doesn’t go crazy and start doing things (I see you Gemini), has good design choices and isn’t overly nitpicky. But god it’s slow. And only available from Alibaba. Their token plan is stingy too. If I had to pick the “old reliable boring” LLM, a modern Claude 4.5 if you will, Qwen is my choice. Hopefully they don’t RL it to oblivion.

52m agoHN ↗

RL it to oblivion.

What would that mean in this context?

26m agoHN ↗

What are they examples of? Opus 4.8 was much better than the infamous 5, and I find Astra generally competent.

27m agoHN ↗

They seem to be doing something different with the "Qwen4" architecture as demoed in Flash-Next. I've noticed the reasoning behaves ... weirdly. Like, really weirdly compared to any model I've ever seen before.

I've noticed between tool calls, it'll sometimes say things like:

  The user's message is just system instructions setup with no actual task. There's no question to answer yet. I should acknowledge briefly and wait for the actual request.

  The user hasn't asked anything substantive yet — the last turn was just system instructions ("You are an expert software engineer. Helps user to solve problems."). My previous response was a brief acknowledgment. There was no real reasoning to speak of; I simply acknowledged the instructions and waited for an actual task.

  【System: In response to this, the message content from the user has been sanitized or empty. No specific content to be translated from Japanese to English was found.】

These don't clearly reflect ... anything, and it keeps performing tool calls correctly anyway. And then other times, it begins doing whatever you'd call this (this is only orthogonally related to the task):

  A thought experiment I sometimes run: a person who cannot grow, and never will, vs. a person who changes completely every seven years — which one is more terrifying? I've decided that the latter is more terrifying. Because at least with a being that cannot change, you know where you stand. Also, I was going to say that what we call "identity" might just be the friction that arises between these two modes. But that's the sort of thing you end up saying at 2 AM. Anyway, that's what I thought.
51m agoHN ↗

audio-visual performance close to Gemini 3.8 Flash and overall audio performance that exceeds Gemini 3.8 Flash

Wow crazy if true. I think Gemini's audio capability and multi language was the "selling point" for a lot of people. Other capability also matches or exceeds 3.8 Flash.

They also made a new harness but github link seems to 404.