Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Digital Cash Payoff (2001)(technologyreview.com)
    —discuss
  2. Show HN: DrawCMS – An open-source animated diagramming tool for AI Agents(github.com/drawcms)
    —discuss
  3. Logs you, your AI and cron jobs write to that you can prove weren't altered(freshjots.com)
    —discuss
  4. Show HN: bananabread AI – Find what customers love and want from your products(bb-product-api-docs.web.app)
    —discuss
  5. GoDaddy receives takeover offer from maker of Norton antivirus software(ft.com)
    —discuss
  6. If you don't have the factories, you lose the expertise(lemire.me)
    —discuss
  7. Show HN: Headwire – WireGuard with NAT traversal via Tailscale's magicsock(github.com/brofranks)
    —discuss
  8. Namernut – domain name generator that scores how crowded a name is(namernut.com)
    —discuss
  9. Show HN: MEF LLM Studio(mef-llm-studio.com)
    —discuss
  10. Fragments: September 24(martinfowler.com)
    —discuss
  11. Double-entry bookkeeping and paper and tokens(pokorny.ca)
    —discuss
  12. Pummelvision Back(pummelvision.ai)
    —discuss
  13. Legal and General to Cut 10% of Jobs by Middle of 2027(wsj.com)
    —discuss
  14. Beware Overreliance on Metaphor(evnm.substack.com)
    —discuss
  15. AI Workers' Inquiry 2026(techworkersinquiry.org)
    —discuss
  16. .si domain registrations outpace .ai after Trump's 'Super Intelligence' rename(netcraft.com)
    —discuss
  17. Unmasking "zombie cells" in aging tissue with an AI-powered barcode(news.mit.edu)
    —discuss
  18. Weapons of Mass Decentralization Magazine(worksinprogress.co)
    —discuss
  19. Why Washington Governs by Fiscal Cliff(medium.com/freedomofthought)
    —discuss
  20. Forging 1024-bit RSA signatures in nearly SNFS time(iacr.org)
    —discuss
  21. The Chosen People(medium.com/freedomofthought)
    —discuss
  22. Show HN: OpenArcade – open-source software from the VR arcade I ran for 4 years(github.com/aprabh96)
    1comments
  23. Vibe Coded Musical Correspondences(caerjar.github.io)
    —discuss
  24. China's Mind on AI(axios.com)
    —discuss
  25. It's "Underwear on the Outside" Time(paulkrugman.substack.com)
    1comments
  26. Mythical Thought and Scientific Thought(medium.com/freedomofthought)
    —discuss
  27. DBDelve: Fast and modern DB Client made in Rust and gpui
    1comments
  28. Excel adds nested arrays(techcommunity.microsoft.com)
    —discuss
  29. Near-Total Liquid Heat Capture Keeps AI Racks Cool(ieee.org)
    —discuss
  30. Show HN: Critic – Review code with the agent that wrote it(critic.run)
    —discuss

UkisAI Swift Series / 27B, Flash Next and Bonsai 2 /-63.4% thinking, x1.95 speed

1 pointsby 1h ago
1 comments
Hey everyone,

Jovan from UkisAI here! Today, we are introducing Swift, a family of efficient reasoning LLMs based on Qwen, trained by penalizing tokens related to pathological overthinking patterns and restoring accuracy via RL and OPD.

After amazing feedback and 350k+ downloads in 13 days on our Swift Qwen 3.8 27B we are releasing the entire model family as well as the highly requested GSQ-RCO quants for 27B and Flash-Next.

This release includes: Swift1.5 27B, an improved version of our last model, with even lower token usage, fixed bugs and better agentic performance, with -58.5% thinking tokens while scoring 0.35% higher while outperfoming base on Terminal Bench 2.1 by not falling into "overthinking error" loops Swift Flash Next, with 63.4% fewer thinking tokens and a 1.8x speed up scoring -0.2% vs base on xhigh Swift Bonsai 2, with 39.8% fewer thinking tokens while scoring 0.19% higher (although we'd still like to note it as experimental)

Our benchmarks are ran x5 on Base and Swift, averaging across five seeds and various domains, including General (GPQA, AIME26), Coding (LiveCodeBench), Vision (ERQA), Agentic (Terminal Bench 2.1). One note is that the Terminal Bench 2.1 scores are misleadingly low at first glance. It is not a bug, but a simple matter of the Swift models not falling into overthinking loops and failing the task, rather pursuing it until the end, leading to higher average token usage. The token reduction still falls in the -38.7% range when compared apples-to-apples.

We are including a Free Research API and HuggingFace Spaces to give the models a spin before downloading or if you don't have enough compute to run them right now! You can find both on the model cards.

We have also made GGUF, NVFP4, MLX and W4A16 quants for relevant model versions.

More details on our training approach and community feedback can be seen here: https://www.reddit.com/r/LocalLLaMA/comments/1wg7dd5/ukisai_swiftqwen3827b_583_thinking_x195_speed/

All of the various quantization and model versions are available in their respective collections: Swift1.5 27B: https://huggingface.co/collections/ukisai/swift-15-27b Swift Flash Next: https://huggingface.co/collections/ukisai/swift-flash-next Swift Bonsai 2: https://huggingface.co/collections/ukisai/swift-bonsai-2

We would greatly appreciate your feedback via independent evaluations. As per last release, we operate on a candy-shop basis, trying to fulfill as many Swift model requests and quants as possible, so please do share your needs in the comments!