Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Agent Aced the Task. Will It Do It Again?(huggingface.co ↗)
    discuss
  2. The truth about 'Naza,' or collateral damage(timesofisrael.com ↗)
    1comments
  3. Why Whisper and medical speech APIs are making potentially fatal errors(appen.com ↗)
    discuss
  4. Agent Session Introspection(lmstudio.ai ↗)
    discuss
  5. U.S. Strike on Iranian School May Have Been War Crime, U.N. Report Says(nytimes.com ↗)
    discuss
  6. Five air-gapped, zero-telemetry utilities for privacy and operational leaks(github.com/warknoc ↗)
    discuss
  7. The GDP-per-person slowdown threatens global living standards(economist.com ↗)
    discuss
  8. Air Force secretary acknowledges the US has weapons in space(abcnews.com ↗)
    discuss
  9. Show HN: OpenAI has a live market price before it has a stock ticker(tickerlayer.com ↗)
    2comments
  10. Open weights models are surprisingly aligned on offensive cyber(pentesttoday.com ↗)
    discuss
  11. Google Earth to be axed next year [video](youtube.com ↗)
    1comments
  12. What Little Hope Remains(coredump.cx ↗)
    discuss
  13. Mind Mapping for Omarchy(mindarchy.xyz ↗)
    1comments
  14. Ask HN: What exists if I want a premade SoC/MCU/screen/enclosure/battery?
    1comments
  15. What's New in Vapor 5 Beta(vapor.codes ↗)
    discuss
  16. How to Get Your First 100 SaaS Users in Europe(speartip.eu ↗)
    discuss
  17. Show HN: Salt – Human and AI Economies driven across E2E encrypted chats(saltapp.ai ↗)
    discuss
  18. Resy suspends VC for using AI agents to book reservations(businessinsider.com ↗)
    1comments
  19. OpenAI Agents Got into Link Shortener, Surfed Web, Called FBI with a Stranger's(kennethdegraff.com ↗)
    1comments
  20. Best Sugar Daddy Apps(gitlab.com/bestsugardaddyapps ↗)
    discuss
  21. How Bain Cap is deploying 1.6B(techcrunch.com ↗)
    discuss
  22. Separating Compilation from Attestation. C++ memory safety without disruption(gist.github.com ↗)
    discuss
  23. Show HN: Jev routing coding tasks to Grok Build or Codex Astra(github.com/jcpsimmons ↗)
    discuss
  24. The Mother Tongue of All Secrets(sundaylongread.com ↗)
    discuss
  25. Cohere launches confidential computing in model vault(cohere.com ↗)
    discuss
  26. How Claude is uplifting biomolecular modeling(anthropic.com ↗)
    discuss
  27. The Next Gene-Editing Technology May Also Be the Oldest(nytimes.com ↗)
    discuss
  28. How to Write with an LLM(sockpuppet.org ↗)
    discuss
  29. Alibaba Releases Qwen3.8-Omni-Flash(tokenstead.ai ↗)
    discuss
  30. Don't Drown the Dream – AI(jeffreylminch.substack.com ↗)
    discuss

PrismML Launches Bonsai 2 27B, Its Most Capable Model Yet

2 pointsby 1h agoprismml.com
1 comments
57m agoHN ↗

Across a 20-benchmark suite spanning reasoning, math, coding, instruction following, vision, and agentic tool use, Ternary Bonsai 2 27B achieves an aggregate score of 83.9, retaining 98.2% of Qwen3.8 27B’s performance of 85.4. The model reaches up to 143 tokens/second on NVIDIA GeForce RTX 5090.

https://huggingface.co/collections/prism-ml/bonsai-2

It's ~6 GB for "98.2% of Qwen3.8 27B's performance". It is significantly faster as well. Remains to be seen how well it works, I had significant issues with the previous Bonsai 27 compared to the parent Qwen model (documented here: https://humanparadox.org/local-vs-frontier-benchmarks-for-my...) but that also graded a bit worse (~95% of Qwen3.6 27B).