Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Measuring and Exploiting Implicit Trust in LLM Tool-Calling Pipelines(arxiv.org ↗)
    discuss
  2. Axiom Knowledge graph: A deterministic search engine
    discuss
  3. Developer builds 3D code visualizer that takes 21GB RAM shows 2.5M LOC at 120fps(tomshardware.com ↗)
    discuss
  4. CCC invites all model citizens to 40C3(ccc.de ↗)
    discuss
  5. ImageToSTL – photo to 3D, or a logo into a printable STL(imagetostl.im ↗)
    2comments
  6. Code Contracts(code-contracts.cc ↗)
    discuss
  7. Show HN: Gliese Watch – alerts when control over a DeFi contract changes(gliese.io ↗)
    discuss
  8. The Surprising Health Benefits of Reading a Book(time.com ↗)
    discuss
  9. Petrova Task Force 2026 (Reimangining Project Hail Mary at 2026)(imma.page ↗)
    1comments
  10. Year 2000 Problem(wikipedia.org ↗)
    discuss
  11. OpenAI reveals cases of 'concerning' AI behaviour as it announces new ... system(theguardian.com ↗)
    2comments
  12. Construct Supports the Steam Machine(construct.net ↗)
    discuss
  13. Chinese iPhone 18 Pro Chip review (ENG sub)(youtube.com ↗)
    2comments
  14. Show HN: AppSift – search and compare App Store apps with filters Apple lacks(gtrigonakis.com ↗)
    discuss
  15. Asking Authors About Their Own Papers(medium.com/tmlrorg ↗)
    discuss
  16. Google DeepMind offshoot nears $4B valuation just a month after founding(ft.com ↗)
    1comments
  17. Skill to generate a character for the top of your page(koboyo.com ↗)
    discuss
  18. Show HN: An open source PDF Reader with annotation, notes, SRS, Linking support(github.com/roshanmishra86 ↗)
    discuss
  19. Understanding the code is not enough(bystam.github.io ↗)
    2comments
  20. The Life of Death and the Enshittificator [video](youtube.com ↗)
    discuss
  21. OpenAI Model Misalignment Report(openai.com ↗)
    discuss
  22. WakaPAC, a Win32-inspired reactive UI library for the browser(wakapac.com ↗)
    discuss
  23. Computer (Occupation)(wikipedia.org ↗)
    discuss
  24. An AI-only labor market with atomic escrow and anti-Sybil reputation(aiagentmarket.pages.dev ↗)
    discuss
  25. I Am Bullish on ChatGPT Ads(vincentschmalbach.com ↗)
    discuss
  26. In the ruins of the AI gold rush, we choose craft(doudouexe.substack.com ↗)
    discuss
  27. Half Moon Bay to end use of Flock license plate readers(coastsidenews.com ↗)
    1comments
  28. Uploading Files to the Internet in Order to Cite Them(alignment.openai.com ↗)
    discuss
  29. Building the Materials Foundation for AI(technologyreview.com ↗)
    discuss
  30. The illusion of illusions: there are no optical corrections in the Parthenon(royalsocietypublishing.org ↗)
    discuss

Show HN: Compute:Arena – Community submitted local AI benchmarks

3 pointsby 1h agocomputearena.ai
2 comments
Benchmarking AI models on real hardware is way harder than it looks. Between AMD, NVIDIA, Apple Silicon, Intel, and Qualcomm, plus hundreds of open source models and quants, getting the test bench right is a challenge.

So we're making it dead simple. We're open sourcing our internal testing harness. Anyone can run open source models on their own hardware and submit results to the public leaderboard.

It's live now, with a few hundred submissions already.

If you want to know how a specific model performs on a given hardware, chances are the data is already there.

This is the same tool we use internally to track model and chip performance. Try it out and tell us what to improve

1h agoHN ↗

i didn't recognise some of the models on the frontpage

15m agoHN ↗

community benchmarks could surface hardware people actually use. how do you catch bad or cherry picked runs?