Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Meta VR Glasses(meta.com)
    85comments
  2. Linux support is coming to Snapdragon X2 Series(qualcomm.com)
    52comments
  3. Claude discovers a novel enzyme system with CRISPR-like repeats(anthropic.com)
    509comments
  4. VSCode's SSH Agent Is Bananas (2025)(fly.io)
    74comments
  5. ArXiv receives multiyear commitments to support it as an independent nonprofit(arxiv.org)
    4comments
  6. We just shipped support for the ugliest part of HTTP: Vary(cloudflare.com)
    7comments
  7. The "Windows XP Box" (2003)(mini-itx.com)
    4comments
  8. Fixing the Portobello Police Station Clock(pointinthecloud.com)
    85comments
  9. Mercury 2.5 LLM hits 770 tokens per second(artificialanalysis.ai)
    14comments
  10. LensVLM: Compressing long context as images, expanding only relevant pages(huggingface.co)
    6comments
  11. Italian parliament votes for return to nuclear energy(apnews.com)
    365comments
  12. The mystery animal on an ancient god's head(signoregalilei.com)
    15comments
  13. A brief history of Windows scroll bar shortcuts(devblogs.microsoft.com/oldnewthing)
    47comments
  14. Gemini 3.8 text-to-speech(blog.google)
    118comments
  15. Making Tailscale Faster(tailscale.com)
    19comments
  16. Radicle: Disclosure of Vulnerability in the Network Protocol(radicle.dev)
    45comments
  17. The Curious Power of Punctuation(newyorker.com)
    3comments
  18. Tokens too cheap to meter(jyn.dev)
    178comments
  19. A refined phylochronology of the second plague pandemic in Western Eurasia(pnas.org)
    discuss
  20. I don't want the details(michaelheap.com)
    193comments
  21. Z80 REPL (2018)(abagames.github.io)
    18comments
  22. Claude Code reads AGENTS.md only when telemetry is on [fixed](szypowi.cz)
    253comments
  23. Swap, ZRAM, Zswap and Hibernate on NixOS(matthewbrunelle.com)
    6comments
  24. Show HN: I built a post-mortem debugger for native Windows x64/x86 crashes(forensicdbg.com)
    4comments
  25. 28% of job postings on company career sites have been open over 90 days(unlisted.careers)
    294comments
  26. Once Claude can measure something, it can make it faster(claude.dev)
    97comments
  27. Seattle City Council votes to ban surveillance pricing in sale of groceries(consumerreports.org)
    190comments
  28. GPT-6 Sol and Luna(openai.com)
    822comments
  29. QuestDB (YC S20) Is Hiring a Sales Engineer(questdb.com)
    discuss
  30. Stripe's Knowledge AI Platform(stripe.dev)
    108comments

Llama-macOS – Agentic and MCP Native macOS Front End for Llama.cpp

6 pointsby 1mo agogithub.com
2 comments
1mo agoHN ↗

If you have llama.cpp installed, Llama uses it. Otherwise, it installs a prebuilt binary for your Mac. Models you've already installed via llama.cpp show up in the app automatically. You can install any GGUF model from Hugging Face, and Llama also recommends models that fit your Mac's hardware.

You can chat with any model in the built-in WebUI, connect other apps (coding agents, chat UIs, editors), or use the API directly. Models load when requested and unload when idle, so they don't take up memory when not in use.

Features

- 100% local — Models run on your Mac; no data ever leaves it

- Small footprint — 4 MB native macOS app

- Zero configuration — models are auto-configured with optimal settings for your Mac

- Model recommendations — a built-in list of models your Mac can run, installable in one click

- Standard storage — models live in the Hugging Face cache, shared with llama.cpp and other tools

- Built on llama.cpp — from the GGML org, developed alongside llama.cpp

Installation

To install llama.cpp , run:

  brew install llama.cpp 

To install llama-app, run:

  brew install --cask llama-app
1mo agoHN ↗

Since this is designed for Apple hardware, is MLX supported?