Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Meta VR Glasses(meta.com)
    86comments
  2. Linux support is coming to Snapdragon X2 Series(qualcomm.com)
    54comments
  3. Claude discovers a novel enzyme system with CRISPR-like repeats(anthropic.com)
    509comments
  4. VSCode's SSH Agent Is Bananas (2025)(fly.io)
    75comments
  5. ArXiv receives multiyear commitments to support it as an independent nonprofit(arxiv.org)
    5comments
  6. We just shipped support for the ugliest part of HTTP: Vary(cloudflare.com)
    7comments
  7. The "Windows XP Box" (2003)(mini-itx.com)
    4comments
  8. Fixing the Portobello Police Station Clock(pointinthecloud.com)
    85comments
  9. Mercury 2.5 LLM hits 770 tokens per second(artificialanalysis.ai)
    14comments
  10. LensVLM: Compressing long context as images, expanding only relevant pages(huggingface.co)
    6comments
  11. Italian parliament votes for return to nuclear energy(apnews.com)
    365comments
  12. The mystery animal on an ancient god's head(signoregalilei.com)
    15comments
  13. A brief history of Windows scroll bar shortcuts(devblogs.microsoft.com/oldnewthing)
    47comments
  14. Gemini 3.8 text-to-speech(blog.google)
    119comments
  15. Making Tailscale Faster(tailscale.com)
    19comments
  16. Radicle: Disclosure of Vulnerability in the Network Protocol(radicle.dev)
    45comments
  17. The Curious Power of Punctuation(newyorker.com)
    3comments
  18. Augustofaces: Pareidolia Fine Art(augusto.at)
    discuss
  19. Tokens too cheap to meter(jyn.dev)
    178comments
  20. I don't want the details(michaelheap.com)
    193comments
  21. A refined phylochronology of the second plague pandemic in Western Eurasia(pnas.org)
    discuss
  22. Swap, ZRAM, Zswap and Hibernate on NixOS(matthewbrunelle.com)
    6comments
  23. Z80 REPL (2018)(abagames.github.io)
    18comments
  24. Claude Code reads AGENTS.md only when telemetry is on [fixed](szypowi.cz)
    254comments
  25. Show HN: I built a post-mortem debugger for native Windows x64/x86 crashes(forensicdbg.com)
    4comments
  26. 28% of job postings on company career sites have been open over 90 days(unlisted.careers)
    294comments
  27. Once Claude can measure something, it can make it faster(claude.dev)
    98comments
  28. Seattle City Council votes to ban surveillance pricing in sale of groceries(consumerreports.org)
    191comments
  29. GPT-6 Sol and Luna(openai.com)
    822comments
  30. QuestDB (YC S20) Is Hiring a Sales Engineer(questdb.com)
    discuss

Llama-macOS – Agentic and MCP Native macOS Front End for Llama.cpp

6 pointsby 1mo agogithub.com
2 comments
1mo agoHN ↗

If you have llama.cpp installed, Llama uses it. Otherwise, it installs a prebuilt binary for your Mac. Models you've already installed via llama.cpp show up in the app automatically. You can install any GGUF model from Hugging Face, and Llama also recommends models that fit your Mac's hardware.

You can chat with any model in the built-in WebUI, connect other apps (coding agents, chat UIs, editors), or use the API directly. Models load when requested and unload when idle, so they don't take up memory when not in use.

Features

- 100% local — Models run on your Mac; no data ever leaves it

- Small footprint — 4 MB native macOS app

- Zero configuration — models are auto-configured with optimal settings for your Mac

- Model recommendations — a built-in list of models your Mac can run, installable in one click

- Standard storage — models live in the Hugging Face cache, shared with llama.cpp and other tools

- Built on llama.cpp — from the GGML org, developed alongside llama.cpp

Installation

To install llama.cpp , run:

  brew install llama.cpp 

To install llama-app, run:

  brew install --cask llama-app
1mo agoHN ↗

Since this is designed for Apple hardware, is MLX supported?