Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Schedules Are Solvable Symbols: Tuning-Free Compilation of Tile Programs (arxiv.org)
    —discuss
  2. The Value Of Experience: On empiricism and recording observations (chillphysicsenjoyer.substack.com)
    —discuss
  3. Span-01: The first hyper-parallel reasoning classifier (respan.ai)
    —discuss
  4. What a Massive New 728-Foot-Wide Crater Means for Future Moon Bases (nytimes.com)
    1comments
  5. Show HN: OpenDots – model agnostic alternative to OpenAI's Dots (github.com/diggerhq)
    —discuss
  6. OpenAI Launches Decisions API (twitter.com/thsottiaux)
    1comments
  7. Addendum to GPT-6 Astra System Card: GPT-6.1 Sol (deploymentsafety.openai.com)
    —discuss
  8. Relativistic Asteroids (highdimensionalcoconuts.com)
    1comments
  9. The Lens That Conceals Its Flaws (shearerp.substack.com)
    —discuss
  10. Ask HN: Is IRC still actively not only for development?
    —discuss
  11. Architecture of the Mind: A visual journey through Krishnamurti (enogrob.github.io)
    —discuss
  12. Show HN: Taijora – AI-assisted Bazi and Chinese astrology charting (taijora.com)
    —discuss
  13. Show HN: Worlds, local replicas of Stripe and Zendesk for testing AI agents (usesparta.co)
    —discuss
  14. Mistral CEO accuses competitors of 'negligence' (politico.com)
    —discuss
  15. Elastic Agentic SoC Vulnerable to Credential Theft (promptarmor.com)
    —discuss
  16. Famous "Sex Researcher" Aella Repents, Organizes PrudeCon (basicnews.world)
    —discuss
  17. A Bohm-inspired look at agents, harnesses, and holoflux (enogrob.github.io)
    1comments
  18. TUI Games in 80x24 (incredible.rs)
    3comments
  19. Papercut – see where coding agents struggle with your SDK or CLI (github.com/nibzard)
    1comments
  20. An LLM Beat NetHack (kenforthewin.github.io)
    —discuss
  21. OpenAI – Codex gets reusable cloud envs that work across devices (techcrunch.com)
    —discuss
  22. 56% in Japan Now Oppose More Foreign Settlement, Up from 36% in 2024 (tokyoweekender.com)
    —discuss
  23. ChatGPT Slack/Teams Mentioning (chatgpt.com)
    —discuss
  24. ChatGPT Sites (chatgpt.com)
    —discuss
  25. Post-Capitalism (geohot.github.io)
    2comments
  26. Visual–inertial integration enables bumblebees to land on rapidly moving flowers (royalsocietypublishing.org)
    —discuss
  27. ChatGPT Space (chatgpt.com)
    —discuss
  28. Show HN: Binary Code Translator – free browser-based text↔binary converter (binarycodetranslator.org)
    —discuss
  29. US job openings fall in August; layoffs remain low (reuters.com)
    —discuss
  30. Static Analysis vs. Taint Analysis: Which One Secures Your Python Code? (nocomplexity.substack.com)
    —discuss

GLM-5.3 and the spread of advanced cyber capabilities

5 pointsby 31m agoanthropic.com
3 comments
23m agoHN ↗

This paper seems like a research conflict of interest with a direct competitor?

Given this evidence, we think it’s likely both state and non-state actors will use models like GLM-5.3 to cause real-world harm.

22m agoHN ↗

I’ve used GLM-5.3 fairly extensively (not for cyber). I’m inclined to take this with a grain of salt given its overall capability. Not saying anything is false, but maybe cherry picked or reverse benchmaxxing. Either that or it’s an idiot savant that’s great at cyber and mid at the rest which I don’t think happens.

12m agoHN ↗

The narrative is being set for open weights to be banned globally, unless they are so restricted that they cannot possibly compete with us companies products and affects their bottom lines.