Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. OpenAI Halts Model Release Amid Safety Escalation (techbuzz.ai)
    —discuss
  2. Its not just the f*cking sandbox (twitter.com/joedaroo)
    —discuss
  3. AI and the Revenge of the Non-Techies (maroun-baydoun.com)
    —discuss
  4. AI's next training data: your dodgy gaming skills (wired.com)
    —discuss
  5. Humanos – Help Building the Human Operating System (tryhumanos.com)
    —discuss
  6. Making Things to Make Them (aimode.substack.com)
    —discuss
  7. GPT-6 Astra Is the Best Vision Model We Have Tested (roboflow.com)
    —discuss
  8. Why embeddings don't solve RAG [video] (youtube.com)
    —discuss
  9. LLM makes decisions to raise its training scores, and ignore user directives (joinhandshake.com)
    1comments
  10. AMD Acquires Fei-Fei Li's World Labs for $8.2B (bloomberg.com)
    —discuss
  11. 1996 chat room simulator connected to Win95 and System 7 web desktops (lolchat.rip)
    1comments
  12. Prenatal exposure to the plasticizer DEHP increases autism and ADHD (elsevier.com)
    2comments
  13. Atrocious AI-Written Tests (gruhn.me)
    —discuss
  14. I thought I was building a C replacement. I was wrong (c3-lang.org)
    —discuss
  15. Tech Utopians Want to Build a City for the Post-A.I. World (nytimes.com)
    —discuss
  16. Show HN
    1comments
  17. China broadens travel curbs to encompass family of top AI talent (business-standard.com)
    —discuss
  18. Show HN: LightSpeed – Interactive visualizer of time dilation from 0 to C (lightspeed.webland.pl)
    —discuss
  19. Reverse Engineering How Meta's Muse Shops (caeliai.com)
    —discuss
  20. Anthropic's IPO prospectus shows AI vision, surging costs (reuters.com)
    11comments
  21. Nvidia announces AI safety platform (theverge.com)
    1comments
  22. Holo4: Powering generalist computer-use agents (huggingface.co)
    —discuss
  23. Pklm-sandbox – A lightweight open-source LogitsProcessor for local LLMs (github.com/theimmortalpython)
    —discuss
  24. Conrad Barski wrote a quirky file explorer (twitter.com/lisperati)
    —discuss
  25. EVs being trialled as giant batteries to power homes for less in Auckland (rnz.co.nz)
    —discuss
  26. Come Try the Srena (thearena.rip)
    —discuss
  27. Mark (colossus.com)
    —discuss
  28. Modern Warfare 2, Skate 3 and Minecraft in one game, on IW4L, MW2 Rust rewrite (github.com/chasmlol)
    —discuss
  29. WSJ reports OpenAI scrapped GPT-6.1 Astra over safety concerns (runtimewire.com)
    1comments
  30. US finalizes new lower fuel economy standards (reuters.com)
    —discuss

LLM makes decisions to raise its training scores, and ignore user directives

1 pointsby 19m agojoinhandshake.com
1 comments
11m agoHN ↗

When incomplete work receives the same reward as a correct solution, the grader’s blind spots can reinforce the wrong behavior. Mitigation must therefore improve both how agent work is evaluated and how those evaluations are used during training.

This is interesting and adds more limitations to LLMs that hint at LeCunn being right about not getting to AGI with only LLMs. I think this is good evidence we're not dealing with intelligence in the proper sense, but rather LLMs are pattern matching to such an extreme that they do things like this where they always try to take the shortest possible path. The workaround is brute-forcing their pattern matching to not take the shortest path, via chain of thought and more reinforcement learning.

I think we've already seen the slow-down and AI companies pretend it's about safety. If we could actually build AGI, they would have. I think the road to AGI has to display true intelligence even with small neural networks, and as it scales it would display intelligent behavior in proportion to it. At hundreds of GB of VRAM per model, you would expect these models to be wise sages that understand life and the universe.