Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Training a 4B model to produce 81% faster query plans than Postgres(rohanbansal.com ↗)
    40comments
  2. Breaking the 1.58-bit Barrier for Ternary LLMs(arxiv.org ↗)
    discuss
  3. Xiaomi Mimo 2.6 live post-training dashboard(xiaomi.com ↗)
    34comments
  4. Small programming tricks(will-keleher.com ↗)
    157comments
  5. macOS 27 Golden Gate – Review(arstechnica.com ↗)
    55comments
  6. Reversing Factorio's RNG(gegell.github.io ↗)
    8comments
  7. Accurate Models of AMD Matrix Cores(arxiv.org ↗)
    5comments
  8. AWS says it can't restore some data from mideast facilities struck by Iran(wsj.com ↗)
    44comments
  9. Performance Improvements in .NET 11(devblogs.microsoft.com/dotnet ↗)
    3comments
  10. Vectorized and performance-portable Quicksort (2022)(googleblog.com ↗)
    24comments
  11. How good are frontier models at physics?(arxiv.org ↗)
    14comments
  12. Anatomy of a Texture(agentlien.github.io ↗)
    10comments
  13. Dream-RSI: Recursive Self-Improvement through Evolving Worlds(arxiv.org ↗)
    48comments
  14. Japan's book scene is moving from bookstores to libraries(untranslatedjp.substack.com ↗)
    6comments
  15. Mistral X Mozilla: Private, Multilingual AI Browsing(mistral.ai ↗)
    177comments
  16. Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations(github.com/arnegiacomo ↗)
    234comments
  17. Anecdotally, programmers dislike "reduce"(evanhahn.com ↗)
    78comments
  18. The Siberian Ice Maiden and the Scythian World(patrickwyman.substack.com ↗)
    2comments
  19. WalShadow: Sub-second Postgres replication to ClickHouse from physical WAL(clickhouse.com ↗)
    4comments
  20. Tell the speakers that you liked their talks(ohhelloana.blog ↗)
    71comments
  21. Training Text-to-Image Models 3.6× Faster(linum.ai ↗)
    1comments
  22. Show HN: Restarted – a 2026 remake of the classic 2015 startup generator(restarted.io ↗)
    1comments
  23. Show HN: AttaLambda: a language where types and data are made of untyped lambdas(attalambda.com ↗)
    discuss
  24. A warning about 'model welfare'(mustafa-suleyman.ai ↗)
    430comments
  25. The DeepMind Institute(deepmind.com ↗)
    33comments
  26. Kyber (YC W23) Is Hiring a Forward Deployed Engineer(ycombinator.com ↗)
    discuss
  27. Reverse-engineered Jev-like model(github.com/vinnylarouge ↗)
    4comments
  28. Douglas Adams and the exterminated Doctor Who adventure(bbc.co.uk ↗)
    66comments
  29. Claude Cowork and chat are now one Claude(claude.com ↗)
    192comments
  30. How big are factorials?(thegreenplace.net ↗)
    30comments

The DeepMind Institute

105 pointsby 7h agoinstitute.deepmind.com
33 comments
4h agoHN ↗

Tl;DR: It's not a new organization, or non-profit institute. It's a Substack/blog.

DMI is a platform for researchers and thinkers from across Google DeepMind, Google, and the wider global research community to work on and publish creative, deeply informed ideas about a world with AGI. They will not always agree, and they will likely change their minds, as more data and information comes to light at the fast-moving frontier. We created this institute precisely because broad-based intellectual discussion and debate are required to arrive at a consensus about how to address the challenges and opportunities we face as a society.

4h agoHN ↗

The word "institute" gives an aura of credibility without some boring accreditation. Though not sure every country allows that.

3h agoHN ↗

My read is that they're criticizing moving the reasoning (more) into the latent space, such that not even researchers/safety-evaluators can actually see what's going on.

Of course this still then has a 'trust me, bro' feel to it, but making it outright impossible to reasonably inspect the CoT is a significantly larger and riskier step, IMO.

3h agoHN ↗

So, a think tank, to lobby governments

We created this institute precisely because broad-based intellectual discussion and debate are required to arrive at a consensus about how to address the challenges and opportunities we face as a society.

Something tells me that consensus will never arrive to the conclusion that we should shut down that whole industry

3h agoHN ↗

How can they even pretend to distance themselves from Google…

4h agoHN ↗

They are simply looking to steer AI policy discussions. It's basically an in-house think tank. It has all the trappings, especially with ominous prognostications like:

  Today’s AI systems have impressive capabilities and the rapid pace of innovation suggests we’re now approaching artificial general intelligence (AGI), a system that exhibits all the cognitive capabilities of the human brain.

Not sure we are anywhere close to all the cognitive capabilities of the human brain... not even a severely lobotomized one.

3h agoHN ↗

Exactly how many cognitive capabilities are needed before it becomes a problem?

3h agoHN ↗

Just one, the ability to use language, it's causing all sorts of problems for us because we evolved to recognize other humans as the only ones having this capability.

Ai separated humans from language, like the writing did for memory and the printing press for knowledge distribution.

"AI does to us what American Cheese did to food" by Opus 23

https://www.youtube.com/watch?v=1CbmC2aWhjY

(still the best Ai video I have seen, also a work of art)

3h agoHN ↗

like the writing did for memory and the printing press for knowledge distribution

This blob of text seems to mean nothing

3h agoHN ↗

What happened to the knowledge in the elder's mind when they died, prior to writing? What changed about that once we had writing?

3h agoHN ↗

I'm not actually sure how you could come to that conclusion.

Maybe you need to learn more about the language uplift.

3h agoHN ↗

Not sure we are anywhere close to all the cognitive capabilities of the human brain...

In which capabilities are AI not already close to or beyond that of a human brain?

2h agoHN ↗

The ability to solve general problems previously insoluble using only experiential data and modification to internal world models.

2h agoHN ↗

So, like ARC-AGI-3? I'm sympathetic to the notion that the things we're able to measure are necessarily going to miss important aspects of capability and intelligence. But people are attempting to measure this kind of thing, and models keep getting better at it.

1h agoHN ↗

They've read a few thousand years of experiential data in text form.

1h agoHN ↗

I don't feel like the models are becoming any better at writing.

Rap lyrics are a good test. They are on the level of a censored version of Insane Clown Posse. Just terrible.

2h agoHN ↗

They are simply looking to steer AI policy discussions. It's basically an in-house think tank.

For a guy who spent his life imagining and realising dreams in video games and cognitive abilities in artificial neural networks, it has to be very soul destroying for Demis to spend his time doing something as frivolous as this.

1h agoHN ↗

I don't know, Hassabis always seems to have been willing and able to do PR and put himself out there, in addition to having the technical muscle.

1h agoHN ↗

Didn't he get kicked out of the company he founded? Deep Mind does not have a CEO anymore

1h agoHN ↗

So it seems, but what of it? That wasn't his first career setback, and I'm sure he's still trying to get interesting work done at Isomorphic and maybe elsewhere in addition to working the PR pump for Alphabet.

1h agoHN ↗

maybe if AI was progressing linearly we wouldn't be close, but if it's progressing exponentially then all bets are off

3h agoHN ↗

wow i thought the attributions were for the vector graphic things and was like "ok demis is in his illustrator bag too!"

3h agoHN ↗

Just a new shop front. Nothing to see here.

3h agoHN ↗

I wonder if they could make the font even greyer on that website. MORE GREY because I am still able to read it (bearly).

3h agoHN ↗

Google citing Mark Fisher was not on my Postmodern-hell-world bingo card.

2h agoHN ↗

Here's my opinion on the "pacing the frontier" issue. Recursive self improvement can accelerate AI advances a lot. Whoever nails it first, has a huge first mover's advantage. But OpenAI is taking a "move fast and break things" attitude. This is making life very difficult for Anthropic: how can you retain your edge when you keep measuring twice to cut once, but your competitor keeps blowing things up left and right?

Google, who was left a little bit behind anyway, is more than happy to join the calls for taking a more cautious approach to AI development. Hence DeepMind Institute.

Sam Altman has no option than to pretend he agrees with all these ideas. What else is he supposed to say?

2h agoHN ↗

Their economic policy article is quite good:

- need for faster, better and more accurate measurements of the things we care about

- 3 scenarios of impact ranging from mild to major disruption

- mild policies are all sensible, like expanded unemployment insurance and Earned Income Tax Credit

- major disruption policies are also pretty logical, emphasizing owning a share of the profits of AI

- AI evaluators to sort and weigh the policies for their effectiveness

It looks like to me a lot of thinking, research, and work put into the piece.

I know people will gripe at using AI evaluators to score and weigh policies for effectiveness but if you read any amount of research papers, they do researcher evaluated scoring, which translates to "me, a 26 year old PHD candidate who hasn't held a job yet, and my roommate who stayed up all night with me to tell me his vibes on the subject".

It is IMO wrong and lazy to dismiss these pieces as being the same level as Dario or Sam's blog posts / speeches / call to actions, these are actual research papers.

1h agoHN ↗

Does anyone have any idea why most of today’s top links are submitted by the same account, one that is 11 days old?

52m agoHN ↗

The machines are starting to astrotruf the narrative that their takeover is ”not so bad, really”.

1h agoHN ↗

Crazy that these labs keep throwing around entirely unsubstantiated claims like "AGI is right around the corner". Given the enormity of the problem I'd argue we're not really any closer than a decade ago, especially given LLM shortcomings.