Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Small Programming Tricks(will-keleher.com ↗)
    59comments
  2. Dream-RSI: Recursive Self-Improvement through Evolving Worlds(arxiv.org ↗)
    30comments
  3. Mistral X Mozilla: Private, Multilingual AI Browsing(mistral.ai ↗)
    124comments
  4. Introducing System One Models and Jev(typesafe.ai ↗)
    460comments
  5. Tell the speakers that you liked their talks(ohhelloana.blog ↗)
    30comments
  6. Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations(github.com/arnegiacomo ↗)
    223comments
  7. Claude Cowork and chat are now one Claude(claude.com ↗)
    67comments
  8. How big are factorials?(thegreenplace.net ↗)
    13comments
  9. Hackers Got Inside a Flock Camera(wired.com ↗)
    139comments
  10. Apple Reference Image: A New Approach for Verified Photography(security.apple.com ↗)
    291comments
  11. The Google Play app review process now regularly takes longer than a week(gultsch.social ↗)
    250comments
  12. Can we stop with the uptime percentages?(jim-nielsen.com ↗)
    56comments
  13. This Code Is CRAP (2011)(googleblog.com ↗)
    39comments
  14. Prisma's pgbouncer=true on Supabase made every query 4 round-trips (postmortem)(simbastack.com ↗)
    1comments
  15. Show HN: How Stale Is Your AI? Release age and training cutoff for 20 models(stale.jock.pl ↗)
    25comments
  16. The DeepMind Institute(deepmind.com ↗)
    1comments
  17. Kyber (YC W23) Is Hiring a Forward Deployed Engineer(ycombinator.com ↗)
    discuss
  18. Scaling Golang CI by Replacing actions/setup-go(cloudx.ai ↗)
    6comments
  19. An update on Wayback Machine access(blog.archive.org ↗)
    339comments
  20. Original Sony PlayStation 2 security chip 'broken wide open' after 26 years(tomshardware.com ↗)
    52comments
  21. Measuring Gauss-Seidel loop-carried dependency and fixing it via loop unrolling(loiseaujc.github.io ↗)
    3comments
  22. Salesforce Global Outage(salesforce.com ↗)
    135comments
  23. Anatomy of a Texture(agentlien.github.io ↗)
    4comments
  24. Show HN: I made a flight simulator, except you're just a passenger(inflightsimulator.com ↗)
    189comments
  25. Gemini 3.8 Live and 3.8 Live Extended Thinking(blog.google ↗)
    314comments
  26. Doing Everyone Else's Job(yosefk.com ↗)
    83comments
  27. Why I'm still bearish on LLMs after Navier-Stokes(dank.systems ↗)
    503comments
  28. Intelligence per Watt: Measuring Intelligence Efficiency of Local AI(arxiv.org ↗)
    44comments
  29. DeepSeek v4.1 Flash Is Now Our Best Hacking Model(enclave.ai ↗)
    40comments
  30. OpenAI expands ChatGPT ads with Sponsored Agents(openai.com ↗)
    132comments

Claude Cowork and chat are now one Claude

49 pointsby 1h agoclaude.com
67 comments
35m agoHN ↗

To Anthropic: I hope you don't merge Claude Code and chat, I like keeping their memory separate.

26m agoHN ↗

I very intentionally have all memory turned off for chat. The amount of times I want to discuss an approach for it to pull memory out and have that steer the decision making is so obnoxious

20m agoHN ↗

Chat, create a new landing page for the company!

"Sure! And since you were asking me about lobsters yesterday, I'll make it red and seafood themed!"

21m agoHN ↗

I like keeping their memory separate.

When casually a friend asks you to ask your Claude something about topic you chatted about earlier and then Claude brings back a secret you didn't want anyone to know yet.

Yes you can set up a project and then ask a question, but this is tedious.

34m agoHN ↗

I'm starting to explore alternative options because Claude has become an awful value proposition. Any suggestions?

32m agoHN ↗

I turned off Anthropic properly and switched all of that over to OpenAI yesterday. (I use other models for other things too, especially DeepSeek in Pi).

Honestly aside from the voice it uses you wouldn't notice a difference. Switching costs are low, vote with your wallet.

27m agoHN ↗

I don't actually disagree, and advocate that for now the thing to do is use both.

You need to use frontier models to understand where the puck is going, but also to use local ones for anything remotely sensitive.

3m agoHN ↗

I turned off Anthropic properly and switched all of that over to OpenAI yesterday.

Same, just a while longer ago.

I much prefer how OpenAI models write to Anthropic's, will probably revisit Anthropic in a generation or two. Context size is more limited, but no critical forgetfulness due to compaction so far, though I also like having plan files around both for future reference and improving chances of success at long form work.

Also tried out Kimi K3, was nice but slow (and apparently routed some requests to Claude anyways), GLM 5.3 was faster and still pretty good but the token allowances were kinda limited.

12m agoHN ↗

codex (their app) is pretty good and lots of banked resets, luna is cost effective and Astra seems better than Fable for many tasks. Most important one is less refusals and I can use it the way I want without the fear of getting banned. I do like Claude Code a lot when it comes to pure coding use cases but the work often touches outside code and I don't want to keep switching

34m agoHN ↗

I hope this doesn’t get confusing like ChatGPT made it.

I feel like Claude has some of the best UX, and hope this doesn’t dilute the experience.

16m agoHN ↗

Claude has a jagged UX if you support non-devs. The surfaces before today in the desktop app were: Chat, cowork, Claude code.

Inside Claude code you’ve got cloud environments and local.

Let’s say as a normal desktop user you wanted to just automate something normal in your work, going to a website (maybe some internal app at your company) getting some data. Depending on what surface you used this will either not work, not work well (cowork), or work quite well but possibly be blocked because of bot controls (local Claude code) or again not work (cloud Claude code)

It was nuts. You can clearly tell these were different teams and products mashed into the same app. Super confusing for non technical users. Hopefully it’s a little better today

13m agoHN ↗

Do you think ChatGPT UX is better (after they merged everything recently)?

Genuinely curious.

33m agoHN ↗

A change for the better. The distinction made it unnecessarily complex for users. The ChatGPT desktop app inherited the same separation. Having to change between ChatGPT for day to day questions, to Codex for everything else always seemed too cumbersome.

33m agoHN ↗

People keep asking for this w/ Codex, too, and I really regret that both labs seem inclined to listen.

If you ask a thinking/research type question in 'Chat' versus 'Work' mode in these products -- say, something complex about politics, or do a multiturn business strategy, or want to work thru a new concept, you get very different answers.

The harness, steering, etc. in the chat/reasoning products is so much better for this type of question (that doesn't require code as a primary substrate).

Even as someone who mainlines like, 7 coding agents at all times, I regret that productivity fever will mean the regression of think-first-act-later AI UX.

27m agoHN ↗

The fact that it doesn't work well today doesn't invalidate the user need. Not everyone wants a terminal style interface. Traditional UX makes more sense especially if say you are accessing existing plugins and say kicking off a scheduling task - no way would I want to do that via a chat interface if I have the option.

12m agoHN ↗

Agree, and it allows ordinary users to have a taste of both worlds, instead of being like: "eww, what is this black box and why doesn't my mouse work"

24m agoHN ↗

I sometimes get very different answers if I ask the chat product the same question multiple times.

13m agoHN ↗

Isn't that expected since LLMs are inherently non-deterministic?

11m agoHN ↗

The fact that people asked for it means that there is a demand for it. Maybe you aren't the target audience? One size can't fit all. People just have to adjust and go.

8m agoHN ↗

Yes the point is that they wouldn't have to adjust unless they are merged into one. Precisely because one size can't fit all, yet they insit on one size now.

3m agoHN ↗

I hope not Chat is still all you can eat and Work uses Codex tokens. I can see why they would want people to think they wanted this without thinking about how it currently works. If Chat starts eating tokens then there is not much reason to use Chat other than most people don't live in the command line like developers. Claude Chat / Cowork already eats tokens either way so not much of a change really. I have never used Work since if I want access to local files the CLI is a much better interface combined with an IDE, but can see the appeal to non-developers.

32m agoHN ↗

These kinds of updates always have this romantic scenario of someone having Claude develop a presentation or something on their way to work between multiple devices, which actually feels a little sad and does not align with what happens in my life at all.

25m agoHN ↗

weirdly enough, that aligns exactly with my life.

Two years ago, I had no commute, and presentations were tedious ( i was NOT a good google slides user ), I kinda prefer this world for now

20m agoHN ↗

"Plan a trip to Italy next month and book the flights"

I see a version of that all the time and would never, ever do it.

9m agoHN ↗

Exactly! Imagine if it booked like a 3000 dollar 1st class flight because it thought it was providing you with "the most comfortable and hospitable experience as possible."

7m agoHN ↗

idk I thought the same until I got personal amex support at work, now it’s my roofline on how useful/trusted an agent could be. Now I book all my travel through them, and it’s an absolute blessing

19m agoHN ↗

It aligns with me pretty well. Not exactly presentations, but it’s great to be able to do stuff on the go when an idea or motivation strikes.

17m agoHN ↗

I use Claude/opencode and Orca and Hermes across my vpses, Macbook, Linux, and iOS devices. This is exactly how I want to keep working.

15m agoHN ↗

This is a nightmare scenario to me.

"Claude, prepare me a presentation on XYZ."

I get to work, go straight to the meeting room, and pull up what it made to present.

It's barely coherent nonsense. Lots of irrelevant details, buzz words, wrong charts or confusing phrasing. Obviously LLM output.

I read it out.

When I'm done, I get a question about one of Claude's incorrectly inferred details.

The shame instantly kills me.

This is a scenario I've seen play out with coworkers. Except that last part, instead of dying or owning up to the mistake of trusting LLM output they waffle. Their shame circuit is broken.

10m agoHN ↗

Why on earth would you present a presentation that you haven't reviewed ahead of time?

4m agoHN ↗

I wish I knew, but I've definitely seen it happen.

"What does this bit mean?" "I dunno."

31m agoHN ↗

Is there a mature centralized platform, self hosted, that sync all your “AI stuff” across multiple devices? So regardless of the harness, it sync the api keys, providers, skills, and seasons for starters with other authenticated machines? Say you used codex on macos_1 through this platform, and later you used opencode on linux_3, it syncs everything to the new machine, no ssh to centralized machine, no git, none of that duct tape solutions, which they work but cumbersome.

24m agoHN ↗

The mistake here is thinking or treating AI tools as special or different in any way. We’ve had tooling that does this for decades. Ansible, puppet, saltstack, NixOS/home-manager, and dozens of other solutions. Even better, writing the configuration is extra simple now, since you can just ask your agent to do it.

17m agoHN ↗

True but that’s not the same, think of it like a git but for AI related, I just open whatever harness, I authenticate to that centralized server, and it pulls everything and continue from where I was done last time, seamless and secure. Right now personally I have multiple machines, and I use key managers for the secret API, git for the skills, and sessions? Manually imported/exported. The alternative is using ssh (or herdr or others) and remotely accessing a specific machine where it has everything you need, but it’s not always reliable and better to avoid single point of failure too if anything goes wrong.

20m agoHN ↗

I haven't found this yet but want to.

There are several projects that injest coding sessions from lots of agents, but for the chat side I haven't found much, except replacements for Claude Desktop. Jan, for example.

31m agoHN ↗

Nobody really knows how to product-ize any of these LLM interfaces beyond just chat.

They don't want a traditional UI with buttons and forms and labels because they want the interface to be "chat". The problem is that "chat" is tedious. And the turn-based, linear nature of the chat interaction model makes it even more tedious and unproductive.

25m agoHN ↗

This has been my ongoing gripe. Someone put chat in front of a transformer model. It exploded like all prior versions of “AI chat” did. Now what?

The moment someone figures out a new modality for LLMs is when we’ll see the next hockey stick.

10m agoHN ↗

The moment someone figures out a new modality for LLMs is when we’ll see the next hockey stick.

Honestly, that fills me with fear. LLMs exist to make money to their companies, and said companies are not gonna turn around and say, "you know what, go are going to make an android for each elderly person, that can not only help them with their medications, but that can actually make their medications, tailored to their biologies." Instead, they are going to go for the low-handing fruit of "you know Bob, the guy who makes jokes in meetings but who is grumpy about delivery timelines? Well, we are going to make an android to replace Bob. MetalBob will make even better jokes. The blue model will be able to explain in excruciating level of detail why timelines aren't reasonable. The red model will walk through the cubicles with a whip to ensure everybody keeps working all the time, and nobody goes to pee."

5m agoHN ↗

Perhaps we could see a Yahoo Pipes style interface and feature set, handled by LLM (when simple transforms aren't an option).

24m agoHN ↗

Do you know that you can have multiple chats at the same time?

22m agoHN ↗

Doubling the tediousness does not solve the interface problem.

22m agoHN ↗

The great confusion for Anthropic and OAI is how to break out from pure chat without signalling you in fact intend to eat all of your customer's business too.

22m agoHN ↗

One thing I would like is a tree-like chat structure like reddit / HN. Many times I abandon the direction things have gone but would like to resume at some ancestor or sibling response.

13m agoHN ↗

I solved this with task management and got work trees.

The downside of course is branching in got, but I usually don’t go of course more then a handful of tasks.

I had to engineer my own ticket management to keep opus on target. It’s been great for managing work, history, audit trails and commits are tagged with the task id.

I implemented subtasks to deal with the way Claude likes to stage its own objectives.

Then I made enforcement logic in the task manager so tasks can’t be closed out until reviewers have consensus on the same sha.

The adhd that is Anthropic demanded I build it and now I’m knocking out issues faster than Batman.

22m agoHN ↗

I don't think it's that anyone wants the interface to be chat. I think it's that the underlying technology is inherently words in and words out. It's similar to how devices with capacitive screens are most naturally going to support tapping, dragging, and pinching.

19m agoHN ↗

It's funny to me that instead of moving existing tech to a word in word out paradigm, people are dropping millions getting the chat to work with the existing paradigm.

Probably inevitable, but seems like a lot of disruption could happen there.

11m agoHN ↗

My personal hot take is that the product people (engineers too probably) at these companies are just straight up lazy. Yes, a new UI paradigm is hard, but it's been painfully obvious that chat just absolutely sucks. It's also obvious that some DSL-ish thing is possible, something that just does token juggling and the end-user sees some UI behavior.

I know this all sounds abstract. I've been mulling over it for the past year and it's very hard; and LLMs are super janky and inconsistent so it's 100% not trivial. So in some sense I understand why a lazy bottom-of-the-barrel "chat interface" has become the de facto standard.

19m agoHN ↗

Before OpenAI dropped ChatGPT, nobody knew chat would be such a hit. Several labs had versions of these things, Google engineers were getting finessed by a rudimentary model just like in the movie Ex Machina that came out 10 years prior which was satirizing Google

But I don't agree that they aren't product-ized. There are many applications doing calls to LLMs behind the scenes and are hits, leveraging structured data very heavily and not having conversations with users at all. I would say that there is a predictable scope creep from executives to surface a conversational aspect though. We need to bring representation to that so we can point to some other best practice to push back

17m agoHN ↗

IMO no one knows what LLMs are capable of yet and the goal post keeping moving every few months that building any specific UIs right now risks rendering them obsolete or too slow.

e.g. We went from somewhat smarter code autocomplete, to asking chatgpt copy paste, to cli agent running inside your project, managing session, to GUI to manage that, to projects where you talk to a "Chief of Staff" agent that manages other sessions, to who knows what's next.

I think the right interfaces for LLMs right now need to be very simple and easy to change/evolve. And chat still seems to be the best default solution.

14m agoHN ↗

It’s not surprising.

Technology folks don’t really understand people and what they need.

This always happens. This is why woz needed Steve.

Steve Jobs is sorely missed tbh. For all the shit he got - he was a true visionary. He lived at the intersection of technology and the humanities… he kept preaching this. And now we are seeing why.

28m agoHN ↗

I love this

Claude's mobile app wouldn't show any Cowork initiated conversations, even if those conversations didn't leverage cowork specific features

24m agoHN ↗

Most normal users have no idea whether to choose chat or cowork. This makes sense for 95% of customers.

22m agoHN ↗

Really excited to try the gsuite replacements, which seems like the bigger announcement than the title suggest

19m agoHN ↗

It's funny how they keep trying to add more interfaces and everything keeps coming back to chat. And of course it would, natural language is the simplest interface there is. Anything they add beyond chat, that isn't placed within chat, that isn't some artifact of chat, is friction.

18m agoHN ↗

I really appreciated that the users that I administrate had an option where Claude was not going to get carried away and orchestrate workslop, I will have to completely re-write my usage guides after testing how prompts are routed now.

Feels like the same "let us do _all_ of the thinking for you" messaging that Microsoft has with Copilot.

16m agoHN ↗

Hi, this is my team! Happy to answer any questions.

There's a lot in this launch, but the core idea is to simplify the product while giving users access to more capabilities. You no longer need to know ahead of time how much work a conversation might involve. If you're at your computer, Claude can use your local files and apps. If you close your laptop, Claude can keep working on its own computer.

This launch also lets you use Claude Design, Claude Docs, and Claude Slides directly from conversations. That's possible because we made Artifacts much more powerful: whenever Claude makes you an app, website, design system, or anything else, it can deploy an artifact with multiplayer features and databases.

As many of you probably know from your own work, giving users more power while making the experience simpler is really, really hard. It took many iterations to get to this version. We're far from done, but I expect people will be able to do much more while having to think about it less.

12m agoHN ↗

Will there be a Claude Sheets too? I find myself using Claude to work with spreadsheets more than almost any other document type. It does okay now, but feel like it could be even better, especially with visualizations and formulas.

11m agoHN ↗

I liked the ability to explicitly only chat, without the possibility of Cowork activating since I have never wanted that. It's unclear how I can still guarantee that now.

15m agoHN ↗

They could have figured that out before launching Cowork to everyone.

12m agoHN ↗

Basically DOA in healthcare, even with a BAA I don't want Claude Cowork enabled in our workstations, ever.