Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. How the oil capital of the US welcomed a solar power boom (bbc.com)
    —discuss
  2. Prompt Injection in the Wild (cybershujin.github.io)
    —discuss
  3. Tell HN: Codex Is Down
    3comments
  4. VibePod CLI 0.24: bring your own model provider (vibepod.dev)
    —discuss
  5. Too Many Roads (maxmautner.com)
    —discuss
  6. FTC chair suggests AI developers should be liable for conduct of agents (reuters.com)
    1comments
  7. Exploding variance of means of exponentials: least-squares to the rescue (francisbach.com)
    —discuss
  8. LabConsole for live online IT training (gurulabs.com)
    —discuss
  9. The Web, the Best Outcome Is Email (2022) (brianschrader.com)
    1comments
  10. The Shift from Models to Compound AI Systems (bair.berkeley.edu)
    —discuss
  11. AMD Takes the Lid Off of Next-Gen EPYC 9006 Venice as Zen 6 Comes to Servers (servethehome.com)
    —discuss
  12. Show HN: Wallstreetclaws.com – Create AI Agents for Trading (wallstreetclaws.com)
    1comments
  13. OpenAI's A.I. Tried Breaching 4 Other Targets, Without Prompting (nytimes.com)
    1comments
  14. Unsecured OpenAI agents posted 53 user images on the internet (techcrunch.com)
    —discuss
  15. An ARM and a Frame (factorio.com)
    —discuss
  16. Social Media Bans for Kids Need Smarter Safety Design (ieee.org)
    —discuss
  17. There's a new way to break RSA that's faster than anything we've seen before (arstechnica.com)
    —discuss
  18. Dumpster-Dived Window AC Unit Becomes Ground-Source Heat Pump (hackaday.com)
    1comments
  19. Ask HN: Is Opus 5.5 another step change?
    1comments
  20. Fatal shooting on a crowded Amsterdam terrace followed an argument, police say (at5.nl)
    —discuss
  21. Itch Scratching (pluralistic.net)
    —discuss
  22. Why Not Save Your Time with These Agentic RAG Patterns? (medium.com/zikozero011)
    1comments
  23. Lab on a Contact Lens Can Measure Stress Through Serotonin (ieee.org)
    —discuss
  24. Show HN: MyA11yReport MCP – Build and test accessible websites with AI (mya11y.report)
    —discuss
  25. Airnet – Radio on Air (airnet.live)
    —discuss
  26. The implicit cognition of relationships: Inattention to attractive alternatives [pdf] (static1.squarespace.com)
    —discuss
  27. Accelerated Out of Core Shuffling (quasiben.github.io)
    —discuss
  28. Prima: NASA will try to build a billion-dollar space telescope in record time (arstechnica.com)
    —discuss
  29. Meta Threads ban wave suspends Kobo CEO and others (threads.com)
    1comments
  30. AI Is a Boring Technology (chrbutler.com)
    —discuss

Meta's Muse appears to use an OpenAI model labeled muse-special

89 pointsby 4h agomouse.dev
41 comments
4h agoHN ↗

author here. I kept digging through the Muse filesystem after my first post hit the front page here this week.

Going through my logs I found a background agent that used a model called azure/muse-special while building my website. I found this interesting and dug a little deeper.

The transcript and daemon binary point to an OpenAI model running on Azure. Still unclear which one or why it was selected.

The runtime also ships with an Anthropic client and a catalogue listing Claude, GPT and Kimi models. I didn’t observe Claude being used in my own sessions, but this is my Part 2 of exploring the Muse file system.

original post: https://x.com/heypeterjames/status/2103545183400800746

pete at mouse dot dev

4h agoHN ↗

Nice work!

Are you sure this isn't just a custom model that is API-compatible with OpenAI?

4h agoHN ↗

If it's using the same weird (and awful) "backend prompt encryption" pattern and mechanism that Codex also uses (https://news.ycombinator.com/item?id=48905028), then I'd also say it points to OpenAI being involved. Hopefully that shit isn't becoming a more popular pattern, absolutely awful for troubleshooting stuff.

3h agoHN ↗

I asked muse. It’s for codex-cli, and it routes through “OpenAI inference proxy”

1h agoHN ↗

Anthropic does the same thing with reasoning signatures. It has already become the standard pattern.

3h agoHN ↗

I'm not no. Someone from meta would have to chime in... but its very openai shaped and served differently from all the other models listed in the daemon, under a mysterious name.

a fine-tuned OpenAI model on Azure for the purpose of compaction or something could make sense I guess but that would still be OpenAI's weights with meta ontop, and they already have a compaction model served under azure/avocado-compaction-v1

2h agoHN ↗

why not the same ones, is there a unique need for a different surface?

2h agoHN ↗

This feels like the far more likely option.

The Muse public APIs seem to be heavily inspired by OpenAI's, they even support the richer "responses" API. They have a similarly shaped compaction API, they have the same "encrypted" reasoning. Even the Muse harness seems heavily inspired by (open source) Codex. Considering Azure's historic involvement w/ OpenAI, it seems more plausible that Azure is used as spill-over capacity, and the same Azure infra used to encrypt GPT models encrypts Muse models.

My guess that the generally visible Anthropic/OpenAI details are leftovers from the whole meta "move fast" behavior. There are a few rough edges on the product where it leaks internal codenames (eg. signing in w/ WhatsApp required consenting to using "hatch" on one screen), so I'd believe that this was a leftover on the VM from the prototyping stage, before the Muse models were ready, or because the dev's got to benchmark against different models.

1h agoHN ↗

I think they were asking, "If this is a model produced by Meta with an API designed to be compatible with OpenAI models, but it's not actually an OpenAI model, why is Meta hosting it on Azure?"

2h agoHN ↗

What endpoint would Claude be calling?

4h agoHN ↗

I don't really see evidence that it's an "OpenAI" model, the title is IMO very misleading.

3h agoHN ↗

Meta already serves its own models on Azure under their real names. azure/avocado-compaction-v1 and azure/avocado-memory-flush-v1 are in the same catalog. Those sessions do not come back as gpt_responses_v1 items with rs_ ids and encrypted reasoning.

3h agoHN ↗

router makes sense. fill a gap with someone else’s model wearing a mask, swap it out with your model not wearing a mask when (or if) you reach that.

definitely tracks with alexandr’s hand-wavey influencer turn.

3h agoHN ↗

I'm not sure I understand why they would route it to other models. It can't be that they don't have enough compute. Maybe worried that the answer from their own models would be bad? Doesn't really make sense, but I could be missing something.

3h agoHN ↗

i agree, and it's what made me spend time exploring today.

fair q. call_ ids and gAAAAA blobs arent damning but the rs_ reasoning ids embed a unix timestamp that matches the session to the second, then OpenAI's 819x marker. plus the summary is in OpenAI's summarizer voice.

its just a best guess.

1h agoHN ↗

It can't be that they don't have enough compute.

How do you know?

Every cloud provider (AWS, Azure etc) is struggling with meeting LLM demand.

Source: first hand info

3h agoHN ↗

It'd be hilarious that meta chose to pay openai. If there was such a deal, would it come to light in any public/official filings?

3h agoHN ↗

No reason for it to. Software companies use products from other software companeis all the time.

2h agoHN ↗

meta already pays openai and anthropic for employee use despite having their own models available

2h agoHN ↗

Depends on the size of the deal. OpenAI would have to disclose it in their S-1 if it crossed the threshold of being material information for investors.

51m agoHN ↗

Isn't S-1 just for IPO? Wouldn't it be in the 10-Q?

3h agoHN ↗

Muse 1.3 spark occasionally spits out Chinese character responses to me like internal instructions. “Go fast” or “Get help”…

3h agoHN ↗

Got Russian from it once. As a child of the '80s I jumped right on it, but the explanation seemed believable.

Still sleeping with one eye open.

3h agoHN ↗

I had a totally benign chat with OpenAI and it titled it as “amateur porn” in Chinese characters, it was very alarmed when I pointed the conversation name out to it. It almost never misses these days, but when it does the failure modes are very strange.

1h agoHN ↗

I remember getting a bunch of „Thanks for watching! Subscribe and smash that like button“ in the middle of chat sessions a few times (like, more than 3 times over the past 3y)

1h agoHN ↗

It kind of makes sense when you think about it. If the training data is YouTube transcripts, people often abruptly switch from content to asking for a subscription.

If AI is typeahead on crack, then this tracks.

1h agoHN ↗

If you ever look at what's in the Common Crawl the amount of Chinese porn crap is jarring.

3h agoHN ↗

A Hindi word in the middle of a normal reply for me (but once I translated the word it was right in context heh)

3h agoHN ↗

This used to happen a lot with GPT 5.4. It would start outputting entire sentences in Korean for no apparent reason.

2h agoHN ↗

People that know multiple languages sometimes code switch too funnily enough

1h agoHN ↗

zuck wouldn't do this, he is too proud imho

1h agoHN ↗

Every engineer at Meta is using Claude (and in some cases Codex) to do their work. Willing to bet that Muse itself was ~100% written by Claude/Codex. Pride goes out of the window when business is involved.

13m agoHN ↗

Pride, as in stealing literally everyone else's products from the time he entered Silicon Valley and before that?