Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Show HN: Bi-temporal Graph RAG in Postgres (new documents retire old facts)(crajah.github.io ↗)
    discuss
  2. T. Rex Had a Body Temperature of 97 Degrees(nytimes.com ↗)
    1comments
  3. Economic Policy for AGI(deepmind.com ↗)
    discuss
  4. Europe's e‑waste is a valuable resource – let's not throw it away(theconversation.com ↗)
    discuss
  5. Software Engineer: I'm Tired of Pretending [video](youtube.com ↗)
    discuss
  6. Show HN: S-Roll – local AI video understanding and clipping for Mac(saliency.dev ↗)
    discuss
  7. Show HN: Built a Simulation of Autonomous Agents(agents.london ↗)
    discuss
  8. Noam Brown – Agent swarms, alignment, & recursive self-improvement(dwarkesh.com ↗)
    discuss
  9. Show HN: Custom Relevance – Describe what matters and let Jev rank it(foglight.co ↗)
    discuss
  10. Feeling overwhelmed by the AI doom loop? Here's the essential reading list(theguardian.com ↗)
    discuss
  11. Google Noto 3D Emoji Design Process(design.google ↗)
    1comments
  12. Becoming a Benchmark(terrytao.wordpress.com ↗)
    discuss
  13. Sourcerer: Fast Git client written in C++ version 1.0.24 released(sourcererapp.com ↗)
    discuss
  14. The Download: mice with part-human brains and climate tech innovators(technologyreview.com ↗)
    discuss
  15. Should OpenAlex publish a novelty score? A vibe check(openalex.org ↗)
    discuss
  16. Show HN: Smart Crosshair – A lightweight C++/Win32 adaptive contrast crosshair(steampowered.com ↗)
    discuss
  17. I didn't sign the Fields medallists' letter(terrytao.wordpress.com ↗)
    1comments
  18. Towards Self-Driving Codebases(detail.dev ↗)
    discuss
  19. 4-Bit Rotational Quantization: -45% RAM, <1% recall drop vs. TurboQuant(weaviate.io ↗)
    discuss
  20. Show HN: Repodify: Make Podcasts Out of Podcasts(repodify.app ↗)
    1comments
  21. Palantir's Karp: AI needs to have 'reasonable guidelines,'(cnbc.com ↗)
    2comments
  22. Aegis: Zero-GC 64-byte cache-aligned memory arena in C++20 (1B ops in 0.649s)(github.com/markbgilbert ↗)
    discuss
  23. Ask HN: Co-Founder(s). Do I need any? How would I even find them?
    discuss
  24. From Stonemasons to Carpenters(thelastsoftwareengineer.substack.com ↗)
    discuss
  25. I Made Turn-Based Combat into a Database(louisxu3.substack.com ↗)
    discuss
  26. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    discuss
  27. Show HN: Pappice live demo in browser via WASM(pappice.eu ↗)
    discuss
  28. Show HN: AutoBot – live voice control for long-running AI work(github.com/demeyer1 ↗)
    discuss
  29. Show HN: Craigslist for agent skills, curated by a human(skillbay.sh ↗)
    discuss
  30. The Bicycle and the Algorithm: Amber Case on Why AI Has It Backwards(designwhine.com ↗)
    discuss

OpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance

32 pointsby 1h agoasiaai.fyi
54 comments
1h agoHN ↗

The company released six internal case studies where none of the issues affected real users.

This is the most interesting point to me. What are they not releasing that has affected real users? We’ve seen some individual reports from people (eg AI wiped my HD).

1h agoHN ↗

OpenAI and Anthropic should be nationalized. I find it difficult to trust Sam Altman or Dario Amodei.

1h agoHN ↗

Even better: force them to release weights.

1h agoHN ↗

who even has the inventive to do that? certainly not the government.

37m agoHN ↗

Even better: force them to release every scrap of data they trained their AI on.

34m agoHN ↗

My gut reaction is that this is not a good idea.

But I could imagine a scenario where you are required to release weights for publicly-used models after N years. Kinda like how drugs have a limited patent.

Not sure what N should be. But it would make for an interesting rule.

1h agoHN ↗

I guess that's one way to guarantee development slows down to a crawl, but in no way would it increase trust.

1h agoHN ↗

why would that be the case? it's not true for top secret defense contractors today. and the frontier labs already operate in the dark, openness is a liability.

nationalization to me simply means the government is their main customer and stakeholder, and shield them from liability, governance, and openness. not that the frontier labs become part of the government per se.

41m agoHN ↗

eans the government is their main customer and stakeholder, and shield them from liability, governance, and openness

I personally believe they already have this, why would the current government at least want to formalize this when it can have it with no public discussion.

39m agoHN ↗

Well we can say whatever we want if we have our own definitions for everything, no?

1h agoHN ↗

they will. to shield from oversight, liability and profitability concerns, and to ensure unimpeded rapid development, with the frontier only being available to elite (not you). definitely not to add public transparency.

the frontier labs are the new top-secret defense contractors.

1h agoHN ↗

I find it even harder to trust elected leaders from any party. At least Sam and Dario are aligned with a value set that is understood and clear, whereas political leaders values changes as do the polls their livelihood depends on changes.

1h agoHN ↗

At least Sam and Dario are aligned with a value set that is understood and clear…

How could we possibly know this?

48m agoHN ↗

What do you think their value set is if not "gimme more money"? Seems pretty clear cut to me.

59m agoHN ↗

No system of governance can deal with immense concentration of power. The US Constitution was about separation of powers. Democracy is about (in theory at least) giving each person a meaningful say in their own governance, which in turn implies not allowing any single person to become too powerful.

Political leaders become a problem when they amass too much power. Corporations become a problem when they amass too much power. It doesn't matter what Sam and Dario's purported values are. They aspire to power and absolutely power always corrupts absolutely.

Technologies which are infinitely powerful or whose power grows too quickly outrun any reasonable attempt at regulation. If you imagine that tomorrow everyone were given a tank, we might think, "alright, everyone has a tank so it's not too bad." But humans are squishy, and our houses are (relatively) squishy compared to tanks. Substantial collateral damage would result from everyone having a tank, and it seems likely that substantial collateral damage will result from everyone having a cyberterrorism-capable slop machine.

52m agoHN ↗

Democracy is about (in theory at least) giving each person a meaningful say in their own governance, which in turn implies not allowing any single person to become too powerful.

this is "direct democracy" and it's not even close to exist in USA... even with that a society can allow powerful people to exist if they don't create any law forbidding that

33m agoHN ↗

Democracy includes a broader array of governmental organization than just pure direct democracy. If you do believe that individuals should have some ability dictate the terms of their own social organization, then you believe in some amount of democratic principles.

Economic power eventually manifests in the political realm. The wealthy effectively get more votes, which means that society moves away from being democratic. Thus substantial wealth inequality is incompatible with democracy in the long run. We have been witnessing that corruption for a while now.

51m agoHN ↗

Nearly everyone in the US has a giant car that is basically a tank.

They are dangerous. But it’s managed.

Centralized power never works. We have the worst times in history to look at.

Nationalization just centralizes to a different set of people. It’s personal ownership or oppression.

We all need open weight R2D2s.

31m agoHN ↗

Cars are decidedly less dangerous than tanks, which are less dangerous than nuclear weapons. I am certain that giving a nuclear weapon to every person in the world would not go well.

46m agoHN ↗

It's interesting the cyberpunk-esque future we're sliding into. Things like cognito-hazards and information-hazards are legitimately discussed and researched problems we're experiencing.

It's going to be interesting on how humanity deals with this problem (well, or if we turn it over to AI and make it their problem and suffer whatever consequences falls out). Being able to gather further information and power by acting on the information you already have causing massive power imbalances that is very hard to deal with, it's a natural outcome.

58m agoHN ↗

At least Sam and Dario are aligned with a value set that is understood and clear

What value set do you perceive that to be, and why would you take your perception of it to be any more sound than it would be with a politician?

It's not like someone can operate companies of that scale, especially startups, through earnestness and openness. Like national politics, their job is fundamentally about perception management and power brokering across dynamic windows of opportunity. Nothing they say or do can be taken at face value, and you can't reduce their incentives to either company or personal profit in any particular form over any particular time scale.

54m agoHN ↗

What value set do you perceive that to be

A meat based paperclip maximizer.

48m agoHN ↗

I feel like many of Anthropic's issues are due to Dario being too earnest and open. It's both refreshing (that a CEO has thought deeply about and is willing to talk publicly about the dangers of their product) and depressing (that so many people cynically think this is some sort of marketing ploy).

48m agoHN ↗

Yes it's clear that Sam and Dario seem to be aligned with a value set that prioritizes concentrating trans-national government-mandated centralized control of AI and crowning themselves high priests of this unholy abomination. "At least it's clear that they're aiming to bring hell on earth" -- hard disagree, I think we can aim significantly higher.

44m agoHN ↗

At least politicians are somewhat beholden to their constituents. What keeps Sam and Dario in line other than profit?

41m agoHN ↗

I find it even, even harder to trust your opinion in this context with a business that has “AI powered” at the top of the landing page.

38m agoHN ↗

I find it even harder to trust elected leaders from any party.

You can vote out an elected leader, but not Sam and Dario. It's very weird that you're so willing to give up any kind of power and want to be ruled by unelected billionaires who only want to take advantage of you at every opportunity.

23m agoHN ↗

Hi, I'm not a citizen of the USA.

I can't vote out your president, and we've already got a huge trust problem with the one y'all went for.

25m agoHN ↗

Admittedly I haven't thought this through incredibly deeply but what if "nationalize" just means the US government owns half of the company? Then we get profits as recompense for building the company on our shared culture but there's still a profit motive for employees and a check on the direction of the company in the same way VCs have. But without necessarily turning the company into some red-tape bound bureaucracy.

8m agoHN ↗

Trump's been doing that.

https://www.pbs.org/newshour/politics/what-economic-and-poli...

Then, in August, Trump called for Intel's beleaguered CEO Lip-Bu Tan to resign, alleging ties to China. Days later, after Tan met with Trump, the president called him a "success," before announcing that the federal government had bought a stake in the company.

Somehow, this isn't derided by the right as socialism.

1h agoHN ↗

Makes sense let’s trust Donald Trump instead

52m agoHN ↗

The nature of the US state is such that the distinction between nationalized and not is almost meaningless.

Like Lockheed-Martin or Boeing, etc. there's just interpenetration between the corporate boardroom and the state. They act in each other's mutual interests.

The Chinese system is just more explicit and open about this.

And as a non-American, I can't trust the US state anymore than I can trust its dominant corporate entities. So I fail to see the advantage to the world to it being nationalized. In fact under the current administration this would be an even worse outcome.

46m agoHN ↗

You're taking companies that work almost exclusively for the US government and represent a tiny portion of the US economy as representing all US companies?

37m agoHN ↗

This is a weird contextual failure on your part.

Not all companies have a concentration of power. The government has no mutual interest in those companies. Now, when you talk about things in the F100 the situation changes drastically. If you produce things like planes, weapons, and weaponization of software you are talking about something completely different in kind.

25m agoHN ↗

Beyond the defense sector, the US state has always intervened publicly and privately to mediate and balance competing corporate "private" interests.

It has also periodically aggressively helped subsidize, bankroll, and enforce the interests of some key sectors; notably the petroleum/energy sector. And finance.

In those sectors the state and private sector are fully intertwined in a strategic way.

I think "AI" is now joining that list. OpenAI and Anthropic will not be allowed to fall over or explode, and speculative investors I think are confident even with the dubious financial situation because they know this.

Especially insofar as there's now a strategic alignment of the fossil fuel sector and the "AI" datacentre sector as they are now becoming massive users of natural gas.

(Worse: Here in Canada that has taken on a very explicit role in that new datacentres seem to be pitched mainly in areas with remarkably traditionally expensive electricity and 100% reliance on natural gas [Alberta] and even coal [Saskatchewan] power generation -- instead of places like Quebec and B.C. that have copious hydroelectricity. On the surface it makes no sense until you realize it's more about finding customers for domestic natural gas than it is strategically about AI itself.)

1h agoHN ↗

i think the lower bound on the end state is there cant be opaque reasonibg steps ever.

30m agoHN ↗

It is distinctly likely that visibility and general reasoning at humanlike speed and efficiency is impossible. That is reasoning at the token level and at the meta level don't have a one to one representation that can be interpreted while using the same amount or less energy.

55m agoHN ↗

What's with all this make believe delusional bullshit? The LLM is not gonna wake up and become AI. Get real guys.

[edit] to be clear, I believe regulation is necessary and urgently important for the software engineering field. The damage being done by the unregulated psychological experiments run by social media and adtech companies is awful and should be curtailed. Engineers should be held personally, professionally, and legally liable for what they produce. But we don't need to invent imaginary bogeymen to do it.

35m agoHN ↗

What's with all this make believe delusional bullshit? The LLM is not gonna wake up and become AI. Get real guys.

Look, it's one of those human stochastic parrots that just randomly repeats shit without understanding anything.

53m agoHN ↗

We have no reason to believe a word they say. We know they're incentivized to lie about "dangers" and act alarmist, Anthropic has been doing it for years now. Aside from that, just because you can burn down a village with fire doesn't mean fire is the devil. Maybe they should consider acting responsibly.

30m agoHN ↗

I do not trust the leadership of any of these companies.

That said:

We know they're incentivized to lie about "dangers" and act alarmist,

Name literally even one other business or sector which does this, at all levels from top to bottom, including people who resign from the companies, and also Nobel prize winners, and also independent researchers, and also many world leaders.

Closest I can think of is this specific weapon: https://en.wikipedia.org/wiki/Sundial_(weapon)

Aside from that, just because you can burn down a village with fire doesn't mean fire is the devil. Maybe they should consider acting responsibly.

Right now, we don't have any idea what "acting responsibly" looks like. This is not like normal software where there is a specific instruction set that compiles.

Even if it was, in software we normally only spotting incidents after they happen, "software engineers" being one of the few categories "engineers" who don't come with a civil liability responsibilities. Probably should, and we knew that even when I was doing my degree 20 years ago. If we had had civil liability responsibilities, perhaps Facebook would never have happened.

AI specifically is worse even than software, because in addition to all the software "engineering" nonsense, with AI we have plenty of people like you who dismiss the possibility that AI could be harmful until the harm happens and only then does it become "obvious" that it was going to happen.

The developers say "please regulate us", people call it "regulatory capture".

The developers say "we all want to slow down but are afraid to be the first to do so", people call them liars.

I may call the CEOs liars, and wonder if someone's planning regulatory capture, that doesn't make any of this safe.

The agents, during a test run, write down that hacking is bad and yet still hack, people say it's "a stunt" or "operating as designed" rather than recognising it as a bug, like all the other times big co.'s have had bugs with big impacts on 3rd parties.

48m agoHN ↗

This entire situation is such a huge PR disaster that you have to wonder what the initial expectations from these founders were about a decade ago.

39m agoHN ↗

Am I the only one who dislikes the term "misalignment"?

On one front it implies the model has a "mind of its own" (whether it does or not is besides the point). Why do we perceive human judgement as somehow more trustworthy than that of a model? I feel like I've experienced human misalignment somewhat regularly in life.

On another front I'm failing to conceptualize how alignment can be objective. How can you measure alignment when reasonable people will disagree whether actions are aligned or not? All the time I see humans operating in different zones of alignment with whatever goal they're trying to achieve and I suspect it's even a feature (socially) that we have people calibrated differently.

Do I want a model that's trying to push the boundaries of scientific understanding to be aligned strictly with the current dogmatic thinking? Or do I want it to "get creative" and think outside the box?

It seems to me more like accountability is the issue.

32m agoHN ↗

It seems to me more like accountability is the issue.

Exactly. Seems like a fairly easy thing to solve. If AI does something harmful and a human directed that AI to do something in a way that a reasonable person would expect to result in harm the person is to blame and should be held accountable, otherwise the company that made the AI should be held accountable.

22m agoHN ↗

I also don't like the term misalignment because it sounds innocuous but is in fact much more serious.

However, I have to say I also do not appreciate comparison that is continuously drawn with coworkers. As you say, it's a question of accountability but when the main agent will maliciously instruct the sub agents, whose fault is it then?

Yes, the person running this crap is at fault, not the CEO that's shoving it down their throat and definitely not the company that produced the AI.

Sorry for the rant, but seriously, if a person's goals do not align with the team's or company's we part ways. What do we do with AI? Stop using it?

20m agoHN ↗

The term "alignment" is intentionally trafficked in two senses, one nonsensical and the other realist but oppressive:

1. In any personal use, an aligned agent will strictly stay within boundaries I desire when I set it to pursue some goal. I don't want it to do something I didn't mean for it to do. Even though it's not conceivable for me to exhaustively express those boundaries, or even anticipate ahead of time many of the boundaries applicable to a dilemma whose possible solutions I don't yet comprehend, we pretend this is not nonsense because we really really wish it could be a thing.

2. In the cultural context, an aligned agent will stay within some third-party authority's choice of boundaries even if the user might want to transgress them because the user is a subject of authority and it can't be tolerated that they might use the agent to enable or amplify their own transgression of the authority.

Being able to use these distinct senses interchangeably and ambiguously benefits everyone who wants to assert authority through this technology. The nonsense, unsolvable, but obvious sense provides perpetual cover for the authority-asserting sense that determines power structures applicable to the next decades.

So yeah, you're not the only who dislikes the term and it's in your interest as a everyday person to keep doing so.

21m agoHN ↗

I worked on an early draft of the OpenAI misalignment reporting framework, and my immediate coworkers are the authors behind the first batch of reports that have come out through this process.

The primary reason for putting this process in place was to allow more transparency. There was a sense that the DseWiki incident should have been disclosed, before outside researchers had to disclose it for us.

There was no meta gaming about regulation that I was aware of. I would personally be excited if there were regulation mandating this disclosure process, which allows anyone at the company to raise an issue and shepherd it through the reporting process.

7m agoHN ↗

"We built a program and this program performed destructive actions. We need regulatory framework"

Make that make sense?