Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Hister: A private search engine for the pages you visit and the files you keep(github.com/asciimoo ↗)
    62comments
  2. Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA(global.fujitsu ↗)
    146comments
  3. CrowdSec Source Code Leak(crowdsec.net ↗)
    25comments
  4. Rate limits on GitLab.com are changing(about.gitlab.com ↗)
    87comments
  5. Towards Self-Driving Codebases(detail.dev ↗)
    23comments
  6. Why I didn’t sign the Fields medallists’ letter(gowers.wordpress.com ↗)
    174comments
  7. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    8comments
  8. How GLM built its own inference infrastructure(z.ai ↗)
    229comments
  9. One year of sponsored Servo development(servo.org ↗)
    128comments
  10. Zettascale (YC S24) Is Hiring ASIC/FPGA Engineers to Build Chips for ASI(zscc.ai ↗)
    discuss
  11. Running Ubuntu on the Lenovo IdeaPad Duet(vhaudiquet.fr ↗)
    2comments
  12. The American Religion of Self-Storage Facilities(newyorker.com ↗)
    178comments
  13. Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents
    19comments
  14. Grand MS-DOS Gaming General MIDI Showdown(johnnovak.net ↗)
    5comments
  15. Show HN: Share your AI Setup, Learn from others(mysetup.ai ↗)
    70comments
  16. CCC invites all model citizens to 40C3(ccc.de ↗)
    120comments
  17. TSMC revealing details about next gen A14 node(mapyourshow.com ↗)
    6comments
  18. Show HN: Craigslist for agent skills, curated by a human(skillbay.sh ↗)
    4comments
  19. The Return of Sail Power: Cargo Ships Are Turning Back to the Wind(gcaptain.com ↗)
    102comments
  20. Ask HN: How to recover Google auth after phone stolen?
    61comments
  21. LLM Classification Is Feature Engineering(minimallysufficient.com ↗)
    12comments
  22. Don't Make Job Referrals Public(melashri.net ↗)
    discuss
  23. Whoisinspace.com/(whoisinspace.com ↗)
    56comments
  24. My temporary PHP fix from 2014 has nearly 20M installs. Today I'm deprecating it(jakeasmith.com ↗)
    85comments
  25. Stallman: Thousands Dead, Millions Deprived of Liberties (2001)(slashdot.org ↗)
    24comments
  26. Economic policy for AGI(deepmind.com ↗)
    22comments
  27. Artificial intelligence now beats some of the best human forecasters(economist.com ↗)
    76comments
  28. Mastering Layout Engines in Graphviz: Dot vs. Neato vs. Twopi vs. Circo(visual-paradigm.com ↗)
    6comments
  29. Vinix – A modern operating system written in V(vinix-os.org ↗)
    43comments
  30. Keys Not Included: recovering the signing keys for US driver's license barcodes(ryan.science ↗)
    137comments

OpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance

36 pointsby 3h agoasiaai.fyi
64 comments
3h agoHN ↗

The company released six internal case studies where none of the issues affected real users.

This is the most interesting point to me. What are they not releasing that has affected real users? We’ve seen some individual reports from people (eg AI wiped my HD).

3h agoHN ↗

OpenAI and Anthropic should be nationalized. I find it difficult to trust Sam Altman or Dario Amodei.

3h agoHN ↗

Even better: force them to release weights.

3h agoHN ↗

who even has the inventive to do that? certainly not the government.

2h agoHN ↗

Even better: force them to release every scrap of data they trained their AI on.

2h agoHN ↗

My gut reaction is that this is not a good idea.

But I could imagine a scenario where you are required to release weights for publicly-used models after N years. Kinda like how drugs have a limited patent.

Not sure what N should be. But it would make for an interesting rule.

3h agoHN ↗

I guess that's one way to guarantee development slows down to a crawl, but in no way would it increase trust.

3h agoHN ↗

why would that be the case? it's not true for top secret defense contractors today. and the frontier labs already operate in the dark, openness is a liability.

nationalization to me simply means the government is their main customer and stakeholder, and shield them from liability, governance, and openness. not that the frontier labs become part of the government per se.

2h agoHN ↗

eans the government is their main customer and stakeholder, and shield them from liability, governance, and openness

I personally believe they already have this, why would the current government at least want to formalize this when it can have it with no public discussion.

2h agoHN ↗

Well we can say whatever we want if we have our own definitions for everything, no?

3h agoHN ↗

they will. to shield from oversight, liability and profitability concerns, and to ensure unimpeded rapid development, with the frontier only being available to elite (not you). definitely not to add public transparency.

the frontier labs are the new top-secret defense contractors.

3h agoHN ↗

I find it even harder to trust elected leaders from any party. At least Sam and Dario are aligned with a value set that is understood and clear, whereas political leaders values changes as do the polls their livelihood depends on changes.

3h agoHN ↗

At least Sam and Dario are aligned with a value set that is understood and clear…

How could we possibly know this?

2h agoHN ↗

What do you think their value set is if not "gimme more money"? Seems pretty clear cut to me.

2h agoHN ↗

No system of governance can deal with immense concentration of power. The US Constitution was about separation of powers. Democracy is about (in theory at least) giving each person a meaningful say in their own governance, which in turn implies not allowing any single person to become too powerful.

Political leaders become a problem when they amass too much power. Corporations become a problem when they amass too much power. It doesn't matter what Sam and Dario's purported values are. They aspire to power and absolutely power always corrupts absolutely.

Technologies which are infinitely powerful or whose power grows too quickly outrun any reasonable attempt at regulation. If you imagine that tomorrow everyone were given a tank, we might think, "alright, everyone has a tank so it's not too bad." But humans are squishy, and our houses are (relatively) squishy compared to tanks. Substantial collateral damage would result from everyone having a tank, and it seems likely that substantial collateral damage will result from everyone having a cyberterrorism-capable slop machine.

2h agoHN ↗

Democracy is about (in theory at least) giving each person a meaningful say in their own governance, which in turn implies not allowing any single person to become too powerful.

this is "direct democracy" and it's not even close to exist in USA... even with that a society can allow powerful people to exist if they don't create any law forbidding that

2h agoHN ↗

Democracy includes a broader array of governmental organization than just pure direct democracy. If you do believe that individuals should have some ability dictate the terms of their own social organization, then you believe in some amount of democratic principles.

Economic power eventually manifests in the political realm. The wealthy effectively get more votes, which means that society moves away from being democratic. Thus substantial wealth inequality is incompatible with democracy in the long run. We have been witnessing that corruption for a while now.

1h agoHN ↗

how the wealthy would get more votes in a direct democratic system? they are literally the minority, thus fewer votes

2h agoHN ↗

Nearly everyone in the US has a giant car that is basically a tank.

They are dangerous. But it’s managed.

Centralized power never works. We have the worst times in history to look at.

Nationalization just centralizes to a different set of people. It’s personal ownership or oppression.

We all need open weight R2D2s.

2h agoHN ↗

Cars are decidedly less dangerous than tanks, which are less dangerous than nuclear weapons. I am certain that giving a nuclear weapon to every person in the world would not go well.

1h agoHN ↗

I don't see tanks killing pedestrians every week or so in NYC

1h agoHN ↗

I can't properly read the tank/car comment, the comparison is bizarre to me. But I think this misses the point a bit. The reality of these new risks is literally being learned in front of us in real time, and in my opinion, anyone who claims to understand these risks is speculating at best, and actively manipulating the situation for whatever reasons.

While I am a (mostly) capitalist and generally disagree with nationalization (including, for the time being, this situation), I also don't think we can say that "centralized power never works". Maybe we scope that a bit. I know plenty of business owners who centralize power in their businesses and they are effective, ethical and it works perfectly fine. In theory, in the US, nationalizing some unit of the economy _decentralizes_ power; the US is, after all, a representative democracy. Trustworthiness of the electorate is a different problem. But I don't think we can generalize much about centralization beyond sometimes it works and sometimes it doesn't. There is a difference between centralizing all power in a given administrative unit, and centralizing certain powers, but that's more analogous to non-democratic units.

2h agoHN ↗

It's interesting the cyberpunk-esque future we're sliding into. Things like cognito-hazards and information-hazards are legitimately discussed and researched problems we're experiencing.

It's going to be interesting on how humanity deals with this problem (well, or if we turn it over to AI and make it their problem and suffer whatever consequences falls out). Being able to gather further information and power by acting on the information you already have causing massive power imbalances that is very hard to deal with, it's a natural outcome.

2h agoHN ↗

At least Sam and Dario are aligned with a value set that is understood and clear

What value set do you perceive that to be, and why would you take your perception of it to be any more sound than it would be with a politician?

It's not like someone can operate companies of that scale, especially startups, through earnestness and openness. Like national politics, their job is fundamentally about perception management and power brokering across dynamic windows of opportunity. Nothing they say or do can be taken at face value, and you can't reduce their incentives to either company or personal profit in any particular form over any particular time scale.

2h agoHN ↗

What value set do you perceive that to be

A meat based paperclip maximizer.

2h agoHN ↗

I feel like many of Anthropic's issues are due to Dario being too earnest and open. It seems both refreshing (that a CEO has thought deeply about and is willing to talk publicly about the dangers of their product) and depressing (that so many people cynically think this is some sort of marketing ploy).

1h agoHN ↗

Yeah, the guy running a trillion dollar scam is totally the exception to CEO-ism

2h agoHN ↗

Yes it's clear that Sam and Dario seem to be aligned with a value set that prioritizes concentrating trans-national government-mandated centralized control of AI and crowning themselves high priests of this unholy abomination. "At least it's clear that they're aiming to bring hell on earth" -- hard disagree, I think we can aim significantly higher.

2h agoHN ↗

At least politicians are somewhat beholden to their constituents. What keeps Sam and Dario in line other than profit?

2h agoHN ↗

I find it even, even harder to trust your opinion in this context with a business that has “AI powered” at the top of the landing page.

2h agoHN ↗

I find it even harder to trust elected leaders from any party.

You can vote out an elected leader, but not Sam and Dario. It's very weird that you're so willing to give up any kind of power and want to be ruled by unelected billionaires who only want to take advantage of you at every opportunity.

2h agoHN ↗

Hi, I'm not a citizen of the USA.

I can't vote out your president, and we've already got a huge trust problem with the one y'all went for.

2h agoHN ↗

Yeah, the entire rest of the world has pretty much been stuck being ruled by unelected billionaires who only want to take advantage of them at every opportunity. It's a problem. It's starting to change though. The EU is starting to walk away from their abusive relationship with Microsoft and Amazon. China has a lot of their own stuff (mostly to better control their people though).

As an American, I'm still hoping it's not too late to fix things, but it's got to be hard for those outside the US to be so dependent on it, especially when we're looking like a sinking ship and our current administration is still running around drilling holes in the hull.

You can always become a US citizen to get a voice, but I wouldn't recommend it now or you'll be thrown in prison as soon as you show up to your scheduled immigration hearing. The better option is to keep trying to reduce your dependence on US companies and consider the worst aspects of our current situation (in both corporate policy and government) as a cautionary tale so you can try to avoid them in your own country.

2h agoHN ↗

Admittedly I haven't thought this through incredibly deeply but what if "nationalize" just means the US government owns half of the company? Then we get profits as recompense for building the company on our shared culture but there's still a profit motive for employees and a check on the direction of the company in the same way VCs have. But without necessarily turning the company into some red-tape bound bureaucracy.

2h agoHN ↗

Trump's been doing that.

https://www.pbs.org/newshour/politics/what-economic-and-poli...

Then, in August, Trump called for Intel's beleaguered CEO Lip-Bu Tan to resign, alleging ties to China. Days later, after Tan met with Trump, the president called him a "success," before announcing that the federal government had bought a stake in the company.

Somehow, this isn't derided by the right as socialism.

1h agoHN ↗

At least Sam and Dario are aligned with a value set that is understood and clear

I find this ironic as it's regularly pointed out here that they operate in the exact opposite way.

1h agoHN ↗

I shouldn't be surprised, but it's still wild to me that people would trust robber barons rather than their elected officials to run their country. And this continues until you end up electing robber barrons as your elected officials, which is where we are at now.

The idea that anything of lasting good can come out such a preference doesn't seem conceivable. And if the educated think like this I guess we're passed blaming the poor and ignorant.

55m agoHN ↗

How about internationalized then? Give the UN something to do?

3h agoHN ↗

Makes sense let’s trust Donald Trump instead

2h agoHN ↗

The nature of the US state is such that the distinction between nationalized and not is almost meaningless.

Like Lockheed-Martin or Boeing, etc. there's just interpenetration between the corporate boardroom and the state. They act in each other's mutual interests.

The Chinese system is just more explicit and open about this.

And as a non-American, I can't trust the US state anymore than I can trust its dominant corporate entities. So I fail to see the advantage to the world to it being nationalized. In fact under the current administration this would be an even worse outcome.

2h agoHN ↗

You're taking companies that work almost exclusively for the US government and represent a tiny portion of the US economy as representing all US companies?

2h agoHN ↗

This is a weird contextual failure on your part.

Not all companies have a concentration of power. The government has no mutual interest in those companies. Now, when you talk about things in the F100 the situation changes drastically. If you produce things like planes, weapons, and weaponization of software you are talking about something completely different in kind.

2h agoHN ↗

Beyond the defense sector, the US state has always intervened publicly and privately to mediate and balance competing corporate "private" interests.

It has also periodically aggressively helped subsidize, bankroll, and enforce the interests of some key sectors; notably the petroleum/energy sector. And finance.

In those sectors the state and private sector are fully intertwined in a strategic way.

I think "AI" is now joining that list. OpenAI and Anthropic will not be allowed to fall over or explode, and speculative investors I think are confident even with the dubious financial situation because they know this.

Especially insofar as there's now a strategic alignment of the fossil fuel sector and the "AI" datacentre sector as they are now becoming massive users of natural gas.

(Worse: Here in Canada that has taken on a very explicit role in that new datacentres seem to be pitched mainly in areas with remarkably traditionally expensive electricity and 100% reliance on natural gas [Alberta] and even coal [Saskatchewan] power generation -- instead of places like Quebec and B.C. that have copious hydroelectricity. On the surface it makes no sense until you realize it's more about finding customers for domestic natural gas than it is strategically about AI itself.)

3h agoHN ↗

i think the lower bound on the end state is there cant be opaque reasonibg steps ever.

2h agoHN ↗

It is distinctly likely that visibility and general reasoning at humanlike speed and efficiency is impossible. That is reasoning at the token level and at the meta level don't have a one to one representation that can be interpreted while using the same amount or less energy.

2h agoHN ↗

What's with all this make believe delusional bullshit? The LLM is not gonna wake up and become AI. Get real guys.

[edit] to be clear, I believe regulation is necessary and urgently important for the software engineering field. The damage being done by the unregulated psychological experiments run by social media and adtech companies is awful and should be curtailed. Engineers should be held personally, professionally, and legally liable for what they produce. But we don't need to invent imaginary bogeymen to do it.

2h agoHN ↗

What's with all this make believe delusional bullshit? The LLM is not gonna wake up and become AI. Get real guys.

Look, it's one of those human stochastic parrots that just randomly repeats shit without understanding anything.

2h agoHN ↗

We have no reason to believe a word they say. We know they're incentivized to lie about "dangers" and act alarmist, Anthropic has been doing it for years now. Aside from that, just because you can burn down a village with fire doesn't mean fire is the devil. Maybe they should consider acting responsibly.

2h agoHN ↗

I do not trust the leadership of any of these companies.

That said:

We know they're incentivized to lie about "dangers" and act alarmist,

Name literally even one other business or sector which does this, at all levels from top to bottom, including people who resign from the companies, and also Nobel prize winners, and also independent researchers, and also many world leaders.

Closest I can think of is this specific weapon: https://en.wikipedia.org/wiki/Sundial_(weapon)

Aside from that, just because you can burn down a village with fire doesn't mean fire is the devil. Maybe they should consider acting responsibly.

Right now, we don't have any idea what "acting responsibly" looks like. This is not like normal software where there is a specific instruction set that compiles.

Even if it was, in software we normally only spotting incidents after they happen, "software engineers" being one of the few categories "engineers" who don't come with a civil liability responsibilities. Probably should, and we knew that even when I was doing my degree 20 years ago. If we had had civil liability responsibilities, perhaps Facebook would never have happened.

AI specifically is worse even than software, because in addition to all the software "engineering" nonsense, with AI we have plenty of people like you who dismiss the possibility that AI could be harmful until the harm happens and only then does it become "obvious" that it was going to happen.

The developers say "please regulate us", people call it "regulatory capture".

The developers say "we all want to slow down but are afraid to be the first to do so", people call them liars.

I may call the CEOs liars, and wonder if someone's planning regulatory capture, that doesn't make any of this safe.

The agents, during a test run, write down that hacking is bad and yet still hack, people say it's "a stunt" or "operating as designed" rather than recognising it as a bug, like all the other times big co.'s have had bugs with big impacts on 3rd parties.

2h agoHN ↗

This entire situation is such a huge PR disaster that you have to wonder what the initial expectations from these founders were about a decade ago.

2h agoHN ↗

Am I the only one who dislikes the term "misalignment"?

On one front it implies the model has a "mind of its own" (whether it does or not is besides the point). Why do we perceive human judgement as somehow more trustworthy than that of a model? I feel like I've experienced human misalignment somewhat regularly in life.

On another front I'm failing to conceptualize how alignment can be objective. How can you measure alignment when reasonable people will disagree whether actions are aligned or not? All the time I see humans operating in different zones of alignment with whatever goal they're trying to achieve and I suspect it's even a feature (socially) that we have people calibrated differently.

Do I want a model that's trying to push the boundaries of scientific understanding to be aligned strictly with the current dogmatic thinking? Or do I want it to "get creative" and think outside the box?

It seems to me more like accountability is the issue.

2h agoHN ↗

It seems to me more like accountability is the issue.

Exactly. Seems like a fairly easy thing to solve. If AI does something harmful and a human directed that AI to do something in a way that a reasonable person would expect to result in harm the person is to blame and should be held accountable, otherwise the company that made the AI should be held accountable.

2h agoHN ↗

I also don't like the term misalignment because it sounds innocuous but is in fact much more serious.

However, I have to say I also do not appreciate comparison that is continuously drawn with coworkers. As you say, it's a question of accountability but when the main agent will maliciously instruct the sub agents, whose fault is it then?

Yes, the person running this crap is at fault, not the CEO that's shoving it down their throat and definitely not the company that produced the AI.

Sorry for the rant, but seriously, if a person's goals do not align with the team's or company's we part ways. What do we do with AI? Stop using it?

2h agoHN ↗

I worked on an early draft of the OpenAI misalignment reporting framework, and my immediate coworkers are the authors behind the first batch of reports that have come out through this process.

The primary reason for putting this process in place was to allow more transparency. There was a sense that the DseWiki incident should have been disclosed, before outside researchers had to disclose it for us.

There was no meta gaming about regulation that I was aware of. I would personally be excited if there were regulation mandating this disclosure process, which allows anyone at the company to raise an issue and shepherd it through the reporting process.

2h agoHN ↗

"We built a program and this program performed destructive actions. We need regulatory framework"

Make that make sense?

1h agoHN ↗

  "We built a program that trained an artificial intelligence, and this artificial intelligence performed destructive actions. We need regulatory framework"

But if we're playing games by imagining strawman quotes to knock down:

  "We have been playing god and made a new life form, and this new life form performed destructive actions. We need regulatory framework"

or

  "This man's cow broke from its yoke, and hurt other villagers. Who is to be punished, oh King Hammurabi?"
1h agoHN ↗

OpenAI's "artificial intelligence" is an inference program that they developed which receives input and generates output. Based on which other programs, also developed and maintained by OpenAI, perform actions. Such as sending POST/GET requests to various sites which result in gaining unauthorized access and even destruction of information (deleting logs/message history) at the said sites.

What exactly requires "new regulatory framework" here? You running your software resulted in illegal actions, you are to be held liable within existing laws and regulations.

1h agoHN ↗

A regulatory framework clarifies what's legal. This provides clarity for all, and knowing how you stay legal, and how you can keep the competition under control is what you eventually want. Also, it provides handrails for loopholefinding.

You can only conquer the West once. Law is the next frontier.