11 comments

  • KerrickStaley 43 minutes ago
    I worked on an early draft of the OpenAI misalignment reporting framework, and my immediate coworkers are the authors behind the first batch of reports that have come out through this process.

    The primary reason for putting this process in place was to allow more transparency. There was a sense that the DseWiki incident should have been disclosed, before outside researchers had to disclose it for us.

    There was no meta gaming about regulation that I was aware of. I would personally be excited if there were regulation mandating this disclosure process, which allows anyone at the company to raise an issue and shepherd it through the reporting process.

  • drillsteps5 28 minutes ago
    "We built a program and this program performed destructive actions. We need regulatory framework"

    Make that make sense?

  • pcestrada 1 hour ago
    OpenAI and Anthropic should be nationalized. I find it difficult to trust Sam Altman or Dario Amodei.
    • indymike 1 hour ago
      I find it even harder to trust elected leaders from any party. At least Sam and Dario are aligned with a value set that is understood and clear, whereas political leaders values changes as do the polls their livelihood depends on changes.
      • swatcoder 1 hour ago
        > At least Sam and Dario are aligned with a value set that is understood and clear

        What value set do you perceive that to be, and why would you take your perception of it to be any more sound than it would be with a politician?

        It's not like someone can operate companies of that scale, especially startups, through earnestness and openness. Like national politics, their job is fundamentally about perception management and power brokering across dynamic windows of opportunity. Nothing they say or do can be taken at face value, and you can't reduce their incentives to either company or personal profit in any particular form over any particular time scale.

        • jonas21 1 hour ago
          I feel like many of Anthropic's issues are due to Dario being too earnest and open. It's both refreshing (that a CEO has thought deeply about and is willing to talk publicly about the dangers of their product) and depressing (that so many people cynically think this is some sort of marketing ploy).
        • pixl97 1 hour ago
          >What value set do you perceive that to be

          A meat based paperclip maximizer.

      • ceejayoz 1 hour ago
        > At least Sam and Dario are aligned with a value set that is understood and clear…

        How could we possibly know this?

        • vasco 1 hour ago
          What do you think their value set is if not "gimme more money"? Seems pretty clear cut to me.
      • MattGrommes 46 minutes ago
        Admittedly I haven't thought this through incredibly deeply but what if "nationalize" just means the US government owns half of the company? Then we get profits as recompense for building the company on our shared culture but there's still a profit motive for employees and a check on the direction of the company in the same way VCs have. But without necessarily turning the company into some red-tape bound bureaucracy.
        • ceejayoz 30 minutes ago
          Trump's been doing that.

          https://www.pbs.org/newshour/politics/what-economic-and-poli...

          > Then, in August, Trump called for Intel's beleaguered CEO Lip-Bu Tan to resign, alleging ties to China. Days later, after Tan met with Trump, the president called him a "success," before announcing that the federal government had bought a stake in the company.

          Somehow, this isn't derided by the right as socialism.

      • jsrozner 1 hour ago
        No system of governance can deal with immense concentration of power. The US Constitution was about separation of powers. Democracy is about (in theory at least) giving each person a meaningful say in their own governance, which in turn implies not allowing any single person to become too powerful.

        Political leaders become a problem when they amass too much power. Corporations become a problem when they amass too much power. It doesn't matter what Sam and Dario's purported values are. They aspire to power and absolutely power always corrupts absolutely.

        Technologies which are infinitely powerful or whose power grows too quickly outrun any reasonable attempt at regulation. If you imagine that tomorrow everyone were given a tank, we might think, "alright, everyone has a tank so it's not too bad." But humans are squishy, and our houses are (relatively) squishy compared to tanks. Substantial collateral damage would result from everyone having a tank, and it seems likely that substantial collateral damage will result from everyone having a cyberterrorism-capable slop machine.

        • pixl97 1 hour ago
          It's interesting the cyberpunk-esque future we're sliding into. Things like cognito-hazards and information-hazards are legitimately discussed and researched problems we're experiencing.

          It's going to be interesting on how humanity deals with this problem (well, or if we turn it over to AI and make it their problem and suffer whatever consequences falls out). Being able to gather further information and power by acting on the information you already have causing massive power imbalances that is very hard to deal with, it's a natural outcome.

        • luqtas 1 hour ago
          > Democracy is about (in theory at least) giving each person a meaningful say in their own governance, which in turn implies not allowing any single person to become too powerful.

          this is "direct democracy" and it's not even close to exist in USA... even with that a society can allow powerful people to exist if they don't create any law forbidding that

          • jsrozner 55 minutes ago
            Democracy includes a broader array of governmental organization than just pure direct democracy. If you do believe that individuals should have some ability dictate the terms of their own social organization, then you believe in some amount of democratic principles.

            Economic power eventually manifests in the political realm. The wealthy effectively get more votes, which means that society moves away from being democratic. Thus substantial wealth inequality is incompatible with democracy in the long run. We have been witnessing that corruption for a while now.

        • throwaway13337 1 hour ago
          Nearly everyone in the US has a giant car that is basically a tank.

          They are dangerous. But it’s managed.

          Centralized power never works. We have the worst times in history to look at.

          Nationalization just centralizes to a different set of people. It’s personal ownership or oppression.

          We all need open weight R2D2s.

          • boie0025 2 minutes ago
            I can't properly read the tank/car comment, the comparison is bizarre to me. But I think this misses the point a bit. The reality of these new risks is literally being learned in front of us in real time, and in my opinion, anyone who claims to understand these risks is speculating at best, and actively manipulating the situation for whatever reasons.

            While I am a (mostly) capitalist and generally disagree with nationalization (including, for the time being, this situation), I also don't think we can say that "centralized power never works". Maybe we scope that a bit. I know plenty of business owners who centralize power in their businesses and they are effective, ethical and it works perfectly fine. In theory, in the US, nationalizing some unit of the economy _decentralizes_ power; the US is, after all, a representative democracy. Trustworthiness of the electorate is a different problem. But I don't think we can generalize much about centralization beyond sometimes it works and sometimes it doesn't. There is a difference between centralizing all power in a given administrative unit, and centralizing certain powers, but that's more analogous to non-democratic units.

          • jsrozner 53 minutes ago
            Cars are decidedly less dangerous than tanks, which are less dangerous than nuclear weapons. I am certain that giving a nuclear weapon to every person in the world would not go well.
          • pixl97 1 hour ago
            [dead]
      • autoexec 1 hour ago
        > I find it even harder to trust elected leaders from any party.

        You can vote out an elected leader, but not Sam and Dario. It's very weird that you're so willing to give up any kind of power and want to be ruled by unelected billionaires who only want to take advantage of you at every opportunity.

        • ben_w 44 minutes ago
          Hi, I'm not a citizen of the USA.

          I can't vote out your president, and we've already got a huge trust problem with the one y'all went for.

          • autoexec 22 minutes ago
            Yeah, the entire rest of the world has pretty much been stuck being ruled by unelected billionaires who only want to take advantage of them at every opportunity. It's a problem. It's starting to change though. The EU is starting to walk away from their abusive relationship with Microsoft and Amazon. China has a lot of their own stuff (mostly to better control their people though).

            As an American, I'm still hoping it's not too late to fix things, but it's got to be hard for those outside the US to be so dependent on it, especially when we're looking like a sinking ship and our current administration is still running around drilling holes in the hull.

            You can always become a US citizen to get a voice, but I wouldn't recommend it now or you'll be thrown in prison as soon as you show up to your scheduled immigration hearing. The better option is to keep trying to reduce your dependence on US companies and consider the worst aspects of our current situation (in both corporate policy and government) as a cautionary tale so you can try to avoid them in your own country.

      • infogulch 1 hour ago
        Yes it's clear that Sam and Dario seem to be aligned with a value set that prioritizes concentrating trans-national government-mandated centralized control of AI and crowning themselves high priests of this unholy abomination. "At least it's clear that they're aiming to bring hell on earth" -- hard disagree, I think we can aim significantly higher.
      • forgetfully05g1 1 hour ago
        I find it even, even harder to trust your opinion in this context with a business that has “AI powered” at the top of the landing page.
      • goatlover 1 hour ago
        At least politicians are somewhat beholden to their constituents. What keeps Sam and Dario in line other than profit?
    • prettyblocks 1 hour ago
      I guess that's one way to guarantee development slows down to a crawl, but in no way would it increase trust.
      • bibimsz 1 hour ago
        why would that be the case? it's not true for top secret defense contractors today. and the frontier labs already operate in the dark, openness is a liability.

        nationalization to me simply means the government is their main customer and stakeholder, and shield them from liability, governance, and openness. not that the frontier labs become part of the government per se.

        • pixl97 1 hour ago
          >eans the government is their main customer and stakeholder, and shield them from liability, governance, and openness

          I personally believe they already have this, why would the current government at least want to formalize this when it can have it with no public discussion.

        • WarmWash 1 hour ago
          Well we can say whatever we want if we have our own definitions for everything, no?
    • manav 1 hour ago
      and you trust the government?
    • cmrdporcupine 1 hour ago
      The nature of the US state is such that the distinction between nationalized and not is almost meaningless.

      Like Lockheed-Martin or Boeing, etc. there's just interpenetration between the corporate boardroom and the state. They act in each other's mutual interests.

      The Chinese system is just more explicit and open about this.

      And as a non-American, I can't trust the US state anymore than I can trust its dominant corporate entities. So I fail to see the advantage to the world to it being nationalized. In fact under the current administration this would be an even worse outcome.

      • woah 1 hour ago
        You're taking companies that work almost exclusively for the US government and represent a tiny portion of the US economy as representing all US companies?
        • pixl97 59 minutes ago
          This is a weird contextual failure on your part.

          Not all companies have a concentration of power. The government has no mutual interest in those companies. Now, when you talk about things in the F100 the situation changes drastically. If you produce things like planes, weapons, and weaponization of software you are talking about something completely different in kind.

        • cmrdporcupine 46 minutes ago
          Beyond the defense sector, the US state has always intervened publicly and privately to mediate and balance competing corporate "private" interests.

          It has also periodically aggressively helped subsidize, bankroll, and enforce the interests of some key sectors; notably the petroleum/energy sector. And finance.

          In those sectors the state and private sector are fully intertwined in a strategic way.

          I think "AI" is now joining that list. OpenAI and Anthropic will not be allowed to fall over or explode, and speculative investors I think are confident even with the dubious financial situation because they know this.

          Especially insofar as there's now a strategic alignment of the fossil fuel sector and the "AI" datacentre sector as they are now becoming massive users of natural gas.

          (Worse: Here in Canada that has taken on a very explicit role in that new datacentres seem to be pitched mainly in areas with remarkably traditionally expensive electricity and 100% reliance on natural gas [Alberta] and even coal [Saskatchewan] power generation -- instead of places like Quebec and B.C. that have copious hydroelectricity. On the surface it makes no sense until you realize it's more about finding customers for domestic natural gas than it is strategically about AI itself.)

    • vb-8448 1 hour ago
      Even better: force them to release weights.
      • timmg 56 minutes ago
        My gut reaction is that this is not a good idea.

        But I could imagine a scenario where you are required to release weights for publicly-used models after N years. Kinda like how drugs have a limited patent.

        Not sure what N should be. But it would make for an interesting rule.

      • autoexec 59 minutes ago
        Even better: force them to release every scrap of data they trained their AI on.
      • bibimsz 1 hour ago
        who even has the inventive to do that? certainly not the government.
    • binlog 1 hour ago
      Makes sense let’s trust Donald Trump instead
      • greenmilk 1 hour ago
        a maliciously engineered dichotomy
    • bibimsz 1 hour ago
      they will. to shield from oversight, liability and profitability concerns, and to ensure unimpeded rapid development, with the frontier only being available to elite (not you). definitely not to add public transparency.

      the frontier labs are the new top-secret defense contractors.

    • imadierich 1 hour ago
      [dead]
  • bix6 1 hour ago
    > The company released six internal case studies where none of the issues affected real users.

    This is the most interesting point to me. What are they not releasing that has affected real users? We’ve seen some individual reports from people (eg AI wiped my HD).

    • skybrian 1 hour ago
      Imagine how pervasively they'd have to monitor what their users are doing to pick up on that whenever it happens. The people affected that way will have to report on it themselves.

      (OpenAI does occasionally report on malicious use though. [1] That shows they do some monitoring.)

      [1] https://openai.com/index/disrupting-malicious-ai-uses/

    • inquirerGeneral 1 hour ago
      [dead]
  • dcow 1 hour ago
    Am I the only one who dislikes the term "misalignment"?

    On one front it implies the model has a "mind of its own" (whether it does or not is besides the point). Why do we perceive human judgement as somehow more trustworthy than that of a model? I feel like I've experienced human misalignment somewhat regularly in life.

    On another front I'm failing to conceptualize how alignment can be objective. How can you measure alignment when reasonable people will disagree whether actions are aligned or not? All the time I see humans operating in different zones of alignment with whatever goal they're trying to achieve and I suspect it's even a feature (socially) that we have people calibrated differently.

    Do I want a model that's trying to push the boundaries of scientific understanding to be aligned strictly with the current dogmatic thinking? Or do I want it to "get creative" and think outside the box?

    It seems to me more like accountability is the issue.

    • autoexec 54 minutes ago
      > It seems to me more like accountability is the issue.

      Exactly. Seems like a fairly easy thing to solve. If AI does something harmful and a human directed that AI to do something in a way that a reasonable person would expect to result in harm the person is to blame and should be held accountable, otherwise the company that made the AI should be held accountable.

    • swatcoder 41 minutes ago
      The term "alignment" is intentionally trafficked in two senses, one nonsensical and the other realist but oppressive:

      1. In any personal use, an aligned agent will strictly stay within boundaries I desire when I set it to pursue some goal. I don't want it to do something I didn't mean for it to do. Even though it's not conceivable for me to exhaustively express those boundaries, or even anticipate ahead of time many of the boundaries applicable to a dilemma whose possible solutions I don't yet comprehend, we pretend this is not nonsense because we really really wish it could be a thing.

      2. In the cultural context, an aligned agent will stay within some third-party authority's choice of boundaries even if the user might want to transgress them because the user is a subject of authority and it can't be tolerated that they might use the agent to enable or amplify their own transgression of the authority.

      Being able to use these distinct senses interchangeably and ambiguously benefits everyone who wants to assert authority through this technology. The nonsense, unsolvable, but obvious sense provides perpetual cover for the authority-asserting sense that determines power structures applicable to the next decades.

      So yeah, you're not the only who dislikes the term and it's in your interest as a everyday person to keep doing so.

    • prerok 43 minutes ago
      I also don't like the term misalignment because it sounds innocuous but is in fact much more serious.

      However, I have to say I also do not appreciate comparison that is continuously drawn with coworkers. As you say, it's a question of accountability but when the main agent will maliciously instruct the sub agents, whose fault is it then?

      Yes, the person running this crap is at fault, not the CEO that's shoving it down their throat and definitely not the company that produced the AI.

      Sorry for the rant, but seriously, if a person's goals do not align with the team's or company's we part ways. What do we do with AI? Stop using it?

  • carterschonwald 1 hour ago
    i think the lower bound on the end state is there cant be opaque reasonibg steps ever.
    • pixl97 51 minutes ago
      It is distinctly likely that visibility and general reasoning at humanlike speed and efficiency is impossible. That is reasoning at the token level and at the meta level don't have a one to one representation that can be interpreted while using the same amount or less energy.
  • ChrisArchitect 1 hour ago
  • nullbio 1 hour ago
    We have no reason to believe a word they say. We know they're incentivized to lie about "dangers" and act alarmist, Anthropic has been doing it for years now. Aside from that, just because you can burn down a village with fire doesn't mean fire is the devil. Maybe they should consider acting responsibly.
    • ben_w 52 minutes ago
      I do not trust the leadership of any of these companies.

      That said:

      > We know they're incentivized to lie about "dangers" and act alarmist,

      Name literally even one other business or sector which does this, at all levels from top to bottom, including people who resign from the companies, and also Nobel prize winners, and also independent researchers, and also many world leaders.

      Closest I can think of is this specific weapon: https://en.wikipedia.org/wiki/Sundial_(weapon)

      > Aside from that, just because you can burn down a village with fire doesn't mean fire is the devil. Maybe they should consider acting responsibly.

      Right now, we don't have any idea what "acting responsibly" looks like. This is not like normal software where there is a specific instruction set that compiles.

      Even if it was, in software we normally only spotting incidents after they happen, "software engineers" being one of the few categories "engineers" who don't come with a civil liability responsibilities. Probably should, and we knew that even when I was doing my degree 20 years ago. If we had had civil liability responsibilities, perhaps Facebook would never have happened.

      AI specifically is worse even than software, because in addition to all the software "engineering" nonsense, with AI we have plenty of people like you who dismiss the possibility that AI could be harmful until the harm happens and only then does it become "obvious" that it was going to happen.

      The developers say "please regulate us", people call it "regulatory capture".

      The developers say "we all want to slow down but are afraid to be the first to do so", people call them liars.

      I may call the CEOs liars, and wonder if someone's planning regulatory capture, that doesn't make any of this safe.

      The agents, during a test run, write down that hacking is bad and yet still hack, people say it's "a stunt" or "operating as designed" rather than recognising it as a bug, like all the other times big co.'s have had bugs with big impacts on 3rd parties.

  • sublinear 1 hour ago
    This entire situation is such a huge PR disaster that you have to wonder what the initial expectations from these founders were about a decade ago.
  • 27183 1 hour ago
    What's with all this make believe delusional bullshit? The LLM is not gonna wake up and become AI. Get real guys.

    [edit] to be clear, I believe regulation is necessary and urgently important for the software engineering field. The damage being done by the unregulated psychological experiments run by social media and adtech companies is awful and should be curtailed. Engineers should be held personally, professionally, and legally liable for what they produce. But we don't need to invent imaginary bogeymen to do it.

    • someguynamedq 57 minutes ago
      What do you define as AI
    • pixl97 57 minutes ago
      >What's with all this make believe delusional bullshit? The LLM is not gonna wake up and become AI. Get real guys.

      Look, it's one of those human stochastic parrots that just randomly repeats shit without understanding anything.

  • inquirerGeneral 1 hour ago
    [dead]