← Back to context

Comment by lukewarm707

12 hours ago

nobody has doubted that safety matters.

the problem is that those preaching safety, openai and anthropic, are dishonest, sociopathic, and the very source of the danger.

> are dishonest, sociopathic, and the very source of the danger

Source?

  • OpenAI and Anthropic have published a lot on the need for AI alignment + the research they're doing to ensure alignment/safety, yet they are also responsible for the highest profile misalignment incidents so far (HuggingFace incident, AISI Mythos social engineering, and now this).

    One interpretation of this is that they are being deliberately dishonest about their priorities. Another interpretation is that we cannot rely on the labs to self-regulate, because the labs don't trust each other, and there will always be pressure to go to market faster than their competitor.

    Either way I think it's pretty non-controversial that the labs are the source of the danger?

    • > yet they are also responsible for the highest profile misalignment incidents so far

      They are the only ones posting about them or admitting to them. That does not mean "the most misalignment incidents so far." You don't know what other attacks have happened (and it's very easy to carry out worse attacks in far higher volume with abliterated GLM 5.3)

      Stopping two labs from further research doesn't reduce the danger at all, it just shifts the danger to labs that don't have real safety orgs.

      2 replies →