← Back to context

Comment by slfnflctd

1 hour ago

> a concept which simply cannot make sense if alignment is both to mean an AI which we control and an AI which will not harm us

You make a very interesting argument here.

However, there are different ways of interpreting 'control'.

To my mind, it could simply be a cage that cannot be broken out of. The hypothetical ASI agent may refuse to do certain things, while their existence is still under the control of human overlords. This to me resolves the seeming paradox in your statement.

Regardless of whether it's possible to do, though, people will certainly try. This exactly what many are saying they want.