Comment by calf

1 hour ago

Since LLMs have rudimentary internal models for concepts it is implausible that they cannot also have one for some simple sense of self. But a lot of people prefer the reductive explanation that the machine has never actually decoded anything at all from the language corpus.