Comment by lazarus01

10 hours ago

I was thinking about this yesterday, while going through some slop documentation generated by deepseek. If you compare claude fable 5, side by side with the deepseek, the differences in prose are glaring. It's not even close. Deepseek prose reads like fragmented shorthand, where claude fable 5 comes fairly close to human, certainly not superior to human quality writing.

I watched a podcast with a cognitive scientist and one of main contributors to the theory of linguistic relativity, Lera Boroditsky.

She said something to the effect that, "in this very moment, we are speaking in ways that were never spoken before. We are saying things that no other person has said before...."

Language models are not sample efficient and cannot adapt to evolving language, unless it's documented in large amounts of examples.

So whatever isn't documented, whatever isn't in the training dataset or the rag corpus, the model will always be incredibly different in expression from humans.