Comment by stereolambda

11 hours ago

It's interesting that there must be a decision behind that, even if it's just appealing to the RLHF judges for some reason. Maybe there's an intention that if you cannot decipher what the chatbot is saying to you, you will have to ask and burn even more tokens.

Naively I would often expect it would talk to me about various niche topics like to a layman, which does occur about some topics an actual normal person would ask.

I used to assume it was the fault of average human annotators. That the people who are paid peanuts to rank chat outputs preferred the pretentious sounding ones. It wasn’t until quite a bit into the LLM boom that companies started to pay for domain experts. Overt watermarking is another possibility.

I’ve noticed that Sol is pretty good most of the time, but with long contexts it’ll start to devolve into Claudish.