← Back to context

Comment by derangedHorse

5 hours ago

The only way it would be able to tell you the ending is if it was somehow given more importance in pretraining, loaded into context, or represented in multiple sets of training samples. I have a blog that I make very LLM friendly and usually load posts up into context when I’m working on something relevant. I’ve also opted to improve models for everyone. Despite this, the model can’t recognize my site or any of my posts when I ask it to recall without internet usage (I also turn memory off btw).

That's not correct for a SOTA model with trillions of parameters. Those have immense amounts of knowledge trained into their parameters. Try it. It "remembers" the ending of random books, Goosebumps and otherwise.

That's the entire point of why knowledge cutoff is so important (if you use them offline).

I find it entirely believable that the Navier-Stokes conversation was auto-flagged as high value training data, and burned-in the models knowlegdge base.