Comment by Bluestein
1 day ago
It really is going to get to the point where mathematicians are going to start inserting obvious "tells" in their proofs - like map makers used to do with "trap streets" and similar non-existent features, to catch copies.-
Do you think the training process will retain these artifacts? I doubt it. If they were simply stealing the content - sure it would make sense - but I suspect they’re feeding it into training data and RL might distill these out.
I am thinking along these lines. You do raise a great point.-
https://www.anthropic.com/research/small-samples-poison?from...