Comment by IanCal
13 hours ago
> AI slop,
Now I know there are issues with the field and how just answering these questions may cause broader problems, but I feel like the posted results is far from slop. We can't just call any output slop, or it loses all meaning.
If it was slop, it'd not be causing the issues the group are concerned about - they're not saying "the problem is we're getting loads of incorrect proofs thrown about that are nonsense".
When you blanket a set of things with a pejorative, and it turns out that some of the members of that set are demonstrably and definitively NOT covered by that pejorative, and that all the pejorative means at bottom is "I don't like", all you've accomplished in the long run is to call into question any future legitimate use of that pejorative. It is tempting, especially when heated, to stretch an invective, but it will ironically only lead to the death of its utility over time.
So the fact that the Library of Babel (i.e. all possible books) contains occasional gems means that you can't object to using it on principle? That would seem to follow from your logic.
What about a filtered set "all well formed books"? Or "all well formed books that are plausible enough that they could convince a reasonable person, regardless of their accuracy"?
It's generally taken that a cup of sewage in a barrel of wine makes a barrel of sewage. Surely a reasonable person could claim that a barrel of sewage was still sewage, even if it contained several cups of wine?
I personally just find it hilarious how the complaints and excuses against AI have slowly marched and changed from 2023 to now.
They're not calling any output slop, they're calling indecipherable output slop. The management class responsible for hiring, firing, and paying people doesn't possess the domain knowledge to say for certain whether or not LLM output is optimal (which, in this context, means correct), but they will trust that it's good enough to justify further automation / fewer grant approvals / etc. So in that sense, slop can and will cause the economic issues people are concerned about.
University boards want the prestige of successful research programs. Doing the hard work to get something demonstrably true is going to lose out economically in this paradigm, where we are all being conditioned to uncritically ooh and aah at the incantations being elicited from these magic boxes. The oracles even have legions of zealots who will berate you for not being sufficiently deferential and reverent, or worse, accuse you of blasphemy. If for no other reason, I agree with using the term to express all of the above succinctly, even if LLMs can be helpful tools generally.
I've read some of the results papers (the Einstein condensate one and the pi exponential one). I'm not an expert but it definitely wasn't AI slop. The introduction sections were particularly well framed and informative.
Also you can see in the papers where an idea is introduced but in the bibliography you can see where the foundational idea comes from. So the narratives are not unmotivated as some claim (proof without intuition claims).
In the context of maths papers, the term has come to refer to papers having the shortcomings that are, for whatever reason, typical of LLM out, including things like using non-standard terminology all over the place, emphasizing easy steps while leaping over harder ones, having bizarre organisation, and, importantly, failing to properly cover existing work and as a result being hard to tell from plagiarism.
The degree to which these issues feature will differ, but it is generally the case that converting the output to proper research requires significant effort, hence the AGMAI recommendations being what they are, and not performing that effort tends to come off as laziness or incompetence, so I can see how slop has become the popular term.
> We can't just call any output slop, or it loses all meaning.
The term “ai slop” is not supposed to discriminate good ai output from bad, the entire purpose of the phrase is a blanket term that delegitimizes all ai output.
That is not how it is generally being used.
That is exactly how I see it generally being used. Why else would people be dismissing work as AI slop without even reading it, discovering what it says, or even looking into how and to what extent AI was used in a project? Saying things like "if you didn't write it I won't read it" at the first whiff of an AI smell is absolutely said to delegitimize all ai output.
Or in this specific case, why would someone call these proofs (no one is saying they are wrong) AI slop if not to delegitimize all AI output?
5 replies →
> The term “ai slop” is not supposed to discriminate good ai output from bad, the entire purpose of the phrase is a blanket term that delegitimizes all ai output.
Then it's a useless term and we should all stop using it.