Comment by qwerpy
5 hours ago
I agree with them. In many of the subs I used to frequent, the outrage is high and the signal is low. LLM really is the best way to consume the content there. Retrieve the new posts of the day/week, filter out the outrage and politics and low-effort Reddit in-jokes/copypasta, then summarize the various comment threads. My time is too valuable to sift through the utterly mind-numbing nonsense to get the occasional new piece of info. If you want to dig deeper, ask the LLM to dig in. Or have it regurgitate verbatim a pre-filtered selection of comments.
Reddit is a good source of new happenings and discussions, but I’m willing to spend some tech company’s compute to clean it up for me. For example, I often want to branch out from a particular musical artist. Reddit has many good recommendations. LLM is perfectly capable of gathering these.
To many bots on reddit, so your solution is to go directly to the bot!!?
If bots bury the signal in noise then an LLM is only going go back and give you the summary of the noise.
The bots (and humans) I’m trying to weed out are the pot-stirring permanently outraged/bitter commenters. For example on a “car overturned on X road” post, people getting into fights about whether BMW or Subaru or Tesla drivers are worse. That’s all noise. I want to know when it happened, which direction of traffic, if it’s cleared up yet, etc. LLM works well to remove the off-topic garbage. If someone is botting the actual content (recommendations, reviews, etc) and the up/downvoters are unable to identify it as a bot, I probably won’t be able to either. So I’m not worse off using the LLM.
Worst case I get some pre-digested garbage and I waste less time, best case the LLM pulled out a few good nuggets of info. It’s not a perfect system yet. There’s a lot to tune with how deep to go, what constitutes garbage vs signal, and so on. At least it’s fun :)