Comment by ashleyn

4 hours ago

If you were wondering the same thing I am - it's not about skills loss, quality, and less about money spent. It's more about frontier AI shops dogfooding their own models.

If it’s anything like AWS there’s hundreds of people making bespoke software factory setups, enhanced interfaces for ai tools, spinning up 10 parallel review agents with the best model available, etc because the budget is basically unlimited.

Meanwhile https://www.reddit.com/r/GeminiAI/comments/1wh0qxq/google_fi...

I dunno, it could be about money, too. $100,000 budget per employee, per month. Why were they willing to spend that much on AI usage? I hope said employees also make that much in salary.

I think if you talk to LLMs and give feedback or openly say what works and what doesn't, you are essentially solving a captcha and produce accurate training data, while you pay for the token spend. I'd be a bit nervous with this lol.

Just one unsanitized input and you leak info. Or one hidden character and code may or may not belong to you anymore. Its very odd on many levels

  • Accurate training data?

    At best you produce some noisy signals that are going to have a tiny impact if even that.

    And that's on a personal plan where you didn't opt out of sharing usage data.

    Business plans offer zero data retention. This is a non-issue.

    • These are companies famous for following the rules when it comes to handling other people’s data and IP after all, totally a non-issue and they would never violate contract law

      3 replies →

Dogfooding their own model ls and not letting their competitors use their data to train their.

I wonder why they allow it at all.

Like its a no brainer to force your employees to use your own models, then RL train them to be better.

  • Only if all you care about is developing models. I assume the rest of the business would rather just use whatever's best in class regardless of who made it, so I'm sure it's not that straightforward of a decision.

  • If you assume that noisy general usage data enables good RL, particularly compared to curated RL training sets.

    I am not convinced that's the case.

if AI never happened there's like zero chance I would've ever used or noticed the usage of the word "dogfooding" lmao i hate this timeline