Comment by sailingparrot
7 hours ago
It’s the famous “flattening the curve” from COVID. But for LLMs. This release is not flattening anything.
7 hours ago
It’s the famous “flattening the curve” from COVID. But for LLMs. This release is not flattening anything.
The word "pacing" (especially in the phrase "pace yourself") to mean go more slowly (at least initially) didn't originate with Covid. It's been around as long as I remember i.e. at least several decades.
I guess I understand why we'd want to flatten a COVID curve, but why do people want to flatten the LLM development curve? Don't we want the opposite? Isn't the goal AGI?
There is a difference between wanting AGI (which not everyone does), and wanting it as fast as possible no matter the side effects and potential for vast harm. Homo sapiens is 300k years old, maybe it’s ok to delay AGI by like… 1 year if it meaningfully improve our ability to align the model?
Well, there's a tension because, depending on who you ask, AGI is how you cure cancer and achieve utopia, but also how you kill all life on earth and turn the solar system into paperclips
I'm good with the odds on those 2 scenarios. I believe humans could kill all life on earth without AI anyway.
I think the goal is different from what "we" want anyway.
This is a preexisting model being optimized. Its absolutely not some unexpected release after that blog post. I won't defend that blog post, but saying THIS release is proof they don't mean they are slowing down is just incorrect, this is a prime example of what i consider horizontal improvements
Releasing a new fable is an example of straight up vertical progress, releasing a more efficient preexisting opus that is more affordable is an example of horizontal progress, more efficient models rather than higher power models.
The blog post about slowing down is still just some weird self interested post, they want to govern themselves and impose distillation restrictions/gpu restrictions and used some weird blog post about slowing down and fear mongering as usual to justify it, its strange, but slowing down and stopping are not the same thing at all.
What does model naming have to do with pacing or not? This is a ~20% relative quality improvement on the frontier (fable) at ~40% of the cost, just 21 days after the last release.
Intelligence per dollar is the only thing that matters, this is what controls how many agents you can run in parallel, how long you can let them run etc. This is absolutely a step improvement on the frontier and not some lipstick on a harmless second tier model.
pretty annoying topic tbh. You're just weaponizing this dumb blog post so anything released is now a contradiction. By your same logic, if all inference was served at 50% less power cost and the savings are passed on somewhat to the user, its also a contradiction of the blog post.
Its an agenda serving blog post, but constantly bringing it up like this is just obnoxious.
2 replies →