Posted by jatins 1 day ago
I knew that somewhere in the future, something has to give because you can't just go around saying anything without losing some credibility. You still have some ardent followers in his cult of a subreddit.
But just to be clear: Ed is part of a bigger problem in tech journalism which is characterised by extreme pessimism and excessive skepticism. It is not correct to view Ed in isolation rather to see it as a part of the culture in which he can thrive.
[1] https://news.ycombinator.com/item?id=48447549
[2] https://www.theargumentmag.com/p/ais-biggest-critic-has-lost...
These companies are all heavily buoyed by their investments in AI which is essentially an oroboros of money.
Can I try?
> But when people bring him up, they're of course not generally citing his anger
Wrong.
> Google has been increasing the relative priority of revenue over the user experience over time
Wrong.
> I'm curious what people do after being on the wrong side
Wrong.
Some of the claims categorized as "wrong" are also completely true, such as training hitting diminishing returns. New models are barely an improvement and most people I know stuck on Opus 4.6 over any newer one for example.
Exact same thing for the claim "the fact we're running out of high quality training data and we're hitting the walls of scaling laws, in the training paradigm, these models aren't getting better. What we're seeing today is pretty much what they're always gonna be like".
If anything, model performance has regressed in actual use (i.e. not benchmarks) for the past half a year.
OK, but the first instance of a claim of diminishing returns was in February 2024, when GPT-4 was the best model available. Do you really think improvement since then has been minimal?
1 - Their numbers have also exploded, so I have no idea of any general rule.
What has improved isn't the models, it's the harnesses.
Give GPT-3.5 a 1M context window and a modern harness, and you won't see any meaningful difference with Opus 5.
It's a bit hard to try with such old models, but for example I use Opus 5 / Fable at work and Sonnet 4.5 at home (because it's free via Amazon Q), and there's absolutely 0 difference in performance. None. Obviously 4.5 is only a year old, not 3, but try with any older model that has a decent context window and you'll get the same results.
In fact I'll go further than this and say that models are currently regressing. Opus 5 is much much worse than Opus 4.6 for example, and it's clear that Anthropic (at least - I don't use OpenAI models much) is just tokenmaxing rather than optimizing for performance.
>models are currently regressing. Opus 5 is much much worse than Opus 4.6
I'm in sheer awe at these takes. Literally beyond parody.
You don’t have to spend effort proving him wrong. Just don’t read it and move on with your life. Regardless of which “side” of AI you’re on it’s kinda ridiculous how much effort gets spent on screaming gotcha at this one commentator.
So then we should be calling Dario out every time he opens his mouth, right?