6 ms·
> Or maybe models get so good that a 30B model is genuinely smart enough to do everything, so nobody really needs a model like Opus or Sol unless they’re trying
by qudat 1mo ago
> Or maybe models get so good that a 30B model is genuinely smart enough to do everything, so nobody really needs a model like Opus or Sol unless they’re trying to solve the Reimann Hypothesis. I don’t really buy this. Models can do frontier mathematical work today while still being not smart enough to refactor large codebases as well as me, so it’s hard to imagine a world where I don’t just want to use the smartest model available.
Idk, I already don’t bother with Opus and stick with sonnet med. I really care more about speed. I use qwen3.6 27b for personal projects and I think it works pretty great.
So like the article mentions, if scaling stalls and small models get better it’s not impossible to imagine a convergence and hardware costs drop.
Having said that, self hosting will be a niche thing like it is today for other services.