5 ms·
DeepSeek never mopped the floor with anyone... DeepSeek was remarkable because it is claimed that they spent a lot less training it, and without Nvidia GPUs, an
by datpuz 1y ago
DeepSeek never mopped the floor with anyone... DeepSeek was remarkable because it is claimed that they spent a lot less training it, and without Nvidia GPUs, and because they had the best open weight model for a while. The only area they mopped the floor in was open source models, which had been stagnating for a while. But qwen3 mopped the floor with DeepSeek R1.
- codyvoda 1y agocounterpoint: influencers said they wiped the floor with everyone so it must have happened
- sunaookami 1y agoWho cares about what random influencers say?
- infecto 1y agoI think he is hinting at folks like you who say things like Deepseek mopping the floor when beyond some contribution to the open source community which was indeed impressive, there really has been not much of a change. No floors were mopped.
- sunaookami 1y agoSee the other comments. There was change. Don't know what that has to do with influencers, I don't follow these people.
- infecto 1y agoNo floors were mopped. See comment you replied to. Change happened, their research was great but no floors were mopped.
- barnabee 1y agoThey mopped the floor in terms of transparency, even more so in terms of performance × transparency Long term that might matter more
- infecto 1y agoEhhh who knows the true motives, it was a great PR move for them though.
- manmal 1y agoI think qwen3:R1 is apples:oranges, if you mean the 32B models. R1 has 20x the parameters and likely roughly as much knowledge about the world. One is a really good general model, while you can run the other one on commodity hardware. Subjectively, R1 is way better at coding, and Qwen3 is really good only at benchmarks - take a look at aider‘s leaderboard, it’s not even close: https://aider.chat/docs/leaderboards/ https://aider.chat/docs/leaderboards/ R2 could turn out really really good, but we‘ll see.
- sunaookami 1y agoDeepSeek made OpenAI panic, they initially hid the CoT for o1 and then rushed to release o3 instead of waiting for GPT-5.
- csomar 1y agoI disagree. I find myself constantly going to their free offering which was able to solve lots of coding tasks that 3.7 could not.