5 ms·
In the recent Dwarkesh Podcast episode Jensen Huang (Nvidia) said that virtually nobody but Anthropic uses TPUs. How does that add up?
by philippta 5mo ago
In the recent Dwarkesh Podcast episode Jensen Huang (Nvidia) said that virtually nobody but Anthropic uses TPUs. How does that add up?
- deleted 5mo ago[deleted]
- modriano 5mo agoMaybe he's never heard of Google.
- sarchertech 5mo agoWho is the other frontier lab other than Anthropic, OpenAI, and Google? I thought they were ahead of everyone else.
- DeathArrow 5mo agoFolks who make Deepseek, Qwen, GLM, MiniMax, Kimi and MiMo.
- SwellJoe 5mo agoThey're at the frontier of last year. They compete with Opus 4.5. They don't yet compete with current frontier models. They'll presumably catch up, there is no monopoly on talent held by the US. And, that's more true than ever now that the US is actively hostile to immigrants. Scientists who might have come to the US three years ago have little reason to do so now.
- DeathArrow 5mo agoSince Gemini 3.1 Pro is considered to be at frontier and GLM 5.1 does better than it in coding benchmarks it would be fair to say GLM 5.1 is a frontier model.
- chronc6393 5mo ago> Scientists who might have come to the US three years ago have little reason to do so now. Been saying that about EU and China for decades now. Yet the top European and Chinese still come to the US. Even in April 2026.
- sfink 5mo agoNit: scientists have the same reasons to do so now, the same as ever. They just have additional reasons to not do so. But even that distinction is only temporary, since we're determined to piss away any remaining research lead that draws people in. Hopefully the next administration will work at actively reversing the damage, with incentives beyond just "we pinky-promise not to haul you at gunpoint to a concrete detention center and then deport you to Yemen".
- sofixa 5mo ago> Hopefully the next administration will work at actively reversing the damage, with incentives beyond just "we pinky-promise not to haul you at gunpoint to a concrete detention center and then deport you to Yemen". Won't be enough to undo the damage. The US would have to do a full about face, prosecute crimes of the current administration and enact serious core reforms to make it impossible for things to drastically change again in 4 years. Also known as, never going to happen because even the current opposition party doesn't actually want structural change. The world has seen how bad the US can get from a single election, and that isn't changing any time soon.
- lanstin 5mo agoIt's kind of hard to say this unless you go out of your way - the scaffolding for interacting with the raw model is a lot better now for many tasks. Is it that 4.7 is so much better than 4.5 or claude 1.119 is so much tuned to squeeze utility out of the LLM despite the hallucinations and lack of self awareness etc. Certainly the current products are great, but I think it's hard to separate the two things, the raw model and the agent workflow constraining the model towards utility.
- DeathArrow 5mo agoI am using Claude Code with GLM, MiniMax, Kimi and MiMo.
- SwellJoe 5mo agoYou can use Claude Code with other models, so one could test that theory. https://openrouter.ai/docs/guides/coding-agents/claude-code-integration https://openrouter.ai/docs/guides/coding-agents/claude-code-...
- sarchertech 5mo agoYeah I thought all of those were generally acknowledged to be a little behind the big 3.
- csunoser 5mo agoI am not sure what context Jensen said that. But midjourney uses tpu. Apple uses tpu. They are no other frontier labs that use it, but Google + Anthropic is 2 out of 3 frontier lab so..... You could reasonably say that "A majority of frontier labs uses TPU to train and serve their model."
- Hendrikto 5mo agoAfaik, TPUs are only used for inference, not training. Maybe that was also what the quote referred to.
- csunoser 5mo agoMayhaps! But I think as far as google, anthropic[1] and apple[2] goes, they do use the tpus for training. Ofc v4 and v5 (older generations of tpus) were more specialized for search related embedding workloads and i could see people not using them for training. [1]: We train and run Claude on a range of AI hardware—AWS Trainium, Google TPUs - April 6th, Anthropic on Google and Broadcom partnership [2]: "[Apple foundation model]... builds on top of JAX and XLA, and allows us to train the models with high efficiency and scalability on various training hardware and cloud platforms, including TPUs and both cloud and on-premise GPUs" - Apple in 2024
- deleted 5mo ago[deleted]
- arw0n 5mo ago> How does that add up? He's been saying whatever is good for Nvidia for years now without any regard for truth or reason. He's one of the least trustworthy voices in the space.
- luckydata 5mo agoJensen hallucinates more than any llm, he just speaks without thinking all that much about what he says and he generalizes a lot. Trying to hold him accountable to imprecisions and gross simplifications is just going to frustrate whoever tries without changing one bit of his behavior.
- bandrami 5mo agoYou're asking why a businessman would downplay the use of a competing product line?
- Zetaphor 5mo agoThis is the same guy who said OpenClaw was the most important software release ever. Statements like this make me question how technically competent these tech CEOs are
- KptMarchewa 5mo agoYou should instead question how honest they are.
- munk-a 5mo agoIs technical competence the primary measure of tech CEOs at this point? Points vaguely at Elon Musk and the upcoming IPO
- VirusNewbie 5mo agoHe forgot one other big company that uses TPUs besides Anthropic...