6 ms·
Google rolled out TPUs in 2015. AWS released Inferentia and Trainium chips in 2020. If companies working on ML-specific chips was evidence that large transform
by calebkaiser 22d ago
Google rolled out TPUs in 2015. AWS released Inferentia and Trainium chips in 2020.
If companies working on ML-specific chips was evidence that large transformer models have fully saturated their potential, the field would have been done circa GPT-2.
- edgyquant 21d agoNeither of those companies core business model was serving llms
- calebkaiser 21d agoWhat? Both of those companies absolutely serve LLMs, and both of them would love for serving LLMs to be an even bigger part of their business. Not only that, AWS is Anthropic's primary compute partner for training and inference. They literally use the newest generation of the Trainium chips I mentioned before: https://www.anthropic.com/news/anthropic-amazon-compute https://www.anthropic.com/news/anthropic-amazon-compute Chips are another axis for improvements in training and inference. Orgs large enough to explore the space have been doing it for at least a decade now. This is just a silly line of reasoning based on the faulty assumption that somehow, looking for increases in efficiency in training/inference means teams have reached some theoretical limit in model capability.
- edgyquant 20d agoNotice how I used the words “core business” but that they didn’t do business at all