6 ms·
Elon mentioned that Grok's 4 image and video understanding capabilities are somewhat limited and he suggested a new version of the foundation model is being tra
by manca 1y ago
Elon mentioned that Grok's 4 image and video understanding capabilities are somewhat limited and he suggested a new version of the foundation model is being trained to address these issues. According to the "Humanity's Last Exam" benchmark, though, it seems to perform reasonably well, if not the best among the SOTA models.
I agree, though - the timing of the release is a bit unfortunate and it felt like rushed a bit, since not even a model card is available.
- joaogui1 1y agoThey used a text-only subset of HLE