5 ms·
This is using think time compute and reinforcement learning. I think this is going to plateau even faster than the initial LLM scaling though.
by Ferrus91 1y ago
This is using think time compute and reinforcement learning. I think this is going to plateau even faster than the initial LLM scaling though.