6 ms·
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at
by Slartie 2mo ago
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.
- EwanToo 2mo agoUsing more tokens is a significant problem if you pay per token?
- Slartie 2mo agoNot if the price per token is significantly lower. Also this arm of the discussion was about speed, not price.
- locknitpicker 2mo ago> Using more tokens is a significant problem if you pay per token? Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.