6 ms·
Finally! I was impressed by the together.ai approach to trying the chat models without having to deploy your own and was curious why Hugging Face didn’t have th
by petegordon 3y ago
Finally! I was impressed by the together.ai approach to trying the chat models without having to deploy your own and was curious why Hugging Face didn’t have that type of interface more consistently available. Now to see HuggingFace make as many models available as together.ai does, they still have a good lead in that regard.
- anybodyz 3y agoTogether serves models optimized for inference speed. They're not Groq but Together (and Perplexity Labs) have the lowest latencies and fastest tokens per second of any commercial services available right now. Also the lowest prices afaik.