5 ms·
I can't imagine this setup will get more than 1 token per second. I would love to see Deepseek running on premise with a decent TPS.
by 1ba9115454 2y ago
I can't imagine this setup will get more than 1 token per second.
I would love to see Deepseek running on premise with a decent TPS.
- thomquaid 2y agoIt says 4.25 TPS in the first para.
- ricardobeat 2y agoHonest mistake. Some people think HN is just a series of short tweets and haven’t realized they are links yet!
- thomquaid 2y ago4.25 is enough tps for a lot of use cases.
- 4ndrewl 2y agoIt's the modern way. Why read when you can just imagine facts straight out of your own brain.
- plagiarist 2y agoI agree but also found your comment funny in the context of LLMs. People love getting facts straight out of their models.
- weatherlight 2y agoThat's still pretty slow, considering there's that "thinking" phase.
- thomquaid 2y agoTrue, but 4.25 is the number we all want to know.
- october8140 2y agoYou can get 1t/s on a raspberry pi. https://youtu.be/o1sN1lB76EA?si=i8ecEBjLdV0zewFQ https://youtu.be/o1sN1lB76EA?si=i8ecEBjLdV0zewFQ