5 ms·
TurboQuant can reduce vector index size by 10x at 100M Row Scale
- Slovian 3mo ago[dead]
- 0-_-0 3mo ago32 bits vs 4 bits it looks like
- mxfeinberg 3mo agoYup, and unlike the original turboquant paper, my implementation is pinned to using a 4 bit code book so I could use SIMD kernels for performance.