← Back to events
ActiveTech

Faiss vs. Turbovec vs. Infino: Comparing 4-bit vector quantization

Photo: Hacker News

What happened

FAISS classic PQ, turbovec, and Infino SQ4 scan the same 100,000 OpenAI embeddings at 4 bits per dimension: recall within two points, latency from 1.5 ms to 45 ms.

Summary assembled by rule from the sources below

Why it's spreading

Sources