Writing
Writing
Retrieval at scale
Six pieces, written in order, that form one argument. It starts at the hardware and works up through storage and index construction, then turns to the embedding model, the economics of running it, and the geometry that decides which index will work at all. The paper on the research page is where the argument currently ends.
-
Part 1
-
Part 2
The Reality of Vector Quantization: Does Product Quantization Deliver on Its Promises?
-
Part 3
Breaking the Single-Thread Bottleneck: Concurrent Vector Graph Construction in OpenSearch
-
Part 4
Teaching Embedding Models New Words: A Deep Dive into Domain Adaptation
-
Part 5
-
Part 6
When Geometry Is Everything: Multi-Vector Retrieval Beyond the MaxSim Monoculture