#benchmark

5 postswith this tag

Storing 31 GB of Text in 3 GB (and Recovering Every Byte): the Storage Proof of Concept

Fourth installment of the benchmark: we built the compressed corpus on 10,000 real documents. sqlite-zstd compresses 10.9x with random access, chunks are stored as 110-byte recipes instead of text, the word-search index ended up 15x smaller than feared, and TurboQuant quantization ties full-vector quality. Everything fits in ~11 GB.

Joaquín Bravo ContrerasRead more

Hybrid search: how to combine BM25 and embeddings (and make them fit on a laptop)

Third installment of the benchmark: we merge the BM25 and embedding rankings and the result beats both individually. We also measure how to index the full corpus from the Mac M3: what works (GGUF/Metal), what does not (large batches, fp16), and why jina's binary quantization decides the storage architecture.

Joaquín Bravo ContrerasRead more