How to Tune Vector Search Without Guessing
Dylan Couzon
August 24, 2026

Google's EmbeddingGemma 2 in Qdrant: 99% of retrieval quality in 30x less vector RAM, or 94.5% in 77x less with rescoring.
Qdrant Team
October 06, 2026

Swap query models without re-embedding documents. Explore Constella Zero, Nano, and Stella, with BEIR-15 results and CPU benchmarks.
Dylan Couzon
September 29, 2026

Learnings from writing a cross-encoder inference library in Rust and comparing it with existing Python packages.
Clelia Bertelli
September 25, 2026

How Lucy built a retrieval layer over 2.8M+ SEC and DART filings on Qdrant: four-lane hybrid search, payload filtering, and TurboQuant, reaching 94.7% Recall@10 on FinanceBench.
Daniel Azoulai
September 21, 2026

How Himalayas moved remote job matching from keyword search to hybrid retrieval on Qdrant Cloud: 25M+ requests a month, around 30 ms median time in Qdrant, and a 19-point recall gain.
Daniel Azoulai
September 18, 2026

Fix multilingual embedding language bias with SHIFT: shift document and query vectors to one pivot language and store them in a single Qdrant index.
Evgeniya Sukhodolskaya
September 16, 2026

How Bayer built myGenAssist on Qdrant Hybrid Cloud: 135M points, hybrid search for deep agents, semantic caching, multitenancy, and a 20% efficiency gain.
Daniel Azoulai
August 13, 2026

Hybrid search, payload filters, and late-interaction reranking in Qdrant plus Minima-optimized inference delivered 2.92x more successful agentic RAG tasks per GPU-hour without reducing grounded quality.
Qdrant and Minima Engineering
August 13, 2026

Compare pre-filtering vs post-filtering in vector search: where each breaks, how Qdrant filters in place, and when ACORN earns its cost.
Dylan Couzon
August 07, 2026