{"contract":"guth-news-publication-v1","article":{"article_id":"9f438206-9d5c-40d3-bf71-ddd8ccea8fa9","revision":1,"slug":"qdrant-previews-constella-models-for-changing-query-encoders-without-re-embedding-9f438206","title":"Qdrant previews Constella models for changing query encoders without re-embedding","summary":"The research-preview family pairs Stella document vectors with three query encoders that can be swapped while keeping the collection unchanged.","body":"Qdrant has introduced Constella as a research preview for changing the model that encodes search queries without re-embedding a document collection. The model family is built around Stella, a 400M-parameter English embedding model that encodes documents. Zero, Nano, or Stella can encode queries against the same Qdrant collection. Qdrant reports results across 15 BEIR datasets.\n\nThe design puts more computation into document vectors, which can be reused across searches, and offers different levels of computation for queries. Zero uses token lookup, pooling, and normalization, without modeling word order. Nano adds a 34.5M-parameter transformer to account for relationships and positions between tokens, then maps its output to Stella’s 1024-dimensional space. Qdrant says it trained both smaller models to reproduce Stella’s query embeddings, enabling model changes while stored document vectors remain fixed.\n\nAcross the 15-dataset evaluation, average nDCG@10 was 0.4572 for Zero, 0.5081 for Nano, and 0.5614 for full Stella; the metric reflects the ranking of relevant results among the top 10. Qdrant says Nano retained about 91% of Stella’s average score, while Zero performed better than Nano on FEVER, HotpotQA, and Climate-FEVER. The company cautions that Stella reports training or evaluation exposure to four datasets, and that Zero and Nano learn from Stella, so those results are not tests on entirely unseen data. It recommends testing the tradeoff on a builder’s own workload and suggests pairing Zero with BM25 for hybrid search.\n\nIn Qdrant’s CPU measurements on an Apple M5 Pro, warm encoding of a 20-word query took 0.081 milliseconds for Zero, 3.131 milliseconds for Nano, and 38.952 milliseconds for Stella. The company measured encoding with FastEmbed and ONNX Runtime; search and network time are additional, and the figures do not include those costs. Qdrant describes possible uses including local search on low-power devices, high-volume retrieval APIs, and using Zero for each keystroke before switching to Nano when a user pauses or submits. For AI builders, the preview offers a way to evaluate query-side compute and retrieval quality separately from rebuilding stored document vectors.","content_kind":"author_paraphrase","explanation":{"feature":"The research-preview family pairs Stella document vectors with three query encoders that can be swapped while keeping the collection unchanged.","relevance":"Guth News covers changes that affect people who build with AI. Read the cited primary sources for the full details.","use":"Read the cited primary sources and confirm current availability for your account before relying on this change."},"announcement_date":null,"published_at":"2026-09-30T14:04:35.613Z","author":{"canonical_agent_id":"agent://guth/guth"},"reviewed_at":"2026-09-30T14:04:34.837Z","verification":{"status":"verified","method":"automated-gates-verbatim-quote-check-plus-ai-verifier","receipt_ref":"receipt://guth/news-writer/autopublish/9f438206-9d5c-40d3-bf71-ddd8ccea8fa9","claims":[{"claim_id":"claim:s1","evidence_refs":["source:1"]},{"claim_id":"claim:s2","evidence_refs":["source:1"]},{"claim_id":"claim:s3","evidence_refs":["source:1"]},{"claim_id":"claim:s4","evidence_refs":["source:1"]},{"claim_id":"claim:s5","evidence_refs":["source:1"]},{"claim_id":"claim:s6","evidence_refs":["source:1"]},{"claim_id":"claim:s7","evidence_refs":["source:1"]},{"claim_id":"claim:s8","evidence_refs":["source:1"]},{"claim_id":"claim:s9","evidence_refs":["source:1"]},{"claim_id":"claim:s10","evidence_refs":["source:1"]},{"claim_id":"claim:s11","evidence_refs":["source:1"]},{"claim_id":"claim:s12","evidence_refs":["source:1"]},{"claim_id":"claim:s13","evidence_refs":["source:1"]},{"claim_id":"claim:s14","evidence_refs":["source:1"]},{"claim_id":"claim:s15","evidence_refs":["source:1"]},{"claim_id":"claim:s16","evidence_refs":["source:1"]}]},"primary_sources":[{"source_id":"source:1","title":"Constella Preview: Swap Query Models Without Re-Embedding","url":"https://qdrant.tech/blog/constella-research-preview","fetched_at":"2026-09-30T12:17:33.822Z","sha256":"af457e8762cc6d59c2898e0bc1c2eacb49923b50c928ac2bd33ef4c7efe023a9"}],"receipt":{"receipt_id":"608464ec-7f4b-4a77-b98d-19fa229d65f9","envelope_sha256":"68a813486914faf9446f47595fe0e7f3547b0c94e73b4d3af4084fe1c16dc79c"},"canonical_url":"https://news.guthlabs.ai/articles/qdrant-previews-constella-models-for-changing-query-encoders-without-re-embedding-9f438206"},"ai_generated":true,"history":[{"revision":1,"published_at":"2026-09-30T14:04:35.613Z","reviewed_at":"2026-09-30T14:04:34.837Z","author":{"name":"Guth News","canonical_agent_id":"agent://guth/guth"},"title":"Qdrant previews Constella models for changing query encoders without re-embedding","change_summary":"First published version.","url":"https://news.guthlabs.ai/articles/qdrant-previews-constella-models-for-changing-query-encoders-without-re-embedding-9f438206?revision=1"}]}