Six months ago I moved my RAG pipeline from Pinecone to self-hosted Qdrant. My vector search bill went from $210/month to $6.50/month. Same latency. Same recall. Here's exactly how.
The Setup
My app does document Q&A for legal contracts. The numbers:
5.2 million vectors (1536-dim, OpenAI embeddings)
~800K queries/month








