
4/15/2025 · Daniel Azoulai
What this post added
This post details how Lyzr migrated from Weaviate and Pinecone to Qdrant to improve the performance and scalability of their AI agents. It quantifies the performance gains, including a >90% reduction in query latency (from 300-500ms to 20-50ms P99), 2x faster indexing, and a 30% reduction in infrastructure costs. The benchmarks show Qdrant handling over 1,000 queries per minute and sustaining throughput of more than 250 queries per second with over 100 concurrent agents, demonstrating significant improvements in query throughput, indexing performance, and resource efficiency compared to their previous solutions. The post also includes case studies from NTT Data and NPD, highlighting Qdrant's role in improving retrieval accuracy and low-latency retrieval for AI applications.