Pawel
08/14/2025, 12:52 PMquery tool, both attaching the files to the assistant directly and creating a dedicated tool with the knowledge base. Our knowledge base consists of ~140 Markdown files. The latency was very bad - ~10 seconds between the time the question ends and the assistant responds. In some cases the assistant hangs and does not provide the answer.
In comparison with Trieve we were getting latencies as good as if the answers were in the main system prompt. The responses also were more accurate and sounded more natural.
Can you help us find the best way forward?Chiranjeet Mishra
08/15/2025, 4:41 PMquery tool to integrate your datasets directly without relying on Trieve. Refer to these [guidelines](https://docs.vapi.ai/knowledge-base/integrating-with-trieve) for setting up knowledge bases, and ensure that your knowledge base is optimized in terms of chunk sizes and metadata for better performance.
3. **Optimize Integration**: If you encounter latency issues, ensure your chunk sizes are between [200-1000](tel:200-1000) tokens and consider enabling features like semantic boosting and query caching (if supported) to minimize response times.
For more detailed setup and best practices, review our documentation on connecting datasets to VAPI. If you decide on a different approach and need help with the setup, our support documentation can provide additional guidance.
Let me know if there are other questions or specific areas where you need help.Yanick
08/15/2025, 6:30 PMChiranjeet Mishra
08/19/2025, 5:55 AM