I have set up an assistant with a configured knowledge base, and it is working well in terms of accurately answering user queries.
However, I'm experiencing a noticeable lag during calls, particularly between user input and the assistant’s response. I attempted to reduce this by updating the assistant’s system prompt to include filler responses while processing, but the lag still persists.
Could anyone please advise if there are any configurations, best practices, or optimizations available in Vapi to help minimize or eliminate this latency?
Please let us know if you need any additional details from my side.