If it started within the last hour, it’s likely either Claude Haiku load or request-side overhead, not Vapi core.
Try switching to a fallback model temporarily to confirm, reduce prompt/tool payload size, and ensure streaming is enabled. Also check if any tool calls are adding extra latency.
If it persists, share a call ID to Vapi support #1211483291191083018