What LLM are you guys using in your Vapi Squads?
# prompt-tip
d
We're finding that nothing but the top ChatGPT model will do, but then we struggle with some latency.
p
gpt mini 4.1 or gpt 4.1 seem fine. also openai dropped new ones with similar pricing, i think it was 5.3
I also recommend gemini flash 3.1 when it is enabled on vapi
b
optimized for latency or deep thinking?
u
It's better to stick with the old models, these 5 has very high latency
d
I wish it was easier to use the Claude models. We're finding that they can't search the knowledge base.