I think gpt4o mini will be perfect for most Vapi use cases.
It looks like it will be on the same level as llama 3 70b, just faster.
I experimented with llama 3 70b in the past and it was definitely intelligent enough for a conversation and reliable tool calls, though the latency was 200-400ms more than gpt4o so had to revert back.