interesting. thanks for the feedback. we have our own text ai application (non vapi - ill have to see if theres a way we can start tracking this for voice in vapi) and across a lot of tests TTFT is fastest on gemini 2.0 flash (by almost 30% on average and with way less variance). Makes me wonder if there is something about the way it is orchestrated in Vapi that slows it down?