Set a fixed language in your assistant config to skip detection, enable streaming responses for faster reply generation, use cached or pre-computed replies for common openers, and fine-tune startSpeakingPlan or use smartEndpointing to reduce delay after the user’s first utterance..