@pravdyvyi Yes, it is absolutely possible to set up a system where your server receives an audio buffer or audio file, processes it (e.g., using voice APIs or AI models like OpenAI or Google Speech-to-Text for transcription), and then retranslates the response back to the client