Hi,
We're using Deepgram's Flux model for transcription and noticed that changing the endpointing value in our configuration directly impacts latency. According to Deepgram's documentation/support, endpointing should not affect behavior when using the Flux model.
However, in our testing, adjusting the endpointing value (e.g., from the default to a higher/lower value) produces measurable changes in latency.
Could you clarify:
Is endpointing expected to be ignored when using the Flux model, or does it still apply?