Over charging and call delay
# support
s
ef2a99f3-2aca-4462-a569-a63fa380e098 @shubham.bajaj, please check and resolve this issue, our calls have stopped, 20s call and 0.20 cents charged! Secondly the delay has now increaassedd!! The endpointing issue or whatever it is please. I am waiting.
z
Same here, what a mess !! It's maybe related to knowledge base Turn latency: 9412ms (transcriber: 298ms, endpointing: 101ms, kb: 7898ms, model: 8249ms, voice: 743ms)
c
Hey, we will be having an Office Hour in 20 minutes. Can you hop in and get your query resolved during it?
s
@Shubham Bajaj
s
@Shahzaib.aiii The call cost is $0.20 because it's made up of different provider components. Here's the breakdown: - LLM (AI Model Processing) – $0.1388 - This is the biggest chunk of the cost. The call involved 21,864 prompt tokens and 87 completion tokens, processed by OpenAI’s GPT-4o. - Speech-to-Text (STT) – $0.0169 - This covers the transcription of the call using Azure for about 0.5 minutes of speech. - Text-to-Speech (TTS) – $0.0146 - The system converted 292 characters into voice using 11Labs’ Eleven Turbo V2. - Voice API (VAPI) – $0.0169 - This cost comes from handling 0.3372 minutes of audio streaming. - Analysis (Summarization & Data Processing) – $0.0138 - This includes summarizing, structuring data, and evaluating the call's success using Claude 3.5 Sonnet from Anthropic. All of this adds up to $0.201, which rounds off to $0.20. Essentially, most of the cost comes from AI processing, followed by voice-related services.
@zxdream can you ping me to your #1211483291191083018 ticket.
@Shahzaib.aiiiAs previously discussed, endpointing is not a concern. It is a measure that requires calculation and application based on your call requirements. I have provided you with the optimal median values, and you should now determine the most suitable values. As per our discussion, the implementation was successful last time. However, as you can observe from the screenshot, its related to the KB. https://cdn.discordapp.com/attachments/1338521531101483040/1339049009666326559/Screenshot_2025-02-12_at_7.05.41_AM.png?ex=67ad4e12&is=67abfc92&hm=770eb97aeb0380fd065b67861d9ecfd49b909ec1d853749ad43507790f33cfbf& https://cdn.discordapp.com/attachments/1338521531101483040/1339049010224435281/Screenshot_2025-02-12_at_7.05.33_AM.png?ex=67ad4e12&is=67abfc92&hm=73b8fc342596d689839680b7d7785daaffa72ea9745363d3259cebc08df51a5f&
@Shahzaib.aiii Could you please confirm whether the final response from the assistant during the call utilized the appropriate number of chunks for both content and quantity? If so, does this necessitate the utilization of 10 chunks? This information will assist me in comprehending the requirements for your KB configuration. https://cdn.discordapp.com/attachments/1338521531101483040/1339049997454278767/message.txt?ex=67ad4efd&is=67abfd7d&hm=c250d1d9dd4f4fb9645ba04caa9ae19ba1ea65fa2c5649158cfa9ae4f7d6d763&
s
Shubham we will have to hop on a call to understand thiss!