JaviG
02/10/2025, 11:22 AM{
'provider': 'deepgram',
'model': 'nova-2-phonecall',
'language': 'en',
}
In this particular case it has transcribed the caller's name by a number "call_sid": "62210c63-92b9-4a22-9074-cae47b394a39"
[Start: 00:00:15] [Voice Assistant] Let's start with gathering some information for the shift coverage process. Could you please provide your full name? [End: 00:00:21]
[Start: 00:00:24] [Caller] 2 0 0 [End: 00:00:25]
[Start: 00:00:27] [Voice Assistant] Could you please clarify if 2 0 0 is your full name if you meant something else. [End: 00:00:31]
We did an initial test using the full recording and view the transcript obtained by Deepgram playground. And the transcript obtained is consistent with the recording.
We thought it might be a streaming problem and with a fragment of the recording it would be different, so we cut out only the part where the name is pronounced and repeated the test.
However, the result obtained in DG was still correct and different from the result provided by VAPI.
The next test was to do the same test but changing to another model of lower capacity nova, and here we got the same result that was obtained during the call.
How can this be possible? Can there be a change of model for any reason?
https://cdn.discordapp.com/attachments/1338470163825033287/1338470164190068746/image1.png?ex=67ab32fa&is=67a9e17a&hm=05ad1b1430b4348589740e7b72911e61ca8eaabf44112b96c1c62a3d7b023873&
https://cdn.discordapp.com/attachments/1338470163825033287/1338470164550783046/image.png?ex=67ab32fa&is=67a9e17a&hm=99126fdcfa57d6aab5184b793a0c3a25156e92dcf26da05475d0ca99f1b7dfe9&
https://cdn.discordapp.com/attachments/1338470163825033287/1338470164903235635/image.png?ex=67ab32fa&is=67a9e17a&hm=62f018ae0a912825e04c36dfb4f186d8d1d5f6b71d0d3025ebab1f9cd525758d&Chiranjeet Mishra
02/10/2025, 12:47 PMJaviG
02/10/2025, 1:10 PMnova model, I have not managed to obtain the value of the transcription above but erroneous results that may be more similar like "Two over, Sean"