Yeah man, that's exactly what I'm talking about. appreciate your perspective. How would that work with Vapi? I have a setup on qdrant which I'd use for the knowledgebase. But I'm pussled how I can intercept the text before it's sent to the LLM to aggregate it with the RAG context