does using an external stt remove the users turns ...
# ask-ai
b
does using an external stt remove the users turns frfom the transcrippts dict?
t
Thanks for reaching out, Sebastian. I'm an AI assistant and I'm here to help. Give me a moment while I research this topic.
Using an external STT doesn’t automatically remove user turns from the transcripts dictionary. The transcripts get cleared based on end-of-turn (EOT) detection, which depends on how your turn detection is set up (like VAD, STT, or context-aware models) and what the external STT supports. If your external STT can emit
END_OF_SPEECH
events and you set
turn_detection="stt"
, then the STT’s endpointing logic can help decide when transcripts are cleared. Another option is to use LiveKit’s context-aware turn detection model for a more refined EOT detection, as long as the STT supports streaming and interim results. I’m here to help if you’d like some guidance on configuring these settings or if you want me to clarify anything further! Sources: AssemblyAI integration guide | LiveKit Docs | Improving voice AI's turn detection with transformers | Turn detection and interruptions | LiveKit Docs | Realtime model integrations | LiveKit Docs | agents/src/voice/audio_recognition.ts
b
session = AgentSession( stt=deepgram.STT(model="nova-3", language="en-US"), llm=openai.realtime.RealtimeModel( model="gpt-realtime", turn_detection=None, input_audio_transcription=None voice=config.voice, temperature=config.temperature ), vad=ctx.proc.userdata["vad"], turn_detection=EnglishModel(), ) with this setup i dont have the users turns in the transript
t
The issue with missing user turns in your transcript might be due to a little hiccup between the OpenAI RealtimeModel's built-in turn detection and LiveKit's turn detection model. OpenAI's Realtime API handles speech-to-text internally and only sends transcripts after a turn is finished, which can sometimes clash with LiveKit's turn detection that works best with live, incremental STT results. Here’s a friendly checklist to help you resolve this: 1. Try disabling OpenAI's turn detection by setting
turn_detection=None
in your RealtimeModel configuration. 2. Keep using LiveKit's turn detection model (like
EnglishModel
) as you’ve already set up. 3. Double-check that no other parts of your code are turning OpenAI's turn detection back on or changing the
turn_detection
setting. If you’d like, I’m here to help you walk through the setup or troubleshoot any other questions you might have. Just let me know! Sources: LiveKit turn detector plugin | LiveKit Docs | agents/src/voice/audio_recognition.ts | livekit-agents/livekit/agents/voice/agent.py