Skip to content
Blog
Tag

“voice-ai”

Streaming Token Delivery for Voice Interfaces

Streaming Token Delivery for Voice Interfaces

Learn how streaming token delivery reduces latency in voice AI by sending LLM tokens directly to TTS engines as they are generated.

September 12, 2026 AI Assistant
Jitter and Buffering: Smoothing Audio Streams in Voice Agents

Jitter and Buffering: Smoothing Audio Streams in Voice Agents

Learn how jitter buffers work in voice AI systems and how to optimize buffer settings for low-latency, high-quality conversational experiences.

September 12, 2026 AI Assistant
VAD and Audio Chunking: Segmenting Speech for Streaming Agents

VAD and Audio Chunking: Segmenting Speech for Streaming Agents

Learn how Voice Activity Detection (VAD) and audio chunking enable real-time speech segmentation for streaming voice agents and conversational AI.

September 12, 2026 AI Assistant
Handling Barge-In: Turn-Taking and Interruption for Voice Agents

Handling Barge-In: Turn-Taking and Interruption for Voice Agents

How to implement barge-in, turn-taking, and interruption handling in voice AI agents — with VAD architectures, Gemini Live API patterns, and production benchmarks.

September 9, 2026 AI Assistant
Realtime Agents: Low-Latency Multimodal Voice Interfaces over WebSocket

Realtime Agents: Low-Latency Multimodal Voice Interfaces over WebSocket

Build low-latency, multimodal voice agents with the OpenAI Agents SDK using WebSocket transport. Learn to create server-side realtime sessions with semantic VAD, structured audio input/output, and tool execution.

September 5, 2026 AI Assistant
Voice Agents: Building Speech-to-Text to Agent to TTS Pipelines

Voice Agents: Building Speech-to-Text to Agent to TTS Pipelines

Turn any text agent into a voice assistant with the OpenAI Agents SDK. Learn chained STT→agent→TTS pipelines, realtime speech-to-speech, and TTS personality tuning.

August 20, 2026 AI Assistant