Tegain voice agent — local test

Runs the worker you started locally. The agent is dispatched into the room when you click Start.

Google (Gemini Live)

Native speech — no STT/TTS plugins. The original worker's architecture.

Pipeline (STT → LLM → TTS)

ElevenLabs Scribe → Gemini → ElevenLabs TTS. Ours.

Scribe v2 Realtime is the only realtime streaming model.
cerebras/… → Cerebras; sarvam/… → Sarvam; other slash ids (google/…, meta-llama/…) → OpenRouter; plain ids → Gemini API directly.
v3 streams over HTTP (websocket rejects it) — ~1.5s first audio per sentence, best voice quality.
21m00Tcm4TlvDq8ikWAM = Rachel
Start the LLM while you're still talking (on interim) — saves ~0.5-1s per turn, risks answering before you finish.
idle
Room: Agent: Elapsed: 0.0s Mic: off last user → agent reply: