speech-to-speech · Issues· 114 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #565
Listener backchannels abort active response mid-playback
Updated Sep 16, 2026 - #563
STT can hallucinate confident, mislabeled transcriptions with no confidence signal, leading to nonsensical input reaching the LLM
Updated Sep 16, 2026 - #519
Preserve Responses API reasoning items during tool continuations
enhancementUpdated Sep 16, 2026 - #559
No user-facing signal while a response is generating on slower local TTS/LLM backends
Updated Sep 13, 2026 - #558
Sending silence on RTC causes noticable per-client background CPU usage
Updated Sep 13, 2026 - #557
Browser realtime demo can underrun TTS playback without a startup jitter buffer
Updated Sep 10, 2026 - #292
Add FunASR/SenseVoice as a self-hosted STT option
Updated Sep 9, 2026 - #555
STT handlers retain the detected language across sessions; only Parakeet resets it
Updated Sep 8, 2026 - #479
Speculative execution during user speech?
Updated Sep 8, 2026 - #547
Support optional mid-session model switching through session.update
for-hf-staff-onlyUpdated Sep 5, 2026 - #235
Feature Request: Great new VAD and ASR models
Updated Sep 5, 2026 - #540
llama-swap support
Updated Sep 5, 2026 - #546
Maintainer on holidays
Updated Sep 4, 2026 - #433
Reduce barge-in latency with reversible pre-confirmation audio ducking
Updated Sep 2, 2026 - #536
Support stateful streaming STT sessions instead of repeated whole-utterance uploads
Updated Aug 28, 2026