Baike.dev
All toolsAI codingTrendingOpen sourceNewsSubmit
Log in
Back to tool

Speech · Issues· 314 open

Open on GitHub

Locally synced open issues (discussions stay on GitHub)

  • #16286

    nemotron-labs-voicechat-11b: Training data

    community-requestUpdated Sep 18, 2026
  • #16224

    SubsamplingReductionModule pooling returns wrong lengths (factor 4 -> [6,5], factor 8 -> [0,0])

    community-requestwaiting-on-maintainersUpdated Sep 18, 2026
  • #16178

    voice_agent: NemoSTTService's default model is misspelled ('nnvidia'), silently defaulting has_turn_taking to False

    community-requestwaiting-on-maintainersUpdated Sep 18, 2026
  • #15820

    nvidia/nemotron-3.5-asr-streaming-0.6b: transcribe() fails with "ValueError: Unknown prompt key: 'None'"

    bugcommunity-requestwaiting-on-customerUpdated Sep 18, 2026
  • #16156

    Verify evals on Papers with Code

    community-requestwaiting-on-customerUpdated Sep 18, 2026
  • #15143

    Mismatch of len(words) and len(word_confidence) in parakeet-tdt-0.6b-v3 transcription

    bugASRcommunity-requestUpdated Sep 17, 2026
  • #16223

    ConvSubsampling forward broken: chunking=-1, striding_conv1d/dw_striding_conv1d, and vggnet all raise TypeError

    community-requestwaiting-on-maintainersUpdated Sep 17, 2026
  • #16278

    Sortformer splits one speaker into two when their distance to the mic changes

    community-requestUpdated Sep 17, 2026
  • #16247

    Early-interruption augmentation can drop the relocated EOS token entirely

    community-requestUpdated Sep 16, 2026
  • #16216

    FastPitchModel_SSL / SSLDisentangler crash on validation: on_validation_epoch_end() missing 1 required positional argument: 'outputs'

    community-requestUpdated Sep 16, 2026
  • #15918

    No Streaming API for streaming sortformer?

    community-requestwaiting-on-maintainersUpdated Sep 16, 2026
  • #15640

    missing punctuation marks about nvidia/nemotron-speech-streaming-en-0.6b

    bugcommunity-requestwaiting-on-customerUpdated Sep 16, 2026
  • #16252

    nemotron-labs-voicechat: Issues with the model's voice

    bugcommunity-requestwaiting-on-maintainersUpdated Sep 16, 2026
  • #15482

    Problem with computing drop_extra_pre_encoded when varying pre_encode_cache_size for SubSampling and VGG frontends

    bugcommunity-requestwaiting-on-maintainersUpdated Sep 15, 2026
  • #16249

    MagpieTTS raw-audio padding adds a spurious silent frame to exact-multiple-length audio

    community-requestwaiting-on-maintainersUpdated Sep 15, 2026