We‘re HIRING!

Author: babysorCreated Jan 7, 2026Updated Mar 2, 2026

If you’re excited about building high-quality speech and multimodal models—and want your work to be used by real creators—we’d love to talk.

Algorithm Engineer (Speech / Multimodal)

Requirements

  • CS or related background (BS+)

  • Strong Python and solid engineering practices

  • Experience with PyTorch (DeepSpeed a plus)

  • Hands-on experience in one or more of:

    • TTS / ASR / Voice Conversion
    • Speaker Recognition / Diarization
    • Speech Enhancement
    • Audio–visual or multimodal modeling

⚙️ Software Engineer (AI / Data)

Requirements

  • CS or related background (BS+)
  • Strong Python or backend skills
  • Experience building AI/data pipelines or model-serving systems
  • Familiarity with PyTorch and ML workflows is a plus

If you’re passionate about audio, multimodal AI, and building things that people actually hear and use, feel free to reach out on [email protected] .