We‘re HIRING!
Author: babysorCreated Jan 7, 2026Updated Mar 2, 2026
If you’re excited about building high-quality speech and multimodal models—and want your work to be used by real creators—we’d love to talk.
Algorithm Engineer (Speech / Multimodal)
Requirements
CS or related background (BS+)
Strong Python and solid engineering practices
Experience with PyTorch (DeepSpeed a plus)
Hands-on experience in one or more of:
- TTS / ASR / Voice Conversion
- Speaker Recognition / Diarization
- Speech Enhancement
- Audio–visual or multimodal modeling
⚙️ Software Engineer (AI / Data)
Requirements
- CS or related background (BS+)
- Strong Python or backend skills
- Experience building AI/data pipelines or model-serving systems
- Familiarity with PyTorch and ML workflows is a plus
If you’re passionate about audio, multimodal AI, and building things that people actually hear and use, feel free to reach out on [email protected] .
Source: babysor/MockingBird