At VoiceFrom we're building real-time speech AI that breaks communication barriers across languages, proficiency levels, and context. Think: a single platform that transforms spoken content to fit the listener's needs, instantly.
We're looking for an ML Researcher who's energized by frontier problems in speech/audio-and wants their work to land in a real-time system where every millisecond and nuance counts.
What you'd work on
- Real-time speech transformation that adapts across linguistic, cultural, and domain context
- Low-latency streaming inference with tone, prosody, and speaker identity in multilingual settings
- High-ownership research in a tight feedback loop: rapid experiments, real user signals, fast iteration
You'd be a great fit if
- You have a strong publication or research track record in speech/audio ML
- You hold a PhD in a relevant field (or have equivalent research experience)
- You're comfortable going from paper to prototype to production
- You thrive in small teams where your work ships fast and is felt immediately
What we offer: Competitive pay · Remote or hybrid · Company laptop · Potential equity
If this sounds like your kind of challenge, DM me with a few lines on what excites you + a link to your Google Scholar / GitHub / LinkedIn.