Speech / Audio AI Engineer

Main mission

Builds speech recognition, synthesis, and audio processing.

5 key responsibilities

  • Develop multilingual speech recognition (ASR) and synthesis (TTS).
  • Build audio pipelines: diarization, denoising, detection.
  • Adapt models to accents, local languages, and domain-specific vocabularies.
  • Optimize latency for real-time uses (assistants, telephony).
  • Evaluate perceived quality and iterate with real users.

Key skills

ASR/TTS (Whisper, ElevenLabs), signal processing, PyTorch, audio streaming.

What's expected

A voice understood and a natural voice in all target languages: audio is the next major interface.

Career paths

Voice AI Architect, Multimodal Engineer, Voice Product Manager.

Openings right now

No opening for this role at the moment.

Get alerted as soon as a Speech / Audio AI Engineer position is published.

Create a job alert →