Voice AI Engineer (all genders)
On-siteBremen, City state Bremen, Germany
Job Summary
Build and improve our own TTS and STT systems from model to streaming inference, optimizing for low latency in Time-to-First-Audio and efficient serving. Enhance voice quality and naturalness through prosody, stress, and robustness under real audio conditions while evaluating new open-source speech models. Measure performance via latency profiling, quality benchmarks, and user listening tests. Work with Python and agile methods to deliver real-time voice agents that understand noisy environments. Join a team driving digital transformation of communication with short decision paths and access to top-tier hardware.
Required Qualifications
- praktische Erfahrung mit Sprachtechnologie: TTS, STT, Echtzeit-Audio oder Speech-Modell-Inferenz
- sehr gute Programmierkenntnisse in Python
- Ideen, wie man Latenz optimiert
Desired Qualifications
- Erfahrung mit dem kommerziellen Einsatz von Speech-Modellen
- Arbeiten gerne im Team und mit agilen Methoden
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.