Über die Stelle
The product company is hiring a Machine Learning Engineer to develop voice synthesis, AI voice conversion, and generative audio technologies. The role covers model development, real-time audio optimization, ML pipelines, and production deployment.
About the project
An innovative product company developing generative audio and voice transformation technologies.
The company evaluates candidates based on hands-on confidence in voice synthesis and Python rather than formal Junior, Mid, or Senior labels.
Responsibilities
- Develop, train, and fine-tune machine learning models for voice synthesis and AI voice conversion
- Design, build, and maintain ML pipelines from data preparation to production
- Optimize audio processing models for streaming and real-time inference with ultra-low latency
- Turn research-grade models into reliable production web services and REST APIs
- Analyze and improve the quality, naturalness, and fidelity of synthesized audio
- Integrate ML models into the final user-facing product
Nice to have
- Experience with voice or audio models in a broader ML stack
- Background in Natural Language Processing or Automatic Speech Recognition
- Experience with FastAPI and Docker
- Familiarity with PyDub, Librosa, or similar audio libraries
- Experience with MLOps practices and tracking tools such as Weights & Biases