
Senior Applied ML Engineer (Speech & Audio)
Remote
Full Time
#Technology
#Artificial Intelligence
#Machine Learning
#Speech Recognition
#Text
#Audio
#Optimization
#Python
#Deep Learning
#Data Pipeline
We are building advanced voice technologies that enable natural and accurate Arabic speech interactions at scale. As a Senior Applied Machine Learning Engineer focused on speech and audio, you will take ownership of the full lifecycle of production-grade models, from constructing specialized data pipelines to deploying optimized inference systems that power real-time conversational experiences.
Key outcomes
- Design, fine-tune, and optimize machine learning models for Arabic text-to-speech and automatic speech recognition applications.
- Benchmark and evaluate TTS and ASR models using Arabic-specific test sets, tracking metrics such as Word Error Rate, naturalness, and dialect coverage.
- Fine-tune generative models to support voice cloning, zero-shot speaker adaptation, and high-quality speech synthesis.
- Build and maintain Arabic-focused data pipelines covering audio collection, preprocessing, diacritization, cleaning, and augmentation.
- Optimize model inference for production using quantization, KV-cache tuning, and streaming techniques to achieve low-latency performance.
- Integrate and evaluate complete speech-to-speech conversational pipelines that deliver reliable user experiences.
- Translate recent research findings into production-ready solutions through structured experimentation and implementation.
- Partner with engineering and product teams to ensure robust, scalable deployment of speech systems.
Requirements
- Senior-level experience designing and deploying machine learning models for speech or audio applications.
- Strong proficiency in Python and deep learning frameworks for model development and optimization.
- Hands-on expertise with data pipeline construction, including audio preprocessing, cleaning, and augmentation.
- Practical experience applying inference optimization methods such as quantization and streaming techniques.
- Fluency in English for technical collaboration and documentation.
Compensation
We offer competitive compensation along with the flexibility of remote work.
How to apply
We invite qualified candidates to submit their application. We look forward to reviewing your background and discussing how your expertise can contribute to our speech technology initiatives.










