About the Role
We're seeking a Senior Speech Recognition Engineer to build and optimize production ASR systems using OpenAI's Whisper model. You'll work on fine-tuning, deployment optimization, and integration of Whisper into our voice AI platform serving millions of users.
This is a unique opportunity to work at the cutting edge of speech technology, optimizing one of the most advanced open-source ASR models for real-world production use cases.
What You'll Do
- Fine-tune Whisper models for domain-specific applications (medical, legal, financial)
- Optimize Whisper for low-latency streaming recognition
- Deploy and scale ASR systems handling 10M+ requests/day
- Reduce inference costs through quantization and model compression
- Build evaluation frameworks to measure accuracy across accents and domains
- Collaborate with product teams to integrate ASR into customer-facing features
- Research and implement improvements to Whisper architecture
Requirements
- 5+ years of experience in speech recognition or audio ML
- Strong experience with Whisper, Wav2Vec, or similar transformer-based ASR models
- Proficiency in PyTorch and production ML deployment
- Experience with model optimization techniques (quantization, distillation, pruning)
- Strong Python programming and software engineering skills
- Experience with cloud infrastructure (AWS, GCP, or Azure)
- Understanding of audio signal processing and feature extraction
Nice to Have
- PhD in Computer Science, Electrical Engineering, or related field
- Experience with Kaldi, ESPnet, or other ASR toolkits
- Published research in speech recognition or NLP
- Experience with multilingual ASR systems
- Contributions to open-source speech projects
- Experience with ONNX, TensorRT, or other inference optimization tools
Benefits
- Competitive salary ($160K-$210K) + equity
- Fully remote work with flexible hours
- Health, dental, and vision insurance
- 401(k) with company match
- Generous PTO and parental leave
- Professional development budget ($5K/year)
- Latest hardware and tools
- Conference attendance and speaking opportunities
How to Apply
Send your resume and a brief note about your experience with Whisper or other ASR systems to: careers@voiceflowai.com
Please include:
- Links to relevant projects or publications
- GitHub profile (if applicable)
- Examples of ASR systems you've built or optimized
Apply via Email