Research Engineer, Language - Wearables Polyglot AI
Meta
- Location
- Onsite (Redmond, Washington)
- Compensation
- $154k - $217k/yr
- Employment
- Full-time
- Level
- Senior Level
Posted 3 days ago
About the Role
Meta's Reality Labs Polyglot AI team is seeking a Research Engineer to develop and deploy Voice LLMs for speech recognition, translation, and synthesis. The role focuses on building evaluation frameworks and maintaining datasets to drive product improvements across server-side and on-device platforms.
Skills
Voice LLMs
Speech Recognition
Translation
Synthesis
Machine Learning
Python
On-device Deployment
Server-side Infrastructure
Data Engineering
Evaluation Frameworks
Multilingual AI
Ablation Studies
Latency Optimization
Efficiency Optimization
Speech Processing
Benefits
- Health Insurance
Perks
- Bonus
- Equity
Full job details
Reality Labs at Meta is building products that make it easier for people to connect with the ones they love most, enjoy top-notch, wire-free VR, and push the future of computing platforms. We are a team of experts developing and shipping products at the intersection of hardware, software and content.
We are seeking a Research Engineer to join our Polyglot AI team within Reality Labs. This role will focus on developing and deploying Voice LLMs that power speech recognition, translation, and synthesis capabilities. You will work on both server-side and on-device deployments, maintain high-quality datasets, and build evaluation frameworks to drive rapid product improvements.
$154,003/year to $217,000/year + bonus + equity + benefits
Responsibilities
- Research and develop state-of-the-art Voice LLM models for speech recognition, translation, and synthesis
- Deploy Voice LLM systems to production environments, including both server-side infrastructure and on-device implementations
- Build and maintain high-quality datasets for training and evaluating Voice LLM systems
- Design and implement evaluation frameworks and metrics to measure model performance and drive improvements
- Conduct rigorous experimentation and ablation studies to optimize model quality, latency, and efficiency across deployment targets
- Collaborate closely with product managers, engineers, and UX designers to align technical solutions with user needs and deliver production-ready features
- Stay at the forefront of research in speech processing, Voice LLMs, and multilingual AI, bringing new methodologies into the team's development pipeline
Minimum Qualifications
- Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
- 5+ years experience developing and deploying machine learning models for speech or language applications
- Experience with Voice LLMs or speech processing systems (speech recognition, translation, or synthesis)
- Demonstrated experience in deploying ML models to production
- Track record of building and maintaining datasets and evaluation pipelines for ML systems
Preferred Qualifications
- Proven ability to communicate complex technical concepts and collaborate with cross-functional teams
- Experience with multilingual speech or language models
- Experience deploying audio models to server-side and/or on-device/edge environments
- Experience with evaluation frameworks and metrics for speech/language systems
- Advanced degree (MS or PhD) in Computer Science, Machine Learning, AI, Speech Processing, or a related technical field
$154,003/year to $217,000/year + bonus + equity + benefits