Research Engineer, Language - Wearables Polyglot AI

MetaRedmond, WashingtonOn-sitePart-timeStaff, 8–12 yearsListed 1 hour ago

Apply now

About this role

Reality Labs at Meta is building products that make it easier for people to connect with the ones they love most, enjoy top-notch, wire-free VR, and push the future of computing platforms. We are a team of experts developing and shipping products at the intersection of hardware, software and content.

We are seeking a Research Engineer to join our Polyglot AI team within Reality Labs. This role will focus on developing and deploying Voice LLMs that power speech recognition, translation, and synthesis capabilities. You will work on both server-side and on-device deployments, maintain high-quality datasets, and build evaluation frameworks to drive rapid product improvements.

Responsibilities

Research and develop state-of-the-art Voice LLM models for speech recognition, translation, and synthesis
Deploy Voice LLM systems to production environments, including both server-side infrastructure and on-device implementations
Build and maintain high-quality datasets for training and evaluating Voice LLM systems
Design and implement evaluation frameworks and metrics to measure model performance and drive improvements
Conduct rigorous experimentation and ablation studies to optimize model quality, latency, and efficiency across deployment targets
Collaborate closely with product managers, engineers, and UX designers to align technical solutions with user needs and deliver production-ready features
Stay at the forefront of research in speech processing, Voice LLMs, and multilingual AI, bringing new methodologies into the team's development pipeline

Qualifications

Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
5+ years experience developing and deploying machine learning models for speech or language applications
Experience with Voice LLMs or speech processing systems (speech recognition, translation, or synthesis)
Demonstrated experience in deploying ML models to production
Track record of building and maintaining datasets and evaluation pipelines for ML systems Proven ability to communicate complex technical concepts and collaborate with cross-functional teams
Experience with multilingual speech or language models
Experience deploying audio models to server-side and/or on-device/edge environments
Experience with evaluation frameworks and metrics for speech/language systems
Advanced degree (MS or PhD) in Computer Science, Machine Learning, AI, Speech Processing, or a related technical field