JobConnect

Research Engineer, Language - Wearables Polyglot AI

  • Meta
  • Redmond, United States
  • $180,000 – $250,000

Reality Labs at Meta is building products that make it easier for people to connect with the ones they love most, enjoy top-notch, wire-free VR, and push the future of computing platforms. We are a team of experts developing and shipping products at the intersection of hardware, software and content.

We are seeking a Research Engineer to join our Polyglot AI team within Reality Labs. This role will focus on developing and deploying Voice LLMs that power speech recognition, translation, and synthesis capabilities. You will work on both server-side and on-device deployments, maintain high-quality datasets, and build evaluation frameworks to drive rapid product improvements.

Responsibilities
Research and develop state-of-the-art Voice LLM models for speech recognition, translation, and synthesis
* Deploy Voice LLM systems to production environments, including both server-side infrastructure and on-device implementations
* Build and maintain high-quality datasets for training and evaluating Voice LLM systems
* Design and implement evaluation frameworks and metrics to measure model performance and drive improvements
* Conduct rigorous experimentation and ablation studies to optimize model quality, latency, and efficiency across deployment targets
* Collaborate closely with product managers, engineers, and UX designers to align technical solutions with user needs and deliver production-ready features
* Stay at the forefront of research in speech processing, Voice LLMs, and multilingual AI, bringing new methodologies into the team's development pipeline

Qualifications
Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
* 5+ years experience developing and deploying machine learning models for speech or language applications
* Experience with Voice LLMs or speech processing systems (speech recognition, translation, or synthesis)
* Demonstrated experience in deploying ML models to production
* Track record of building and maintaining datasets and evaluation pipelines for ML systems Proven ability to communicate complex technical concepts and collaborate with cross-functional teams
* Experience with multilingual speech or language models
* Experience deploying audio models to server-side and/or on-device/edge environments
* Experience with evaluation frameworks and metrics for speech/language systems
* Advanced degree (MS or PhD) in Computer Science, Machine Learning, AI, Speech Processing, or a related technical field

Skills

  • Speech Recognition
  • Large Language Models
  • Machine Translation
  • Text-to-Speech
  • PyTorch
  • Model Deployment
  • Evaluation Frameworks

Related jobs

MetaApply for this job