Software Engineer, Multimedia & Multimodal AI

Meta
US - California - Menlo Park
View Company Profile / << Go Back

  • Job Type: Full time
  • Yesterday

Job Description

Build multimodal data-production and evaluation pipelines, agentic and human-in-the-loop workflows, and generative or representation models for speech, sound, and music. Improve training and evaluation quality, implement and extend research methods, and mentor engineers while setting standards for the team.

Requirements: Requires a bachelor's degree or equivalent practical experience, at least six years of relevant programming experience (or three years plus a PhD), and at least three years building ML systems. Candidates should have strong Python and PyTorch skills, experience with speech, audio, or music machine learning, large-scale data pipelines and distributed training, and a record of turning research into measurable systems.

Key Skills: Python, PyTorch, Speech, Audio, And Music Machine Learning, Large-Scale Data Pipelines, Distributed Training, Multimodal AI, Generative Modeling, Audio Signal Processing, Evaluation Infrastructure, Human-In-The-Loop Systems, Agentic Workflows, Research Implementation, GPU Optimization, Audio Codecs, Responsible AI, Mentoring

Benefits: Bonus, Equity, Benefits




Fast Track Upload