Software Engineer, Multimedia & Multimodal AI

Meta Careers

Bellevue (WA)

On-site

USD 154,000 - 217,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

Meta is seeking a Software Engineer for the Multimedia & Multimodal AI team in Bellevue, WA. You will own data and evaluation pipelines spanning image, video, audio, and speech, translating research needs into scalable engineering solutions and driving quality through robust task creation and evaluation tooling.

The role sits near research and emphasizes setting standards, collaboration with cross-functional partners, and building expert-in-the-loop workflows to scale data production and model

Qualifications

  • Bachelor's degree in Computer Science, Computer Engineering, or related field or equivalent practical experience.
  • 8+ years of software engineering experience, or PhD +5 years with media modality depth.

Responsibilities

  • Design, create, and quality-review expert multimedia tasks and reference outputs for model training and evaluation.
  • Define task guidelines, rubrics, and quality criteria; calibrate reviewers to apply them consistently.
  • Build and harden evaluations; diagnose failures and improve data pipelines.
  • Mentor engineers, participate in hiring, and raise the bar on evaluation practice.

Skills

Software engineering
Multimedia pipelines
Model evaluation
Quality review

Education

Bachelor's degree in CS or related

Tools

Premiere
After Effects
Blender
Pro Tools

Job description

Applied AI (AAI) is Meta’s organization focused on making our AI models best-in-class, starting with coding. Within AAI, the MultiMedia & MultiModality team covers the multimedia domain across every modality, on both the input and the output side of a model: image, video, audio, speech and music. We work directly with research, model-training and engineering partners across MSL, TBD and FAIR. Current problems include evaluating video experiences, diagnosing multimedia model behavior, producing domain-expert agent tasks, and building the data and measurement pipelines multimodal capabilities are trained and judged against.About the roleYou will take a modality or a capability area, decide what data is worth producing and how it should be measured, and carry it from an open question through to a pipeline that runs and a measurement the org relies on.This is a multimodal role, not a text-only role. You will work across image, video, audio and speech, as model inputs and as model outputs, and the data and evaluations you own will cover media, not text alone.The role sits close to research. You will translate what researchers need into data and evaluation the team can produce at scale, and bring their findings back into what we build next.You will join a newly formed team, so setting direction, standards and review practice are a critical part of the job.Software Engineer, Multimedia & Multimodal AI Responsibilities:Design, create, and quality-review expert multimedia tasks and reference outputs for model training and evaluation.Define task guidelines, rubrics, and quality criteriacalibrate reviewers to apply them consistently.Build and harden evaluations.Make graders reliable, separate model failure from instrumentation failure, and reproduce and debug quality issues to resolution.Analyze failure modes in model outputs and propose new task types or data to close gaps.Design and build agentic workflows and pipelines, including human-in-the-loop and expert-in-the-loop designs, to automate data production and scale output past what manual authoring supports.Mentor engineers on the team, contribute to hiring and onboarding, and raise the bar on evaluation and quality practice.Minimum Qualifications:Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience8+ years of software engineering experience, or a PhD plus 5 years, including significant depth in one or more media modalities (video, image, audio, speech, or music)Demonstrated experience designing and building multimedia pipelines, workflows or toolingWorking understanding of how models are trained and evaluated, and of how data quality and coverage shape model behaviorExperience owning software components or systems end to end, and driving work with cross-functional partnersPreferred Qualifications:Hands-on experience evaluating or red-teaming multimodal models, or creating the data used to improve themExperience designing benchmarks or evaluations for model capability, with attention to grading reliability, reproducibility and label qualityWorking knowledge of common video, audio and streaming standards (e.g. H.265, MPEG-DASH, WebRTC)Fluency with professional media tooling (e.g. Premiere, After Effects, Blender, Pro Tools). We care about this because authoring tasks a model cannot solve requires knowing what expert media work actually looks likeExperience building data pipelines for image, video, audio, speech or complex media formats, including versioning, lineage and provenanceExperience designing AI agents, orchestration, or human-in-the-loop systemsExperience working directly with researchers and translating research needs into engineering and evaluation workUnderstanding of Responsible AI practices and building quality controls into AI outputExperience with zero-to-one work: forming a charter and standing up process while priorities are still movingDemonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologiesAbout Meta:Meta builds technologies that help people connect, find communities, and grow businesses. When Facebook launched in 2004, it changed the way people connect. Apps like Messenger, Instagram and WhatsApp further empowered billions around the world. Now, Meta is moving beyond 2D screens toward immersive experiences like augmented and virtual reality to help build the next evolution in social technology. People who choose to build their careers by building with us at Meta help shape a future that will take us beyond what digital connection makes possible today—beyond the constraints of screens, the limits of distance, and even the rules of physics.Meta is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Meta participates in the E-Verify program in certain locations, as required by law. Please note that Meta may leverage artificial intelligence and machine learning technologies in connection with applications for employment.Meta is committed to providing reasonable accommodations for candidates with disabilities in our recruiting process. If you need any assistance or accommodations due to a disability, please let us know at accommodations-ext@meta.com.$154,003/year to $217,000/year + bonus + equity + benefitsIndividual compensation is determined by skills, qualifications, experience, and location. Compensation details listed in this posting reflect the base hourly rate, monthly rate, or annual salary only, and do not include bonus, equity or sales incentives, if applicable. In addition to base compensation, Meta offers benefits. Learn more about benefits at Meta.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer, Multimedia & Multimodal AI
Software Engineer, Multimedia & Multimodal AI

Meta Careers • New York (NY)

On-site
USD 184,000 - 257,000
Software Engineer, Multimedia & Multimodal AI
Software Engineer, Multimedia & Multimodal AI

Meta Careers • Menlo Park (CA)

On-site
USD 184,000 - 257,000
Software Engineer, Multimedia & Multimodal AI
Software Engineer, Multimedia & Multimodal AI

Meta • Bellevue (WA)

On-site
USD 184,000 - 257,000
Health insurance
Equity compensation
Research Scientist, Contextual AI and Multimodal Agents
Research Scientist, Contextual AI and Multimodal Agents

Meta • Redmond (WA)

On-site
USD 219,000 - 301,000
Software Engineer, Audio SWE
Software Engineer, Audio SWE

Meta • Burlingame (CA)

On-site
USD 154,000 - 217,000
Bonus
Equity
Benefits
Lead, Product Content Engineering
Lead, Product Content Engineering

Meta Careers • Menlo Park (CA)

On-site
USD 193,000 - 269,000
Software Engineer, Machine Learning
Software Engineer, Machine Learning

Meta • Menlo Park (CA)

On-site
USD 347,000 - 403,000
Research Engineer, Safety Evaluation
Research Engineer, Safety Evaluation

Meta Careers • Menlo Park (CA)

On-site
USD 219,000 - 301,000
Software Engineer, Machine Learning
Software Engineer, Machine Learning

Meta • Burlingame (CA)

On-site
USD 184,000 - 257,000
Bonus
Equity
Benefits
Software Engineer, Systems ML (Technical Leadership)
Software Engineer, Systems ML (Technical Leadership)

Meta • New York (NY)

On-site
USD 219,000 - 301,000
Bonus
Equity
Benefits