Machine Learning Engineer, Multimodal Perception and Authentication

Triwill Group

San Francisco, Northern (CA, KY)

Hybrid

USD 150,000 - 200,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Relocation assistance

Job summary

OpenAI in San Francisco is seeking a machine learning engineer to shape how future AI systems understand the physical world and people within it. You will work on multimodal perception and authentication, integrating signals from cameras, microphones, and other sensors.

You will collaborate with hardware, firmware, software, and product teams to bring research into real-world systems, with relocation assistance offered and a hybrid work model (office three days per week in SF).

Qualifications

  • Background in computer vision, audio or speech ML, multimodal learning, or sensing.
  • Experience developing specialized ML models or larger multimodal models.
  • Experience bringing research ideas into practical prototypes or products.
  • Proficient in Python and PyTorch; comfortable with C++ or systems integration.
  • Experience with authentication, biometrics, or privacy-sensitive applications.
  • Ability to design experiments and evaluate model behavior.

Responsibilities

  • Research and develop multimodal perception and authentication methods across visual, audio, and other sensing signals.
  • Explore integration of specialized perception models and larger multimodal models.
  • Design data, training, and evaluation approaches that work in real-world conditions.
  • Study model behavior, robustness, and failure modes across sensing and deployment environments.
  • Integrate and validate new capabilities in real-time or resource-constrained systems.
  • Collaborate with hardware, firmware, software, and product teams to ship research into working systems.

Skills

Python
PyTorch
C++
Computer vision
Multimodal learning
Speech/Audio ML
Systems integration
Authentication/ biometrics

Job description

Description: About the Team

The Future of Computing Research team is an applied research team within OpenAI’s Consumer Devices group. We study how AI systems perceive people and their surroundings, and we turn that research into capabilities for future products.

Our work spans machine learning, sensing, and hardware, with a focus on building systems that work beyond controlled environments.

About the Role

We’re looking for a machine learning engineer to help shape how future AI systems understand the physical world and the people in it. The role focuses on multimodal perception and authentication, bringing together signals from cameras, microphones, and other sensors.

You’ll work with specialized perception models and larger multimodal models, and partner with hardware, firmware, software, and product teams to bring new research into real-world systems.

This role is based in San Francisco. We work in the office three days per week and offer relocation assistance.

In this role, you will:

  • Research and develop multimodal perception and authentication methods across visual, audio, and other sensing signals.
  • Explore how specialized perception models and larger multimodal models can work together.
  • Design data, training, and evaluation approaches that improve performance in real-world conditions.
  • Study model behavior, robustness, and failure modes across sensing, data, and deployment environments.
  • Integrate and validate new capabilities in real-time or resource-constrained systems.
  • Work with hardware, firmware, software, and product teams to turn research into working systems.

You might thrive in this role if you:

  • Have a strong background in computer vision, audio or speech machine learning, multimodal learning, or sensing.
  • Have experience developing specialized machine learning models, larger multimodal models, or both.
  • Have brought research ideas into practical systems, prototypes, or products.
  • Know how to design experiments, build evaluations, and investigate model behavior.
  • Have worked with sensing hardware, real-time systems, or other deployment constraints.
  • Are proficient in Python and PyTorch and comfortable with C++ or systems integration.
  • Have experience with authentication, biometrics, or other privacy-sensitive applications.
  • Enjoy working across disciplines on research problems that are still taking shape.
About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affi ...

Background checks ...

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Multimodal Perception & Authentication ML Engineer
Multimodal Perception & Authentication ML Engineer

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 200,000
Relocation assistance
Researcher, Multimodal Safety
Researcher, Multimodal Safety

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Relocation assistance
Hybrid work model
Researcher, Multimodal Safety
Researcher, Multimodal Safety

OpenAI • San Francisco (CA)

Hybrid
USD 150,000 - 190,000
Hybrid work model
Relocation assistance
Software Engineer, API Multimodal
Software Engineer, API Multimodal

Ritual Ads ® • San Francisco (CA)

On-site
USD 180,000 - 260,000
Research Scientist - Multimodal Agent, Consumer Devices
Research Scientist - Multimodal Agent, Consumer Devices

OpenAI • United States

Hybrid
USD 140,000 - 230,000
Relocation assistance
Hybrid work model (3 days in office)
Software Engineer, API Multimodal
Software Engineer, API Multimodal

OpenAI • San Francisco (CA)

On-site
USD 200,000 - 270,000
Research Vision Expertise
Research Vision Expertise

Thinking Machines Lab • San Francisco (CA)

On-site
USD 350,000 - 475,000
Generous health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Research, Vision Expertise
Research, Vision Expertise

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Generous health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Research Scientist - Multimodal Agent, Consumer Devices
Research Scientist - Multimodal Agent, Consumer Devices

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 445,000
Hybrid work model
Relocation assistance
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

SupportFinity™ • San Francisco (CA)

On-site
USD 100,000 - 180,000