Get more replies from employers
Send a job-specific resume in minutes.
Jobzhr is seeking an AI Engineer to help build a vision-language-action foundation model that operates on real robots. The role focuses on turning perception, reasoning, and motor control into robust, real-time decisions on embedded edge hardware.
You will join a small founding team drawn from robotics, autonomous systems, and foundation-model research, moving concepts from prototype to production-grade deployments on physical platforms.
AI Engineer - Vision-Language-Action Foundation Models
Location: On-site, Bay Area (CA)
A well-funded, early-stage physical AI company is building a vision-language-action (VLA) foundation model that turns real-world robots into intelligent, autonomous agents. Think of it as a large multimodal model that outputs motor control instead of text - perceiving the environment, reasoning about it, and acting on it in real time.
The model runs onboard, at the edge, in real time, even in GPS- and comms-denied environments. Prototypes across ground and air platforms are already operational in the field today. You'd be joining a small, hand-picked founding team drawn from leading robotics, autonomous-vehicle, and foundation-model backgrounds - solving genuinely unsolved problems in embodied intelligence, and shipping them onto real hardware rather than benchmarks.
This is applied AI, not pure research. If you want your models controlling physical machines in the real world, this is that role.