An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Nexus is hiring a Platform Engineer to own the infrastructure enabling AI Agents in US factories, from cloud to plant networks. You’ll craft deployment architecture, observability, and industrial connectivity that keeps on-prem and edge deployments reliable and scalable.
You will own containerization, networking, and secure connectivity across AWS, VPCs, and plant environments, ensuring robust uptime and predictable deployments for customers shipping billions in product every year.
We’re looking for a Platform Engineer to own the infrastructure that puts long-horizon AI Agents inside America’s factories. This includes the deployment architecture, observability stack, and industrial connectivity layer that let our agents run anywhere from our cloud to a locked-down plant network. This role is core to how Nexus ships. Every agent investigation, every customer go-live, every on-prem deployment runs on the systems you’ll build. Manufacturing doesn’t forgive flaky software: our customers run around the clock and their factories collectively ship billions of dollars of product a year.
Nexus is building the Self-Improving OS for manufacturing. Our platform seamlessly embeds Industrial Agents into decades-old equipment and proprietary industrial control systems to monitor production lines, trace every loss to its root cause, and autonomously act. We’re growing quickly, with customers ranging from Fortune 100 manufacturers to startups building cutting-edge supply chains in the United States. Our team is comprised of manufacturing and software experts from Tesla, Distyl, and Convoy, and we’re backed by A*, a16z, BoxGroup, and more.
You’ll work across a few tightly coupled areas. We’re flexible on entry points and will match scope to your strengths.
Deployment architecture. Define how Nexus ships into wildly different environments: our cloud, customer VPCs, and on-prem servers sitting inside plant networks. You’ll own the architecture and processes that make "deploy to a factory" as routine as "deploy to prod." Most of our deployments go live in under two weeks - your work is what keeps that true as we scale.
Observability. Architect and run the stack that tells us what every agent is doing, across all three layers: application (agent traces, evals, logs), platform (workflow runner, backend services), and infrastructure (AWS, customer VPCs, on-prem). When something misbehaves inside a customer’s plant, your telemetry is how we know before the customer does.
Industrial connectivity. Instead of Salesforce and Slack, our connectors speak Rockwell, Siemens, SAP, and UNS. You’ll build the integration substrate that connects agents natively to every layer of the factory stack - from proprietary control-system file formats and SDKs on up.
Containerization and networking. Build and refine the containerization and internal networking stack that lets one platform run identically in AWS and inside a plant DMZ - secure, updatable, and observable even where internet access isn’t a given.
This role is based in San Francisco. Occasional travel may be required for events and customer visits.