Platform Engineer

Engg

San Francisco (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Nexus is hiring a Platform Engineer to own the infrastructure enabling AI Agents in US factories, from cloud to plant networks. You’ll craft deployment architecture, observability, and industrial connectivity that keeps on-prem and edge deployments reliable and scalable.

You will own containerization, networking, and secure connectivity across AWS, VPCs, and plant environments, ensuring robust uptime and predictable deployments for customers shipping billions in product every year.

Qualifications

  • 2+ years building and operating production systems with uptime ownership.
  • Exceptional software development skills, from greenfield to large codebases.
  • Strong fundamentals in containers, networking, and cloud infrastructure.

Responsibilities

  • Define deployment architecture for cloud, customer VPCs, and on-prem environments to enable factory deployments.
  • Architect and run the observability stack across application, platform, and infrastructure layers.
  • Build the integration substrate that connects agents to factory stacks (Rockwell, Siemens, SAP, UNS).
  • Develop containerization and internal networking so the platform runs identically in cloud and plant DMZ.

Skills

Production systems
Software development
Containers
Networking
Cloud infrastructure
AI tooling
Communication

Tools

AWS
Kubernetes
Docker

Job description

We’re looking for a Platform Engineer to own the infrastructure that puts long-horizon AI Agents inside America’s factories. This includes the deployment architecture, observability stack, and industrial connectivity layer that let our agents run anywhere from our cloud to a locked-down plant network. This role is core to how Nexus ships. Every agent investigation, every customer go-live, every on-prem deployment runs on the systems you’ll build. Manufacturing doesn’t forgive flaky software: our customers run around the clock and their factories collectively ship billions of dollars of product a year.

ABOUT NEXUS

Nexus is building the Self-Improving OS for manufacturing. Our platform seamlessly embeds Industrial Agents into decades-old equipment and proprietary industrial control systems to monitor production lines, trace every loss to its root cause, and autonomously act. We’re growing quickly, with customers ranging from Fortune 100 manufacturers to startups building cutting-edge supply chains in the United States. Our team is comprised of manufacturing and software experts from Tesla, Distyl, and Convoy, and we’re backed by A*, a16z, BoxGroup, and more.

WHAT YOU'LL DO

You’ll work across a few tightly coupled areas. We’re flexible on entry points and will match scope to your strengths.

Deployment architecture. Define how Nexus ships into wildly different environments: our cloud, customer VPCs, and on-prem servers sitting inside plant networks. You’ll own the architecture and processes that make "deploy to a factory" as routine as "deploy to prod." Most of our deployments go live in under two weeks - your work is what keeps that true as we scale.

Observability. Architect and run the stack that tells us what every agent is doing, across all three layers: application (agent traces, evals, logs), platform (workflow runner, backend services), and infrastructure (AWS, customer VPCs, on-prem). When something misbehaves inside a customer’s plant, your telemetry is how we know before the customer does.

Industrial connectivity. Instead of Salesforce and Slack, our connectors speak Rockwell, Siemens, SAP, and UNS. You’ll build the integration substrate that connects agents natively to every layer of the factory stack - from proprietary control-system file formats and SDKs on up.

Containerization and networking. Build and refine the containerization and internal networking stack that lets one platform run identically in AWS and inside a plant DMZ - secure, updatable, and observable even where internet access isn’t a given.

WHAT YOU BRING
  • 2+ years building and operating production systems, with real ownership of uptime. You know how scalable systems should be architected, built, and maintained.
  • Exceptional software development skills, from greenfield to large existing codebases. You can implement systems in about one-third the time most competent people think possible.
  • Strong fundamentals in containers, networking, and cloud infrastructure. You can reason about how data and traffic flow through a distributed system, and what breaks first.
  • You use AI aggressively in your own engineering workflow. You've tried a bunch of AI developer tools and have strong opinions about what actually works.
  • You communicate clearly with teammates, partners, and customers, and you'll respectfully challenge (and be challenged) so that only the best ideas get executed.
  • BONUS - You've shipped agentic or LLM systems to production and think clearly about evals, reliable agent loops, tool design, and context management.
  • You've worked with industrial or OT systems (PLCs, SCADA, historians, MES, UNS, etc.) or in other secure, regulated, or air-gapped environments.
  • You've built internal tooling consumed by both humans and agents.
WHO YOU ARE
  • Excited about the promise of Physical AI in manufacturing, logistics, robotics, and energy.
  • Biased toward action: you believe the fastest way to validate an idea is to build a prototype and put it in front of a customer.
  • An independent thinker with an owner's mentality. No job is too big or too small, and developing the product, debating strategy, and talking with customers all in the same day is energizing.
  • A low-ego, high-integrity operator with an innate hunger to achieve goals in a startup environment.
INTERVIEW PROCESS
  1. Application review:
  2. Intro chat: Share what you're looking for next and learn more about what we're building.
  3. Founder interview: Talk with our founder in more detail about the role.
  4. Technical interview: We'll have you complete a short exercise specific to the role.
  5. Onsite: Come to San Francisco and meet the team through a series of 1:1 interviews.
  6. Decision: We'll move fast.

This role is based in San Francisco. Occasional travel may be required for events and customer visits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Forward Deployed Engineer
Forward Deployed Engineer

Aionia Group • San Francisco (CA)

On-site
USD 200,000 - 325,000
Delivery Manager
Delivery Manager

Nexus One • Atlanta (GA)

On-site
USD 150,000 - 210,000
Technical Program Manager
Technical Program Manager

Nexus Cognitive Technologies • Atlanta (GA)

On-site
USD 100,000 - 130,000
Collaborative team culture
Challenging work environment
Leadership investment in development
Principal Solutions Architect
Principal Solutions Architect

Nexus One • Atlanta (GA)

On-site
USD 180,000 - 240,000
Delivery Manager
Delivery Manager

Nexus Cognitive Technologies • Atlanta (GA)

On-site
USD 120,000 - 150,000
Collaborative team culture
Learning and development opportunities
Platform Engineer—AI Factory Infrastructure
Platform Engineer—AI Factory Infrastructure

Engg • San Francisco (CA)

On-site
USD 150,000 - 210,000
Forward Deployed Engagement Manager
Forward Deployed Engagement Manager

Nexus One • Atlanta (GA)

On-site
USD 140,000 - 200,000
Platform Engineer
Platform Engineer

Harper • San Francisco (CA)

On-site
USD 140,000 - 280,000
Uber commuter benefits
Meals provided (breakfast, lunch, and/
Snacks, drinks and coffee daily
+2
Platform Engineer
Platform Engineer

Antimetal Inc • New York (NY)

On-site
USD 140,000 - 210,000
Competitive salary & equity
Forward Deployed Engagement Manager
Forward Deployed Engagement Manager

Nexus Cognitive Technologies • Atlanta (GA)

On-site
USD 150,000 - 210,000