Platform Engineer

Getnexus

San Francisco, Northern (CA, KY)

Hybrid

USD 150,000 - 230,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Getnexus is seeking a Platform Engineer to own the infrastructure enabling AI Agents across cloud and factory networks. You will design deployment architecture, observability, and the connectivity stack that allows on-prem and AWS deployments to run identically in security-constrained plant environments.

You will work on end-to-end systems that support production investigations, go-lives, and continuous uptime in manufacturing settings.

Qualifications

  • 2+ years building and operating production systems with uptime ownership.
  • Strong software development skills for greenfield and legacy codebases.
  • Solid fundamentals in containers, networking, and cloud infrastructure.
  • Experience using AI tools and integrating them into engineering workflows.
  • Ability to communicate clearly with teammates, partners and customers.

Responsibilities

  • Define deployment architecture for cloud, customer VPCs, and on-prem plants.
  • Architect and run observability stack across application, platform, and infrastructure.
  • Build integration substrates to connect agents to factory stacks (devices, file formats, SDKs).
  • Refine containerization and internal networking for AWS and plant DMZs.

Skills

Production uptime
Software engineering
Containers & cloud
AI tooling
Clear communication

Job description

We're looking for a Platform Engineer to own the infrastructure that puts long-horizon AI Agents inside America's factories. This includes the deployment architecture, observability stack, and industrial connectivity layer that let our agents run anywhere from our cloud to a locked-down plant network.

This role is core to how Nexus ships. Every agent investigation, every customer go-live, every on-prem deployment runs on the systems you'll build. Manufacturing doesn't forgive flaky software: our customers run around the clock and their factories collectively ship billions of dollars of product a year.

About Nexus

Nexus is building the Self-Improving OS for manufacturing. Our platform seamlessly embeds Industrial Agents into decades-old equipment and proprietary industrial control systems to monitor production lines, trace every loss to its root cause, and autonomously act.

We're growing quickly, with customers ranging from Fortune 100 manufacturers to startups building cutting-edge supply chains in the United States. Our team is comprised of manufacturing and software experts from Tesla, Distyl, and Convoy, and we're backed by A*, a16z, BoxGroup, and more.

What you'll do

You'll work across a few tightly coupled areas. We're flexible on entry points and will match scope to your strengths.

Deployment architecture. Define how Nexus ships into wildly different environments: our cloud, customer VPCs, and on-prem servers sitting inside plant networks. You'll own the architecture and processes that make "deploy to a factory" as routine as "deploy to prod." Most of our deployments go live in under two weeks - your work is what keeps that true as we scale.

Observability. Architect and run the stack that tells us what every agent is doing, across all three layers: application (agent traces, evals, logs), platform (workflow runner, backend services), and infrastructure (AWS, customer VPCs, on-prem). When something misbehaves inside a customer's plant, your telemetry is how we know before the customer does.

Industrial connectivity. Instead of Salesforce and Slack, our connectors speak Rockwell, Siemens, SAP, and UNS. You'll build the integration substrate that connects agents natively to every layer of the factory stack - from proprietary control-system file formats and SDKs on up.

Containerization and networking. Build and refine the containerization and internal networking stack that lets one platform run identically in AWS and inside a plant DMZ - secure, updatable, and observable even where internet access isn't a given.

What you bring
  • 2+ years building and operating production systems, with real ownership of uptime. You know how scalable systems should be architected, built, and maintained.

  • Exceptional software development skills, from greenfield to large existing codebases. You can implement systems in about one-third the time most competent people think possible.

  • Strong fundamentals in containers, networking, and cloud infrastructure. You can reason about how data and traffic flow through a distributed system, and what breaks first.

  • You use AI aggressively in your own engineering workflow. You've tried a bunch of AI developer tools and have strong opinions about what actually works.

  • You communicate clearly with teammates, partners, and customers, and you'll respectfully challenge (and be challenged) so that only the best ideas get executed.

Bonus
  • You've shipped agentic or LLM systems to production and think clearly about evals, reliable agent loops, tool design, and context management.

  • You've worked with industrial or OT systems (PLCs, SCADA, historians, MES, UNS, etc.) or in other secure, regulated, or air-gapped environments.

  • You've built internal tooling consumed by both humans and agents.

Who you are
  • Excited about the promise of Physical AI in manufacturing, logistics, robotics, and energy.

  • Biased toward action: you believe the fastest way to validate an idea is to build a prototype and put it in front of a customer.

  • An independent thinker with an owner's mentality. No job is too big or too small, and developing the product, debating strategy, and talking with customers all in the same day is energizing.

  • A low-ego, high-integrity operator with an innate hunger to achieve goals in a startup environment.

Interview process
  1. Intro chat: Share what you're looking for next and learn more about what we're building.

  2. Founder interview: Talk with our founder in more detail about the role.

  3. Technical interview: We'll have you complete a short exercise specific to the role.

  4. Onsite: Come to San Francisco and meet the team through a series of 1:1 interviews.

  5. Decision: We'll move fast.

This role is based in San Francisco. Occasional travel may be required for events and customer visits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer
Platform Engineer

Engg • San Francisco (CA)

On-site
USD 150,000 - 210,000
Forward Deployed Engineer
Forward Deployed Engineer

Getnexus • San Francisco (CA), Northern (KY)

Hybrid
USD 160,000 - 240,000
Technical Program Manager
Technical Program Manager

Nexus Cognitive Technologies • Atlanta (GA)

On-site
USD 100,000 - 130,000
Collaborative team culture
Challenging work environment
Leadership investment in development
Platform Engineer—AI Factory Infrastructure
Platform Engineer—AI Factory Infrastructure

Engg • San Francisco (CA)

On-site
USD 150,000 - 210,000
Delivery Manager
Delivery Manager

Nexus Cognitive Technologies • Atlanta (GA)

On-site
USD 120,000 - 150,000
Collaborative team culture
Learning and development opportunities
Forward Deployed Engagement Manager
Forward Deployed Engagement Manager

Nexus Cognitive Technologies • Atlanta (GA)

On-site
USD 150,000 - 210,000
Forward Deployed Engineer
Forward Deployed Engineer

Aionia Group • San Francisco (CA)

On-site
USD 200,000 - 325,000
Fullstack/AI Engineer
Fullstack/AI Engineer

Poka Labs • San Francisco (CA)

On-site
USD 140,000 - 210,000
AI Staff Engineer
AI Staff Engineer

Software Defined Automation • Boston (MA)

On-site
USD 120,000 - 150,000
Equity options
Flexibility in working hours
Opportunities for rapid career growth
AI Staff Engineer
AI Staff Engineer

Software Defined Automation • Boston (MA)

On-site
USD 120,000 - 150,000
Equity in a growing company
Flexible hours
Dynamic work environment