Senior Platform Engineer, Cloud (Xora Portfolio Company)

Xora Innovation

San Diego (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Elemynt, built by Xora Innovation, is hiring a senior platform engineer to own cloud infrastructure foundations, core services, and observability. The role focuses on reproducible, secure deployments across customer environments and emphasizes reliability from the ground up.

You will design and ship containerized building blocks, deploy pipelines, and monitoring to keep services fast and available in multi-cloud or on-prem setups. This is a hands-on senior role in a fast-moving startup.

Qualifications

  • Bachelor’s or Master’s degree in CS or related engineering field, with 6+ years shipping prod software.
  • Hands-on cloud infra with Kubernetes and infrastructure-as-code (Terraform or similar) on AWS/GCP/Azure.
  • Experience designing and operating core backend services and APIs used by other engineers.
  • Ownership of cloud deployment processes: pipelines, versioned artifacts, and reproducible rollouts.
  • Observability in production: metrics, logs, traces, dashboards, alerts.
  • Strong software fundamentals; coding in Go, Python, or Rust.
  • Experience defining SLOs and reliability-first design.
  • Comfort in early-stage, ambiguous environments with sensible trade-offs.

Responsibilities

  • Build and own the cloud infrastructure foundations (networking, IAM, Kubernetes, IaC).
  • Design and build core platform services and internal APIs to run reliably anywhere.
  • Own end-to-end deployment pipeline: build, versioned artifacts, staged rollouts, rollbacks.
  • Stand up observability: metrics, logs, traces, dashboards, alerts.
  • Instrument SLOs and health signals to measure reliability and detect regressions.
  • Keep foundations portable and reproducible across environments.
  • Deliver deployment-ready building blocks (container images, Helm charts, IaC modules).
  • Harden foundations: secrets, certificates, and network boundaries.

Skills

Cloud infrastructure
Platform engineering
Backend services design
Observability
Go
Python
Rust
SRE fundamentals
Reliability engineering
SLA/SLO

Education

Bachelor’s or Master’s degree in Computer Science or related engineering

Tools

Kubernetes
Terraform
AWS
GCP
Azure
Prometheus
Grafana
OpenTelemetry

Job description

ABOUT ELEMYNT

ELEMYNT is an early-stage startup built by Xora Innovation. We develop applied intelligence that brings AI into the real world. Our platform combines advanced machine learning, high-performance simulation, and modern software engineering to accelerate the design, validation, and deployment of new materials. Our work sits at the intersection of AI, physics, and large-scale computation. The problems are hard, the stakes are high, and the impact is tangible.

ABOUT THE ROLE

This role builds what the rest of Elemynt’s engineering runs on: the cloud infrastructure, the core platform services, and the observability that keeps the product reliable. The platform runs wherever each customer chooses, whether that’s their cloud, their own compute, or a hybrid, and usually inside environments they operate themselves. That’s the hard part. The foundations you build have to be reproducible, measurable, and carry their own security wherever they land.

This is a deeply hands-on, build-focused senior role. In this role, you’ll design and own the cloud substrate, the services at the core of the product, and the metrics, logs, and traces that keep all of it visible. Those building blocks are what gets packaged and deployed wherever the platform needs to run, so the care you put in shows up in every install. How fast and how safely the whole company can ship depends on how solid these foundations are.

WHAT YOU WILL DO

  • Build and own the cloud infrastructure foundations (networking, identity and access, Kubernetes, and infrastructure-as-code) that every service and workload runs on.
  • Design and build the core platform services and internal APIs the product is made of, meant to run reliably wherever it’s deployed.
  • Own the cloud deployment process end to end: the pipeline that turns platform services into versioned, reproducible artifacts and rolls them out with staged releases and clean rollbacks.
  • Stand up the observability layer: metrics, logs, traces, dashboards, and alerting that make a failing service quick to find and diagnose.
  • Instrument service-level objectives and health signals so reliability is measurable and regressions show up before they reach a customer.
  • Keep the foundations portable and reproducible, so the platform stands up the same way across every environment it runs in.
  • Produce the deployment-ready building blocks (container images, Helm charts, infrastructure-as-code modules) that make installs and upgrades clean and repeatable wherever the platform runs.
  • Harden the platform’s foundations: secrets, certificate handling, and network boundaries that protect the software and its data wherever it runs.

WHAT WE ARE LOOKING FOR

  • Bachelor’s or Master’s degree in Computer Science or a related engineering field, and 6+ years building and shipping production software, with real depth in cloud infrastructure and platform engineering.
  • Deep hands-on experience building cloud infrastructure with Kubernetes and infrastructure-as-code (Terraform or similar) on at least one major cloud (AWS, GCP, or Azure).
  • Experience designing and operating core backend services and APIs that other engineers and systems depend on.
  • Direct ownership of a cloud deployment process: build and release pipelines, versioned artifacts, and safe, reproducible rollouts.
  • Hands-on experience building observability into production systems (metrics, logs, and traces) and using it to debug real incidents (Prometheus, Grafana, OpenTelemetry, or similar).
  • Strong software-engineering fundamentals and hands-on coding in a systems or backend language (Go, Python, Rust, or similar): this role builds the platform in code.
  • Experience defining service-level objectives and designing for reliability, making systems observable and reproducible from the start.
  • Comfort building foundational systems others depend on in an early-stage, ambiguous environment, making sensible scope, speed, and quality trade-offs.

NICE TO HAVE

  • Experience building platform components that run across varied deployment environments, including customer-controlled ones.
  • Experience running workloads across more than one runtime: cloud Kubernetes plus HPC schedulers (Slurm or similar) or bare metal.
  • GPU scheduling or multi-tenant cluster experience.
  • GitOps and progressive-delivery patterns (ArgoCD, Flux, staged rollouts).
  • Experience packaging or serving ML models, or supporting ML and data workloads on a shared platform.
  • Exposure to scientific computing, simulation, or other large-scale technical workloads.

LOCATION

Singapore or United States. We’re hiring in both to reach the right person. Work model is on-site or hybrid, set per location.

CLOSING NOTE

If you don’t tick every box but this is clearly your kind of work, get in touch.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Platform Engineer, Core Services (Xora Portfolio Company)
Senior Platform Engineer, Core Services (Xora Portfolio Company)

Xora Innovation • San Diego (CA)

Hybrid
USD 150,000 - 190,000
Health benefits
Senior Data & ML Infrastructure Engineer (Xora Portfolio Company)
Senior Data & ML Infrastructure Engineer (Xora Portfolio Company)

Xora Innovation • San Diego (CA)

Hybrid
USD 180,000 - 250,000
Engineer II, Platform
Engineer II, Platform

Nightingale Education Sole Mb • Salt Lake City (UT)

On-site
USD 120,000 - 180,000
Staff Platform Engineer
Staff Platform Engineer

Horizon3 • United States

Remote
USD 180,000 - 230,000
Growth opportunities
Innovation-driven culture
Flexible remote work
+1
Platform Engineer
Platform Engineer

The HT Group • Round Rock (TX)

Hybrid
USD 130,000 - 190,000
Platform Engineer
Platform Engineer

Synopsys Inc • Sunnyvale (CA)

On-site
USD 150,000 - 190,000
Platform Site Reliability Engineer
Platform Site Reliability Engineer

Specter • San Francisco (CA)

On-site
USD 180,000 - 230,000
Principal Cloud Engineer - Scientific Data and AI
Principal Cloud Engineer - Scientific Data and AI

Evolution USA • Boston (MA)

On-site
USD 120,000 - 180,000
Platform Engineer
Platform Engineer

Synergy • Chicago (IL)

On-site
USD 100,000 - 150,000
Platform Architect
Platform Architect

Halo Recruiting • Memphis (TN)

On-site
USD 140,000 - 190,000