Senior Platform Reliability Engineer (Kubernetes & CI/CD)

Optomi

Orlando (FL)

Hybrid

USD 120,000 - 180,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Optomi, in partnership with a client, seeks a Senior Systems Reliability Engineer to design and automate enterprise platform infrastructure powering AI platforms and observability solutions. You will build scalable Kubernetes-based systems, standardize CI/CD pipelines, and enable self-service platform capabilities across large-scale environments.

The ideal candidate has deep Kubernetes, cloud, Linux administration, automation, and platform engineering experience, plus leadership to influence

Qualifications

  • 5+ years of experience in Systems Reliability Engineering, Platform Engineering, Infrastructure Engineering, or DevOps within large enterprise environments.
  • Deep expertise managing and supporting production Kubernetes environments, including monitoring, capacity planning, performance tuning, and high availability.
  • Strong experience with Infrastructure as Code using Terraform, OpenTofu, AWS CDK, Ansible, or similar automation frameworks.
  • Hands-on experience building and maintaining CI/CD pipelines using GitLab CI/CD, GitHub Actions, Jenkins, or similar tools.
  • Strong Linux systems administration experience, including performance monitoring, troubleshooting, configuration management, and automation.
  • Experience developing reusable platform services, internal developer tools, shared libraries, or self-service infrastructure modules consumed across engineering teams.
  • Experience utilizing AI-assisted development tools such as Claude Code, Cursor, GitHub Copilot, or AWS Bedrock to improve development, automation, and documentation workflows.

Responsibilities

  • Design, build, and support highly available platform infrastructure across cloud and on-premises environments with a focus on reliability, scalability, and operational excellence.
  • Manage and optimize Kubernetes clusters through monitoring, capacity planning, performance tuning, troubleshooting, and automation.
  • Develop and maintain Infrastructure as Code solutions using Terraform and other automation frameworks to standardize infrastructure deployments.
  • Design, build, and maintain reusable CI/CD pipeline components, deployment templates, and developer tooling to improve engineering productivity.
  • Partner with software engineering, platform, and infrastructure teams to automate operational processes and enhance platform capabilities.
  • Implement monitoring, logging, telemetry, and observability solutions to proactively identify performance issues and improve system health.
  • Create technical documentation, platform standards, and engineering best practices while serving as a technical leader across cross-functional teams.

Skills

Platform engineering
Automation
Observability
Technical leadership
AI-assisted development

Tools

Kubernetes
Terraform
OpenTofu
AWS CDK
Ansible
GitLab CI/CD
GitHub Actions
Jenkins

Job description

Optomi, in partnership with a client, seeks a Senior Systems Reliability Engineer to design and automate enterprise platform infrastructure powering AI platforms and observability solutions. You will build scalable Kubernetes-based systems, standardize CI/CD pipelines, and enable self-service platform capabilities across large-scale environments.

The ideal candidate has deep Kubernetes, cloud, Linux administration, automation, and platform engineering experience, plus leadership to influence

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE (Contract/Hybrid)
Senior SRE (Contract/Hybrid)

Optomi • Orlando (FL)

Hybrid
USD 120,000 - 180,000
Senior Site Reliability Engineer - Kubernetes & AI Platform
Senior Site Reliability Engineer - Kubernetes & AI Platform

Optomi • Seattle (WA)

On-site
USD 140,000 - 180,000
Senior Platform Reliability Engineer - Kubernetes & Observability
Senior Platform Reliability Engineer - Kubernetes & Observability

PLP Group • New York (NY)

On-site
USD 75,000 - 130,000
Senior Kubernetes Platform Engineer: Reliability & Observability
Senior Kubernetes Platform Engineer: Reliability & Observability

Veriipro • Phoenix (AZ)

On-site
USD 120,000 - 170,000
Health insurance
401(k) plan
Paid time off
Senior Director, Site Reliability and Platform Engineering
Senior Director, Site Reliability and Platform Engineering

Optomi • Tacoma (WA)

Hybrid
USD 150,000 - 200,000
Medical insurance
Vision insurance
Senior OpenShift & Kubernetes Platform Engineer
Senior OpenShift & Kubernetes Platform Engineer

Optum • Eden Prairie (MN)

On-site
USD 112,000 - 194,000
Benefits package
401k contribution
Equity stock purchase
Senior Platform Engineer - Kubernetes, CI/CD & Automation
Senior Platform Engineer - Kubernetes, CI/CD & Automation

Motion Recruitment • Philadelphia

On-site
USD 120,000 - 160,000
Medical, Dental, and Vision Insurance
401(k) with company match
Paid Time Off and Holidays
+1
Lead SRE, Generative AI Platform (Remote)
Lead SRE, Generative AI Platform (Remote)

Optomi • United States

On-site
USD 150,000 - 190,000
Remote flexibility
Cutting-edge AI platform experience
Mentoring engineers and shaping cloud架
AWS Platform Engineer: Cloud, CI/CD & Observability
AWS Platform Engineer: Cloud, CI/CD & Observability

Optomi • Naperville (IL)

On-site
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Optomi • Seattle (WA)

On-site
USD 140,000 - 180,000