Mid-to-Senior Cloud SRE & DevOps Engineer

Mondrian Alpha

New York (NY)

On-site

USD 170,000 - 230,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Mondrian Alpha is seeking a mid-to-senior Cloud SRE & DevOps Engineer to design, operate, and evolve AI, cloud infrastructure, and data-intensive platforms supporting trading and business-critical systems.

You will work across AI infrastructure, cloud engineering, SRE, and platform engineering to ensure reliable, scalable production environments and improved developer experience through modern DevOps and GitOps practices.

Qualifications

  • Hands-on experience with public-cloud platforms in production environments.
  • Experience in hybrid or multi-cloud environments.
  • Deep experience with Kubernetes and containerized systems.
  • Familiarity with Kubernetes deployment and management tooling.
  • Infrastructure as code using Terraform or similar.
  • Experience building and managing CI/CD pipelines.
  • Git-based and GitOps workflows experience.
  • Proficiency in scripting languages (Python, Shell, PowerShell).
  • Understanding of distributed systems, networking, and cloud architecture.
  • Experience operating production systems with high availability and performance.
  • Experience diagnosing production incidents in distributed environments.
  • Experience supporting relational databases and cloud data platforms.
  • Knowledge of backup, disaster recovery, resiliency, and business continuity.

Skills

Cloud platforms
Kubernetes
Terraform
CI/CD pipelines
GitOps workflows
Python scripting
Distributed systems
Networking
Databases
Cloud architecture

Tools

Kubernetes tooling
Terraform

Job description

Overview

A Global Asset Manager is looking for a mid-to-senior level engineer to join a Cloud SRE & DevOps Engineering team focused on building, operating, and evolving AI, cloud infrastructure, and software delivery platforms that support trading, investment, and other business-critical systems.

This role sits at the intersection of AI Infrastructure Engineering, Cloud Engineering, Site Reliability Engineering (SRE), DevOps, Platform Engineering, and Production Engineering. It is designed for individuals who take ownership of systems running in production.

You will be responsible for designing and operating resilient, scalable environments across public cloud and Kubernetes platforms while enabling engineering teams through modern DevOps and GitOps practices.

The position reflects a strong SRE mentality, with an emphasis on reliability, observability, automation, security, and operational excellence. You will contribute to cloud transformation initiatives, ensuring systems are built for performance, stability, scalability, and resilience.

This is a hands-on role requiring deep technical expertise, accountability for production systems, and a mindset oriented toward continuous improvement, risk management, and engineering efficiency. The role also contributes to evolving platform capabilities supporting AI, machine learning, and data-intensive workloads.

You will work closely with software engineering, cybersecurity, data, and quantitative technology teams to deliver secure, scalable, and high-performing systems while improving developer experience and platform maturity across the organization.

What We're Looking For
Core Technical Expertise
  • Strong hands-on experience with public-cloud platforms in production environments.
  • Experience working in hybrid or multi-cloud environments.
  • Deep experience with Kubernetes and containerized systems.
  • Strong familiarity with Kubernetes deployment and management tooling.
  • Proven experience with infrastructure as code, preferably Terraform or a comparable technology.
  • Strong experience building and managing CI/CD pipelines.
  • Experience with Git-based and GitOps workflows.
  • Proficiency in Python, Shell, PowerShell, or similar scripting languages for automation.
  • Strong understanding of distributed systems, networking, and cloud architecture.
  • Experience operating and supporting production systems with high availability and performance requirements.
  • Experience diagnosing and resolving complex production incidents in distributed environments.
  • Experience supporting relational databases and cloud-based data platforms.
  • Understanding of backup, disaster recovery, resiliency, and business continuity practices.
  • Strong troubleshooting and problem-solving skills.
Monitoring & Observability
  • Hands-on experience with enterprise monitoring and observability platforms or comparable technologies.
  • Experience designing metrics, alerts, dashboards, and operational monitoring for production systems.
  • Strong understanding of logging, tracing, alerting, and proactive system health monitoring.
  • Experience improving system observability and reducing operational risk.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps / Cloud Infra Engineer
Senior DevOps / Cloud Infra Engineer

Reqiva • New York (NY)

On-site
USD 120,000 - 170,000
Senior Cloud Platform Engineer - AWS, Kubernetes, IaC
Senior Cloud Platform Engineer - AWS, Kubernetes, IaC

Mission Staffing • New York (NY)

Hybrid
USD 130,000 - 160,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink, Inc. • Northern (KY)

Hybrid
USD 140,000 - 210,000
Senior Cloud SRE & DevOps Engineer - Kubernetes & AI Infra
Senior Cloud SRE & DevOps Engineer - Kubernetes & AI Infra

Mondrian Alpha • New York (NY)

On-site
USD 170,000 - 230,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink • United States

On-site
USD 140,000 - 190,000
Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 210,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

State of Wisconsin Investment Board • Madison (WI)

On-site
USD 140,000 - 180,000
Sr SRE Automation Engineer
Sr SRE Automation Engineer

Compunnel, Inc. • Austin (TX), Northern (KY)

Hybrid
USD 130,000 - 180,000
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior SRE Engineer
Senior SRE Engineer

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 140,000 - 190,000