Senior Cloud Reliability Solutions Engineer

MathWorks

Natick (MA)

Hybrid

USD 140,000 - 200,000

Full time

48 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

MathWorks seeks a Senior Cloud Reliability Solutions Engineer to design resilient hybrid-cloud platforms across Azure, AWS, GCP, and on-premises VMware. You will collaborate with application, security, finance, and infrastructure teams to improve reliability, automation, and observability.

The role emphasizes Terraform, Packer, PowerShell, Python, and CI/CD practices, with a focus on Azure-centric architectures and cost-aware engineering.

Qualifications

  • Bachelor's degree or higher with professional experience.
  • Extensive Azure platform knowledge across networking, compute, storage, identity, and governance.
  • Strong multi-cloud familiarity with AWS and GCP services.
  • Proven IaC and automation implementation experience.

Responsibilities

  • Design and deliver hybrid-cloud solutions across Azure, AWS, GCP, and on-prem VMware.
  • Partner with customers to understand requirements, assess tradeoffs, and estimate cost/effort.
  • Build infrastructure automation with Terraform, Packer, PowerShell, Python, Git-based workflows, and CI/CD.
  • Apply reliability practices including incident response, RCA, capacity planning, and DR.
  • Create observability through dashboards, alerts, telemetry, health checks, and reporting.
  • Support secure platform operations with RBAC, secrets management, tagging, and patching.
  • Contribute to Kubernetes and platform services: AKS/EKS/GKE, networking, storage, Helm, GitOps.
  • Bridge traditional and cloud-native infra, supporting VMware, Windows, Linux, DNS, and networking.
  • Collaborate openly, document designs, mentor others, and learn with the team.
  • Participate in production support including on-call rotations.

Skills

Azure platform expertise
Multi-cloud literacy
Infrastructure as Code and automation
Reliability engineering
Containers and Kubernetes
VMware and enterprise infrastructure
Security and governance
Observability tools
Systems administration
Communication and consulting skills

Education

Bachelor's degree or higher

Tools

Terraform
Packer
Ansible
PowerShell
Python
YAML
GitHub/GitLab
CI/CD pipelines

Job description

Summary

MathWorks has a hybrid work model that enables staff members to split their time between office and home. The hybrid model provides the advantage of having both in-person time with colleagues and flexible at-home life optimizations. Learn More: https://www.mathworks.com/company/jobs/resources/applying-and-interviewing.html#onboarding.

MathWorks is a company built around problem solving, collaboration, and helping customers do their best technical work. The SSG Hosting team provides reliable, secure, and scalable infrastructure services that support teams across the company. We are looking for a friendly, curious, and technically strong Senior Cloud Reliability Solutions Engineer who enjoys designing practical solutions, automating repeatable work, and partnering with application, security, finance, and infrastructure teams to make our platforms easier to use and operate. In this role, you will help shape Azure-centered solutions that support workloads across Azure, AWS, GCP, and on-premises VMware environments, while advancing reliability engineering, automation, observability, and operational excellence practices.

MathWorks nurtures growth, appreciates inclusivity, encourages initiative, values teamwork, shares success, and rewards excellence.

Responsibilities
  • Design and deliver resilient hybrid-cloud solutions across Azure, AWS, GCP, and on-premises VMware environments, with deep emphasis on Azure architecture and operations.
  • Partner with internal customers to understand requirements, evaluate tradeoffs, estimate cost and effort, and recommend solutions that are reliable, secure, supportable, and aligned with business needs.
  • Build and improve infrastructure automation using Terraform, Packer, PowerShell, Python, Git-based workflows, and CI/CD pipelines to reduce manual work and improve consistency.
  • Apply reliability engineering practices including service ownership, operational readiness, incident response, root cause analysis, capacity planning, disaster recovery, and continuous improvement.
  • Create and maintain observability through useful dashboards, actionable alerts, telemetry standards, health checks, and reliability reporting across cloud and data center platforms.
  • Support secure and governed platform operations through RBAC, least privilege access, secrets management, policy enforcement, patching practices, vulnerability remediation, tagging, and cost visibility.
  • Contribute to Kubernetes and platform services including AKS, EKS, GKE, container networking, ingress, persistent storage, Helm, and GitOps-oriented operating models.
  • Bridge traditional and cloud-native infrastructure by supporting VMware, Windows, Linux, storage, networking, DNS, backup, and hybrid connectivity patterns.
  • Collaborate openly and constructively with peers and partner teams, sharing knowledge, documenting designs, mentoring others, and learning from the team.
  • Participate in production support including a rotating on-call schedule for infrastructure services and critical operational issues.
Minimum Qualifications
  • A bachelor's degree and 6 years of professional work experience (or a master's degree and 3 years of professional work experience, or a PhD degree, or equivalent experience) is required.
Additional Qualifications

A strong candidate will bring a combination of the following skills and experiences:

  • Azure platform expertise: Azure networking, compute, storage, identity, Azure Arc, Azure Monitor, Log Analytics, Azure Policy, RBAC, Azure Update Manager, Key Vault, Backup, and Site Recovery.
  • Multi-cloud literacy: working knowledge of AWS services such as EC2, VPC, IAM, S3, EKS, CloudWatch, Systems Manager, Route 53, and Organizations; familiarity with GCP Compute Engine, VPC, IAM, Cloud Storage, GKE, and Cloud Logging/Monitoring.
  • Infrastructure as Code and automation: Terraform, Packer, Ansible or similar configuration tools, PowerShell, Python, YAML, GitHub or GitLab, and CI/CD pipeline practices.
  • Reliability engineering: SLA/SLO, error budgets, incident management, RCA, disaster recovery, high availability, capacity planning, and operational readiness reviews.
  • Containers and Kubernetes: Docker, Kubernetes, AKS, EKS, GKE, Helm, ingress controllers, container networking, persistent storage, secrets management, and cluster lifecycle practices.
  • VMware and enterprise infrastructure: vSphere, ESXi, vCenter, virtual networking, storage, backup, migration planning, and hybrid integration patterns.
  • Security and governance: identity and access management, least privilege, network segmentation, certificate management, vulnerability remediation, cloud policy, compliance reporting, and cost/tagging governance.
  • Observability tools: Azure Monitor, Log Analytics, Application Insights, Prometheus, Grafana, Splunk, Datadog, or similar monitoring and logging platforms.
  • Systems administration: Windows Server, Linux distributions such as Ubuntu, RHEL, or Rocky Linux, DNS, networking, patching, performance troubleshooting, and service management.
  • Communication and consulting skills: clear written documentation, design reviews, stakeholder engagement, influence without authority, and the ability to explain technical options to both engineering and business audiences.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Reliability Engineer — Azure & Multi-Cloud
Senior Cloud Reliability Engineer — Azure & Multi-Cloud

MathWorks • Natick (MA)

Hybrid
USD 140,000 - 200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Chandler (AZ)

On-site
USD 140,000 - 190,000
Azure Platform Reliability Engineer IV
Azure Platform Reliability Engineer IV

ViziRecruiter,LLC. • Quincy (MA)

Hybrid
USD 125,000 - 188,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Chandler (MN)

On-site
USD 100,000 - 130,000
Reliability Engineer
Reliability Engineer

Compunnel, Inc. • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Senior Cloud Engineer (Azure)
Senior Cloud Engineer (Azure)

GPRS • Kentucky

On-site
USD 100,000 - 130,000
Medical, dental, and vision insurance
401(k) with company matching
Paid holidays
+2
Senior Cloud Engineer / Architect
Senior Cloud Engineer / Architect

Accord Technologies Inc • Columbia (SC)

Hybrid
USD 120,000 - 180,000
Hybrid work model
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

Hybrid
USD 100,000 - 135,000
Systems Engineer - Azure
Systems Engineer - Azure

Dunhill Professional Search & Government Solutions • Germantown (MD)

On-site
USD 100,000 - 130,000
Senior Cloud Engineer
Senior Cloud Engineer

Compunnel, Inc. • McLean (VA)

On-site
USD 120,000 - 150,000