Lead Software Engineer (DevOps)

Mastercard

Bray (OK)

On-site

USD 140,000 - 190,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mastercard is seeking a Lead Site Reliability Engineer to design, build, and maintain scalable AWS cloud infrastructure. You will lead a geo-diverse team, mentor juniors, and drive automation using Terraform, Ansible, and Jenkins.

The role emphasizes uptime, security, and fast delivery of services. Ideal candidates have deep Linux knowledge, strong Go/Python/Bash coding skills, and experience with Kubernetes, Prometheus, and cloud networking.

Qualifications

  • Led or acted as a leader within a team and projects.
  • Experience with provisioning tools such as Ansible/Chef/Terraform.
  • Strong background in cloud infrastructure and DevOps practices.

Responsibilities

  • Plan, design, and scale infrastructure across AWS environments.
  • Develop and maintain Infrastructure as Code with Jenkins, Terraform, and CloudFormation.
  • Improve security, monitoring, and availability for all services.
  • Deploy workloads to cloud environments and manage AWS core services.
  • Build and scale Kubernetes clusters (EKS) with reliability in mind.
  • Create and improve monitoring with Prometheus and related tools.
  • Provide on-call support and troubleshoot performance issues.
  • Write code in Go, Python, or Bash to automate tasks and integrations.

Skills

Leadership experience
Team collaboration
English communication

Education

BS in Computer Science or equivalent

Tools

AWS
Terraform
Ansible
Jenkins
Go
Python
Bash
Kubernetes
Prometheus
Docker
VPC
Direct Connect

Job description

Our Purpose

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

Title and Summary

Lead Software Engineer (DevOps)

Our Purpose: We work to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion (for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team - one that makes better decisions, drives innovation and delivers better business results.

Overview: The Identity Solutions program within the Services group works to prevent fraud by securing a user's identity through their devices, identity information, and usage patterns such as passive Biometrics. Comprised of best-in-class solutions from Ekata (a Mastercard company), we use complex machine learning, combined with device, identity, and transaction information from billions of transactions to reduce user friction to enable highly secure real time transactions.

The Identity Solutions - 'Cloud Infrastructure - Edge' team is looking for a Lead Site Reliability Engineer to help support, design, and build our highly available and scalable AWS Cloud infrastructure which our products are built on. The team builds infrastructure for Engineering projects, extends our platform for the future, and improves the availability and performance of our applications. This position also designs, develops systems, automation, and tools to help make it easier for Engineering teams to deploy services in a fast, automated and reliable fashion.

In this role You will
  • Be a team and highly accountable project leader in a geo-diverse team. Mentoring the more junior members and being an SME.
  • Plan, design, build, and scale our infrastructure.
  • Work with tools such as Jenkins, Ansible, Argo CD, Terraform, CloudFormation, Resource Manager and many more to ensure that our stack is well represented as Infrastructure as Code.
  • Manage, design, and improve security and availability monitoring and processes for all services, ensure defined security policies are consistently implemented across all environments.
  • Deploy and supervise workloads to cloud environments, proven experience with all of the core services within AWS including instance management, IAM configuration, Database, Caching and general support/troubleshooting.
  • Have a deep understanding of the core components required to run Kubernetes (EKS) and be able to build a cluster from scratch or scale out if needed.
  • Have perfected load balancing, service mesh and always looking for ways to improve availability, response time, and uptime.
  • Maintain, and manage quality documentation for systems owned by the Infrastructure team.
  • Design, build, and improve monitoring tools to identify and resolve issues before they happen. Mainly with Prometheus.
  • Help and advise other teams, troubleshoot and solve failures and performance problems, participate in on-call rotations.
  • Have a basic passion for working with Go, Python, Rust or even Bash to build custom tools and improve system integration. Take code ownership to the next level and act as an advocate for writing code that aligns with industry best practice.
All About You
  • Excellent spoken and written English skills. Is a team player and values collaboration.
  • BS degree in Computer Science or equivalent experience.
  • Led / been a leader within a team and projects.
  • AWS Networking Services
Proficiency with
  • Amazon Virtual Private Cloud (VPC)
  • AWS Direct Connect
  • Transit Gateway
  • Elastic Load Balancing (ELB)
  • Route 53, Resolver endpoints, DNS firewall
  • Security Groups and NACLs
  • AWS Firewall
  • VPC flow log analysis
  • Understanding of hybrid connectivity (VPNs, Direct Connect) and multi-region architectures
  • Understand implementation of latency reduction techniques and tools - iperf, sockperf
  • Proven deep skills with Linux or UNIX systems and related protocols/software with 3+ years' experience. Robust DevOps and/or BizOps experiecne.
  • A command of Linux systems including troubleshooting, memory management, tuning, I/O subsystem, RAID, and security.
  • Experience with provisioning tools such as Ansible/Chef/Terraform.
  • Experience with Jenkins or other CI/CD tools.
  • Programming aptitude in Go, Python, and Bash.
  • Knowledge of database systems such as MySQL or PostgreSQL.
  • Experience building and deploying Containers, including orchestration tools such as Kubernetes, Mesos, or Docker Swarm.
  • Experience with AWS cloud providers (AWS, Azure, GCP)
Corporate Security Responsibility
  • Abide by Mastercard's security policies and practices;
  • Ensure the confidentiality and integrity of the information being accessed;
  • Report any suspected information security violation or breach,
  • Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Software Engineer
Lead Software Engineer

Mastercard • Bray (OK)

On-site
USD 140,000 - 190,000
Lead Site Reliability Engineer2
Lead Site Reliability Engineer2

Mastercard • O’Fallon (MO)

On-site
USD 120,000 - 150,000
Senior Software Engineer - Java
Senior Software Engineer - Java

Mastercard • Bray (OK)

On-site
USD 120,000 - 180,000
Software Engineer (Full Stack)
Software Engineer (Full Stack)

Mastercard • United States

Hybrid
USD 120,000 - 190,000
Lead Site Reliability Engineer-2
Lead Site Reliability Engineer-2

Mastercard • O’Fallon (MO)

On-site
USD 122,000 - 207,000
Lead Software Engineer
Lead Software Engineer

Master Card • Town of Montana (WI)

On-site
USD 140,000 - 190,000
Medical Insurance
Paid time off
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Socket.dev • O’Fallon (MO)

On-site
USD 110,000 - 170,000
Lead Platform Engineer – AWS Cloud DevOps Engineer
Lead Platform Engineer – AWS Cloud DevOps Engineer

lwtsquad • O’Fallon (MO)

On-site
USD 112,000 - 187,000
Medical, dental, and vision insurance
401k with company match
Tuition reimbursement
+1
Senior Principal, Data Engineering
Senior Principal, Data Engineering

Mastercard • Arlington (VA)

Hybrid
USD 244,000 - 390,000
Medical, dental, and vision insurance
401k with company match
25 days vacation time
+1
Lead Software Engineer (Java/AI)
Lead Software Engineer (Java/AI)

Mastercard • Bray (OK)

Hybrid
USD 105,000 - 151,000