Cloud Site Reliability Engineer

Stefanini, Inc

Dallas, Northern (TX, KY)

Hybrid

USD 117,000 - 124,000

Part time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Stefanini Group is hiring a Senior Cloud Site Reliability Engineer to design, build, and maintain reliability solutions for the Cloud Foundation Services. You will work with Terraform to automate AWS resources, develop CI/CD pipelines, and implement SRE metrics.

This remote-friendly contract role emphasizes incidents, postmortems, and cross-functional collaboration. The successful candidate will have 7+ years in software development with reliability focus, strong Python skills, and hands-on AWS

Qualifications

  • :

Responsibilities

  • Design, develop, and maintain reliability solutions and SRE utilities to reduce toil, improve cloud platform reliability, and industrialize SRE practices across the system
  • Build and optimize Infrastructure as Code (IaC) using Terraform to manage AWS resources related to SRE solutions, incorporating cost-efficient design principles
  • Develop CI/CD pipelines and automated testing to ensure code quality, reliability, and rapid delivery of the solutions
  • Define SRE standards, best practices, and guidelines for adoption across teams; establish SRE metrics like SLI, SLOs, etc.
  • Apply software engineering best practices including version control, code reviews, test-driven development, and documentation to all development
  • Participate in incident management and on-call rotation, providing technical support for SRE tools, troubleshooting production issues, and collaborating with teams to reduce incident recurrence through proactive detection and pattern analysis
  • Stay current with emerging AWS services, SRE methodologies, and cloud-native development technologies, and drive adoption of innovative solutions
  • Collaborate within Agile and Scaled Agile frameworks with cross-functional teams to deliver integrated cloud automation solutions
  • Produce clear, blameless postmortems with actionable items and documented failure scenarios

Skills

Python
Terraform
AWS
SRE
CI/CD
GoLang
Observability
DevOps
Incident Management
Postmortems

Education

Bachelor's degree in Computer Science or related field
Equivalent experience

Tools

Grafana
AWS CloudWatch
Terraform

Job description

Join us to co-create solutions for a better future!
Job Details
Information Technology

Cloud Site Reliability Engineer Dallas,TX

Posted:6/5/2026

Job ID#:64014

Job Category:Information Technology

Position Type:Contract

Duration:Long Term

Remaining Positions:1

Stefanini Group is hiring! Stefanini is looking for Cloud Site Reliability Engineer - Remote W2 Candidates only!

As a Senior Cloud Engineer in the Cloud SRE team, you will be responsible for designing and developing cloud solutions and engineering reliability tools for the Cloud Foundation Services (CFS) platform in the Infrastructure, Platforms & Operations organization. You will apply software engineering practices to build scalable, reusable solutions and utilities that enhance platform reliability.

Responsibilities
  • Design, develop, and maintain reliability solutions and SRE utilities to reduce toil, improve cloud platform reliability, and industrialize SRE practices across the system
  • Build and optimize Infrastructure as Code (IaC) using Terraform to manage AWS resources related to SRE solutions, incorporating cost-efficient design principles
  • Develop CI/CD pipelines and automated testing to ensure code quality, reliability, and rapid delivery of the solutions
  • Define SRE standards, best practices, and guidelines for adoption across teams; establish SRE metrics like SLI, SLOs, etc.
  • Apply software engineering best practices including version control, code reviews, test-driven development, and documentation to all development
  • Participate in incident management and on-call rotation, providing technical support for SRE tools, troubleshooting production issues, and collaborating with teams to reduce incident recurrence through proactive detection and pattern analysis
  • Stay current with emerging AWS services, SRE methodologies, and cloud-native development technologies, and drive adoption of innovative solutions
  • Collaborate within Agile and Scaled Agile frameworks with cross-functional teams to deliver integrated cloud automation solutions
  • Produce clear, blameless postmortems with actionable items and documented failure scenarios
Qualifications
  • Bachelor's degree in computer science, Information Systems, or equivalent background or equivalent experience
  • 7+ years of extensive experience in software development with focus on reliability and platform engineering
  • 5+ Years of advanced Python development skills with proven experience building enterprise-grade, highly available tools, APIs, and utilities
  • 3+ years of hands-on experience developing solutions in AWS environments with deep understanding of core services (EC2, VPC, S3, Lambda, IAM, CloudFormation, EventBridge, Step Functions etc.) and resource cost optimization
  • 3+ years of experience applying SRE principles including observability, toil automation, SLIs/SLOs and reliability engineering
  • Expert-level proficiency with Infrastructure as Code (IaC) using Terraform, including module development and state management
  • Strong experience with CI/CD pipelines, automated testing frameworks, and DevOps practices
  • Experience with observability tools and practices including Grafana, AWS CloudWatch, AWS Canary
  • Experience defining, implementing, and managing SLOs/SLIs and error budgets; familiarity with conducting RCAs and producing postmortem documentation
  • Working experience in Agile and Scaled Agile environments and familiarity with ITSM processes (incident, change, and problem management), resilience testing and chaos engineering practices
  • Experience with GoLang or additional programming languages is a plus

Stefanini takes pride in hiring top talent and developing relationships with our future employees. Our talent acquisition teams will never make an offer of employment without having a phone conversation with you. Those face-to-face conversations will involve a description of the job for which you have applied. We also speak with you about the process including interviews and job offers.

About Stefanini Group

The Stefanini Group is a global provider of offshore, onshore and near shore outsourcing, IT digital consulting, systems integration, application, and strategic staffing services to Fortune 1000 enterprises around the world. Our presence is in countries like the Americas, Europe, Africa, and Asia, and more than four hundred clients across a broad spectrum of markets, including financial services, manufacturing, telecommunications, chemical services, technology, public sector, and utilities. Stefanini is a CMM level 5, IT consulting company with a global presence. We are CMM Level 5 company.

Pay Range: $ 85.00 - $ 90.00

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AWS Cloud Engineer
AWS Cloud Engineer

Stefanini, Inc • San Francisco (CA)

On-site
USD 124,000 - 129,000
Onsite
AWS Cloud Engineer
AWS Cloud Engineer

Stefanini Group • San Francisco (CA)

On-site
USD 140,000 - 210,000
Cloud Engineer
Cloud Engineer

Stefanini, Inc • Dearborn (MI), Northern (KY)

Hybrid
USD 174,790,000 - 189,117,000
AWS DevOps Engineer
AWS DevOps Engineer

Stefanini Group • San Francisco (CA)

On-site
USD 140,000 - 190,000
AWS DevOps Engineer
AWS DevOps Engineer

Stefanini, Inc • San Francisco (CA)

On-site
USD 138,000 - 146,000
Lead Cloud Data Engineer (Data Mesh & AI)
Lead Cloud Data Engineer (Data Mesh & AI)

Stefanini, Inc • San Francisco (CA)

On-site
AWS Platform Engineer
AWS Platform Engineer

Stefanini Group • San Francisco (CA)

On-site
USD 120,000 - 160,000
Linux Server Administrator
Linux Server Administrator

Stefanini, Inc • Pittsburgh, Northern (KY)

Hybrid
USD 114,616,000 - 123,213,000
Cloud Infrastructure Site Reliability Engineer
Cloud Infrastructure Site Reliability Engineer

Robotics Prcocess Automation, LLC • Berkeley Heights (NJ)

On-site
Project Coordinator
Project Coordinator

Stefanini, Inc • Dallas (TX), Northern (KY)

Hybrid
USD 62,000 - 69,000