Senior Cloud Infrastructure / SRE Engineer (AWS & Hybrid)

Infinite Computer Solutions

Frisco (TX)

Hybrid

USD 125,000 - 180,000

Full time

5 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Infinite Computer Solutions in Frisco, TX seeks a Senior Cloud Infrastructure / SRE Engineer to design, deploy, and operate scalable cloud platforms in a hybrid AWS/on-premises environment. You will lead Terraform-driven provisioning, manage Linux systems, and oversee incident response to ensure high availability across production services.

The role requires hands-on work with AWS services (EC2, ASG, ALB/NLB, VPC, IAM, S3, EBS, EFS, CloudWatch) and modern CI/CD pipelines (GitHub Actions,

Qualifications

  • Strong experience provisioning and managing AWS infrastructure using Terraform.
  • Hands-on experience with Ansible for configuration management, application deployment, and infrastructure automation.
  • Strong Linux administration and operating system troubleshooting skills.
  • Experience supporting both on-premises and AWS cloud-based enterprise applications and infrastructure.
  • Hands-on experience with AWS services including EC2, Auto Scaling Groups (ASG), Elastic Load Balancers (ALB/NLB), Route 53, VPC, IAM, S3, EBS, EFS, CloudWatch, Systems Manager (SSM), KMS, Secrets Manager, and AWS Backup.
  • Proven experience in cloud transformation and migration initiatives from on-premises environments to AWS.
  • Strong Site Reliability Engineering (SRE) experience supporting business-critical production systems.
  • Experience managing production incidents, problem management, root cause analysis (RCA), disaster recovery, and service restoration.
  • Hands-on experience supporting Java/J2EE applications, WebLogic, Tomcat, or similar middleware platforms in production environments.
  • Experience with DevOps practices and CI/CD pipelines using GitHub Actions, Harness, Jenkins, AWS CodePipeline, CodeBuild, or similar tools.
  • Strong scripting and automation skills using Bash, Python, Shell scripting, or PowerShell.
  • Solid understanding of Layer 4 and Layer 7 networking concepts, load balancing, DNS, SSL/TLS certificates, WAF, CDN, firewall technologies, and hybrid connectivity.
  • Hands-on experience with containerization and orchestration technologies, including Docker, Kubernetes, and Amazon EKS.
  • Experience implementing observability, monitoring, logging, alerting, and performance management using tools such as Dynatrace, Splunk, CloudWatch, Prometheus, or Grafana.
  • Experience with AWS security services, IAM roles and policies, Secrets Manager, KMS encryption, CloudTrail, Security Hub, and security best practices.
  • Strong understanding of high availability, disaster recovery, backup, resiliency, and multi-region architectures.
  • Experience supporting and optimizing Auto Scaling, elastic workloads, and cloud cost optimization initiatives.
  • Knowledge of release management, change management, incident management, and ITIL service management processes.
  • Strong troubleshooting, analytical, and problem-solving skills across infrastructure, application, database, middleware, and network layers.
  • Excellent verbal and written communication skills with the ability to drive technical discussions, operational excellence, and cloud transformation initiatives.

Skills

AWS Cloud
SRE Practices
Linux Administration
CI/CD Automation
Networking
Troubleshooting
Scripting (Bash/Python)
Incident Management

Tools

Terraform
Ansible
Docker
Kubernetes
Jenkins
GitHub Actions
Harness
AWS CodePipeline/CodeBuild
Dynatrace/Splunk

Job description

and new BR. Please consider this on priority

Job Description – Senior Cloud Infrastructure / SRE Engineer (AWS & Hybrid)
Required Skills & Experience
  • Strong experience provisioning and managing AWS infrastructure using Terraform.
  • Hands-on experience with Ansible for configuration management, application deployment, and infrastructure automation.
  • Strong Linux administration and operating system troubleshooting skills.
  • Experience supporting both on-premises and AWS cloud-based enterprise applications and infrastructure.
  • Hands-on experience with AWS services including EC2, Auto Scaling Groups (ASG), Elastic Load Balancers (ALB/NLB), Route 53, VPC, IAM, S3, EBS, EFS, CloudWatch, Systems Manager (SSM), KMS, Secrets Manager, and AWS Backup.
  • Proven experience in cloud transformation and migration initiatives from on-premises environments to AWS.
  • Strong Site Reliability Engineering (SRE) experience supporting business-critical production systems.
  • Experience managing production incidents, problem management, root cause analysis (RCA), disaster recovery, and service restoration.
  • Hands-on experience supporting Java/J2EE applications, WebLogic, Tomcat, or similar middleware platforms in production environments.
  • Experience with DevOps practices and CI/CD pipelines using GitHub Actions, Harness, Jenkins, AWS CodePipeline, CodeBuild, or similar tools.
  • Strong scripting and automation skills using Bash, Python, Shell scripting, or PowerShell.
  • Solid understanding of Layer 4 and Layer 7 networking concepts, load balancing, DNS, SSL/TLS certificates, WAF, CDN, firewall technologies, and hybrid connectivity.
  • Hands-on experience with containerization and orchestration technologies, including Docker, Kubernetes, and Amazon EKS.
  • Experience implementing observability, monitoring, logging, alerting, and performance management using tools such as Dynatrace, Splunk, CloudWatch, Prometheus, or Grafana.
  • Experience with AWS security services, IAM roles and policies, Secrets Manager, KMS encryption, CloudTrail, Security Hub, and security best practices.
  • Strong understanding of high availability, disaster recovery, backup, resiliency, and multi-region architectures.
  • Experience supporting and optimizing Auto Scaling, elastic workloads, and cloud cost optimization initiatives.
  • Knowledge of release management, change management, incident management, and ITIL service management processes.
  • Strong troubleshooting, analytical, and problem-solving skills across infrastructure, application, database, middleware, and network layers.
  • Excellent verbal and written communication skills with the ability to drive technical discussions, operational excellence, and cloud transformation initiatives.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer – AWS/Datacenter
Senior Site Reliability Engineer – AWS/Datacenter

Jobtailor • Bellevue (WA)

On-site
USD 160,000 - 220,000
Senior Cloud Engineer
Senior Cloud Engineer

Compunnel, Inc. • McLean (VA)

On-site
USD 120,000 - 150,000
Senior DevOps Enginner (HYBRID)
Senior DevOps Enginner (HYBRID)

Resolution Technologies, Inc. • Plano (TX)

On-site
USD 100,000 - 130,000
Senior Cloud DevOps Engineer
Senior Cloud DevOps Engineer

Sulekha.com New Media Pvt Ltd • Dallas (TX), Northern (KY)

Hybrid
USD 140,000 - 180,000
AWS Cloud Engineer
AWS Cloud Engineer

PSEG • Town of Oyster Bay (NY)

On-site
USD 120,000 - 160,000
Senior Cloud Engineer
Senior Cloud Engineer

ESG • Houston (TX)

Hybrid
USD 100,000 - 130,000
AWS Cloud Platform Engineer
AWS Cloud Platform Engineer

Glint Tech Solutions • Reston (VA)

Hybrid
USD 120,000 - 150,000
Sr. Cloud DevOps Engineer
Sr. Cloud DevOps Engineer

Accord Technologies Inc • Boston (MA)

On-site
USD 120,000 - 150,000
Systems Engineer – AWS, Linux/Windows, Terraform & Kubernetes
Systems Engineer – AWS, Linux/Windows, Terraform & Kubernetes

Motion Recruitment Partners LLC • Los Angeles (CA), Northern (KY)

Hybrid
USD 125,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Motion Recruitment Partners LLC • Chicago (IL), Northern (KY)

Hybrid
USD 140,000 - 170,000