Remote SRE Engineer: Cloud Reliability & Automation

Noctua Technology

United States

Remote

USD 107,000 - 178,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Noctua Technology is seeking a Site Reliability Engineer to apply software engineering to operations, ensuring the reliability, scalability, and performance of production systems across cloud-native environments.

You will define SLOs, build IaC with Terraform/CloudFormation, deploy containerized and serverless apps with Docker and Kubernetes, and automate CI/CD pipelines to reduce toil while maintaining high availability. US citizenship and active Secret clearance eligibility are required.

Qualifications

  • 1-5 years of experience in site reliability engineering, cloud engineering, or related fields.
  • Strong software engineering skills for automation and system management.
  • Proficiency with IaC tools like Terraform or CloudFormation.
  • Experience with Docker and Kubernetes.
  • Knowledge of networking concepts, cloud security best practices, and IAM.
  • Experience with programming or scripting languages such as Python, Bash, or Go.
  • Familiarity with CI/CD pipelines and DevOps methodologies.
  • Strong problem-solving skills and the ability to troubleshoot complex cloud environments.
  • Effective communication skills and a willingness to learn and collaborate.

Responsibilities

  • Define, measure, and report on SLIs and SLOs to ensure system reliability and uptime.
  • Develop and deploy Infrastructure as Code (IaC) using Terraform, CloudFormation, or similar tools, with an emphasis on repeatability and change management.
  • Implement and manage containerized and serverless architectures using Docker, Kubernetes, and cloud-native services, focusing on performance and error budgets.
  • Build and maintain reliable and self-healing CI/CD pipelines to automate deployments and improve development workflows.
  • Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively.
  • Eliminate toil by extensively automating operational tasks, including provisioning, patching, and deployments, using scripting and configuration management tools such as Python, Bash, or Ansible.
  • Conduct post-incident reviews (blameless postmortems) to drive continuous improvement in system reliability and operational processes.
  • Collaborate with development teams to improve operability and production readiness of applications from design through deployment.
  • Create and maintain documentation for cloud architectures, deployment processes, and best practices.
  • Provide technical guidance and support to clients and internal teams on cloud infrastructure and reliability best practices, with a focus on defining SLAs.

Skills

Python
Bash
Go
Terraform
CloudFormation
Docker
Kubernetes
CI/CD pipelines
Networking concepts
Cloud security

Education

Bachelor's degree in Computer Science or related field

Tools

Terraform
CloudFormation
Docker
Kubernetes

Job description

Noctua Technology is seeking a Site Reliability Engineer to apply software engineering to operations, ensuring the reliability, scalability, and performance of production systems across cloud-native environments.

You will define SLOs, build IaC with Terraform/CloudFormation, deploy containerized and serverless apps with Docker and Kubernetes, and automate CI/CD pipelines to reduce toil while maintaining high availability. US citizenship and active Secret clearance eligibility are required.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote SRE Engineer — Cloud Reliability & Automation
Remote SRE Engineer — Cloud Reliability & Automation

Noctua Technology • Virginia (MN), California (MO), Washington

Remote
USD 106,500 - 177,500
Senior SRE Engineer: Remote Cloud Reliability & Automation
Senior SRE Engineer: Remote Cloud Reliability & Automation

Noctua Technology • Virginia (MN), California (MO), Washington

Remote
USD 149,000 - 202,000
Software Engineer- Site Reliability Engineering (SRE)
Software Engineer- Site Reliability Engineering (SRE)

Noctua Technology • United States

Remote
USD 107,000 - 178,000
Software Engineer- Site Reliability Engineering (SRE)
Software Engineer- Site Reliability Engineering (SRE)

Noctua Technology • Virginia (MN), California (MO), Washington

Remote
USD 106,500 - 177,500
Senior Software Engineer- Site Reliability Engineering (SRE)
Senior Software Engineer- Site Reliability Engineering (SRE)

Noctua Technology • Virginia (MN), California (MO), Washington

Remote
USD 149,000 - 202,000
Senior SRE: Cloud Reliability, Terraform & Kubernetes (Remote)
Senior SRE: Cloud Reliability, Terraform & Kubernetes (Remote)

Motion Recruitment • Chicago (IL)

On-site
USD 140,000 - 190,000
Remote SRE II: Cloud-Native Reliability & Automation
Remote SRE II: Cloud-Native Reliability & Automation

NationsBenefits, LLC • Plantation (FL)

On-site
USD 110,000 - 160,000
Unlimited PTO
Competitive compensation & benefits
Career growth opportunities
+1
Site Reliability Engineer
Site Reliability Engineer

Motion Recruitment Partners LLC • Chicago (IL), Northern (KY)

Hybrid
USD 140,000 - 170,000
Remote Senior Network Reliability Engineer (SRE)
Remote Senior Network Reliability Engineer (SRE)

Gainbridge • Zionsville (IN), Northern (KY)

On-site
USD 135,000 - 190,000
Health Insurance
Dental Insurance
Vision Insurance
+4
Remote SRE — Cloud Reliability & Performance
Remote SRE — Cloud Reliability & Performance

DevOpsChat • United States

Hybrid
USD 120,000 - 170,000
Healthcare options
Professional development
Flexible work location
+1