DevOps & Site Reliability Engineer

VoltaGrid

Houston (TX)

On-site

USD 140,000 - 180,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

VoltaGrid is seeking a DevOps & SRE Engineer to design and evolve the infrastructure, deployment pipelines, and reliability posture of our systems. You will collaborate with engineering to build scalable, observable, and resilient infrastructure while promoting operational excellence.

You will manage cloud infrastructure, containerized workloads, and IaC with Terraform, and contribute to monitoring, incident response, and capacity planning.

Qualifications

  • 4+ years of experience in DevOps, SRE, or infrastructure engineering roles.
  • Hands-on with Kubernetes and Docker in production environments.
  • Proficiency with infrastructure-as-code tools, particularly Terraform.
  • Experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, or similar).
  • Strong Linux systems administration skills (Ubuntu, RHEL/CentOS, or similar).
  • Solid understanding of networking, DNS, load balancing, and security fundamentals.

Responsibilities

  • Design, build, and maintain cloud infrastructure.
  • Manage and optimize Kubernetes clusters and containerized workloads in production.
  • Develop and maintain infrastructure-as-code using Terraform (or equivalent tooling).
  • Build and improve CI/CD pipelines to enable fast, safe, and reliable deployments.
  • Implement and maintain monitoring, alerting, and observability systems (Prometheus, Grafana, Datadog, or similar).
  • Define and track SLIs/SLOs, participate in incident response, root cause analysis, and blameless postmortems.
  • Identify and eliminate toil through automation and self-service tooling.
  • Configure and maintain on-prem bare-metal servers and Linux-based infrastructure.
  • Configure, maintain, and optimize virtualized assets.
  • Collaborate with development teams on system design, capacity planning, and performance optimization.
  • Participate in on-call rotations and ensure production readiness of new services.

Skills

Scripting
Incident management
Capacity planning
Linux administration
Networking fundamentals
Observability

Tools

Terraform
Kubernetes
Docker
GitHub Actions
GitLab CI
Jenkins

Job description

Position Title: DEVOPS & SRE ENGINEER

Location: HOUSTON, TX

FLSA Class: EXEMPT

Responsible to: Directo of Software Engineering

Position Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure, deployment pipelines, and reliability posture of our systems. You'll work closely with engineering teams to build scalable, observable, and resilient infrastructure while driving a culture of operational excellence.

Essential Duties And Responsibilities
  • Design, build, and maintain cloud infrastructure
  • Manage and optimize Kubernetes clusters and containerized workloads in production
  • Develop and maintain infrastructureascode using Terraform (or equivalent tooling)
  • Build and improve CI/CD pipelines to enable fast, safe, and reliable deployments
  • Implement and maintain monitoring, alerting, and observability systems (Prometheus, Grafana, Datadog, or similar)
  • Define and track SLIs/SLOs, participate in incident response, root cause analysis, and blameless postmortems
  • Identify and eliminate toil through automation and selfservice tooling
  • Configure and maintain onprem baremetal servers and Linuxbased infrastructure
  • Configure, maintain, and optimize virtualized assets
  • Collaborate with development teams on system design, capacity planning, and performance optimization
  • Participate in oncall rotations and ensure production readiness of new services
Other Requirements
  • 4+ years of experience in DevOps, SRE, or infrastructure engineering roles
  • Strong experience with at least one major cloud provider (AWS, GCP, or Azure AWS preferred)
  • Deep hands-on experience with Kubernetes and Docker in production environments
  • Proficiency with infrastructureascode tools, particularly Terraform
  • Experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, or similar)
  • Solid understanding of monitoring and observability (metrics, logs, traces)
  • Strong scripting skills (Bash, Python, or Go)
  • Experience with incident management, SLObased reliability practices, and capacity planning
  • Strong Linux systems administration skills (Ubuntu, RHEL/CentOS, or similar)
  • Experience with virtualization platforms including VM provisioning, storage, networking, and cluster management
  • Solid understanding of networking, DNS, load balancing, and security fundamentals
Nice To Have
  • Contributions to internal developer platforms or platform engineering initiatives
  • Proxmox VE experience
  • Certifications in cloud platforms (AWS SA, CKA, etc.)

The above statements are intended to describe the general nature and level of work being performed by employees assigned to this classification. All personnel may be required to perform duties outside of their normal responsibilities from time to time, as needed.

VoltaGrid is an Equal Opportunity Employer that does not discriminate on the basis of actual or perceived race, creed, color, religion, alienage or national origin, ancestry, citizenship status, age, disability or handicap, sex, marital status, veteran status, sexual orientation, genetic information, arrest record, or any other characteristic protected by applicable federal, state or local laws.

Our management team is dedicated to this policy with respect to recruitment, hiring, placement, promotion, transfer, training, compensation, benefits, employee activities, and general treatment during employment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Engineering Manager, SRE & DevOps
Systems Engineering Manager, SRE & DevOps

VoltaGrid, LLC • Houston (TX)

On-site
USD 140,000 - 190,000
DevOps
DevOps

Crawford Thomas Recruiting • Dallas (TX)

On-site
USD 100,000 - 160,000
Medical, Dental, Vision Insurance
Life Insurance
Paid Time Off
+2
DevOps Engineer
DevOps Engineer

VTG Defense • McLean (VA)

On-site
USD 100,000 - 140,000
Senior DevOps Cloud Engineer (Houston, TX)
Senior DevOps Cloud Engineer (Houston, TX)

ESG Global Ltd • Houston (TX)

Hybrid
USD 130,000 - 170,000
DevOps Engineer
DevOps Engineer

VT Group (VTG) • Chantilly (VA)

On-site
USD 120,000 - 150,000
DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine • Chicago (IL)

Hybrid
USD 120,000 - 150,000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine • Atlanta (GA)

Hybrid
USD 100,000 - 130,000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Technology Architect
DevOps Technology Architect

Robotics Prcocess Automation, LLC • Alpharetta (GA)

On-site
USD 90,000 - 120,000
Senior Technical Solutions Engineer – Platform
Senior Technical Solutions Engineer – Platform

VoltaGrid • Houston (TX)

On-site
USD 100,000 - 140,000
DevOps Engineer
DevOps Engineer

ComResource • Columbus (OH)

Hybrid
USD 100,000 - 140,000