Site Reliability Engineer

Flanksource Inc.

Mission (KS)

Remote

USD 90,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

100% remote work
Flexible hours
Opportunity to work with cutting-edge technology

Job summary

A technology company is seeking a Kubernetes DevOps Engineer to design, maintain, and optimize Kubernetes clusters. Responsibilities include enhancing the reliability and scalability of cloud infrastructure, managing CI/CD pipelines, and collaborating with development teams. Ideal candidates will have experience in Infrastructure as Code tools and cloud platforms. This position offers 100% remote work with flexible hours.

Qualifications

  • Experience with 2+ Infrastructure as Code tools like Terraform or Pulumi.
  • Proficiency in monitoring tools such as Prometheus and Grafana.
  • Hands-on experience with cloud platforms like AWS, Azure, or GCP.
  • Knowledge of CI/CD processes using GitHub Actions, GitLab CI, or Azure DevOps.
  • Understanding of network fundamentals and security best practices.
  • Strong analytical and troubleshooting abilities.
  • Fluent English for remote asynchronous communication.
  • Ability to work independently in an agile manner.

Responsibilities

  • Design and maintain Kubernetes clusters across multiple environments.
  • Build automation for cluster deployment and management.
  • Monitor and troubleshoot clusters for high availability.
  • Implement security best practices for Kubernetes infrastructure.
  • Participate in incident response to reduce Mean Time To Recovery (MTTR).
  • Enhance reliability and scalability of Kubernetes infrastructure.
  • Manage CI/CD pipelines and DevOps tooling.
  • Collaborate with development teams.

Skills

Infrastructure as Code (IaC) tools
Proficiency with Prometheus
Cloud Platforms (AWS, Azure, GCP)
CI/CD processes
Networking fundamentals
Problem-solving
Fluent English
Self-motivated

Tools

Terraform
Pulumi
Grafana
GitHub Actions
GitLab CI
Azure DevOps

Job description

  • Design and maintain Kubernetes clusters across multiple environments (development, staging, production)
  • Build automation for cluster deployment, configuration, and management
  • Monitor and troubleshoot clusters to ensure high availability and optimal performance
  • Implement security best practices for Kubernetes and underlying infrastructure
  • Participate in incident response and work to reduce Mean Time To Recovery (MTTR)
  • Enhance the reliability and scalability of our Kubernetes infrastructure
  • Manage CI/CD pipelines and DevOps tooling
  • Collaborate with development teams on deployment strategies and best practices
Requirements
  • Infrastructure as Code - Experience with 2+ IaC tools (Terraform, Pulumi, etc.)
  • Monitoring & Observability - Proficiency with Prometheus, Grafana, and related tools
  • Cloud Platforms - Hands-on experience with AWS, Azure, or GCP
  • CI/CD - Knowledge of GitHub Actions, GitLab CI, or Azure DevOps
  • Networking & Security - Understanding of network fundamentals and security best practices
  • Problem-solving - Strong analytical and troubleshooting abilities
  • Communication - Fluent English for remote asynchronous work
  • Self-motivated - Ability to work independently with an agile approach
Nice-to-haves
  • Experience with GitOps tools (Flux, ArgoCD)
  • Go programming knowledge or willingness to learn
  • Active open-source contributions
  • Experience developing Kubernetes operators or controllers
  • 100% remote work with flexible hours
  • Work with cutting-edge cloud-native technologies
  • Contribute to open-source projects
  • Collaborative, distributed team environment
  • Opportunity to shape the future of Kubernetes tooling
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Circle Internet Financial, LLC • United States

On-site
USD 140,000 - 190,000
Principal Security Engineer
Principal Security Engineer

Mass Digital Health • Boston (MA)

On-site
USD 120,000 - 150,000
Senior DevOps Architect for Cloud-Native CI/CD & Kubernetes
Senior DevOps Architect for Cloud-Native CI/CD & Kubernetes

Veriipro • Westlake (TX)

On-site
Site Reliability Engineer
Site Reliability Engineer

NextGen | GTA: A Kelly Telecom Company • Mount Laurel Township (NJ)

On-site
Lead Site Reliability Engineer - Infrastructure & DevOps
Lead Site Reliability Engineer - Infrastructure & DevOps

SRI Tech Solutions Inc. • Orlando (FL)

On-site
USD 140,000 - 190,000
DevOps Engineer, Openshift
DevOps Engineer, Openshift

Jobtailor • United States

On-site
USD 120,000 - 180,000
Kubernetes Engineer
Kubernetes Engineer

Veriipro • Phoenix (AZ)

On-site
USD 120,000 - 170,000
Health insurance
401(k) plan
Paid time off
Senior Site Reliability Engineer Kubernetes Platform : 26-02283
Senior Site Reliability Engineer Kubernetes Platform : 26-02283

Akraya, Inc. • California (MO)

Hybrid
USD 76,000 - 83,000
Senior Developer Operations Engineer
Senior Developer Operations Engineer

BMA Group Global • San Juan (PR)

On-site
USD 120,000 - 180,000
Lead Kubernetes SRE
Lead Kubernetes SRE

TechDigital Group • Minneapolis (MN)

On-site
USD 80,000 - 120,000