Site Reliability Engineer

Flanksource Inc.

Mission (KS)

Remote

USD 90,000 - 130,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

100% remote work
Flexible hours
Opportunity to work with cutting-edge technology

Job summary

A technology company is seeking a Kubernetes DevOps Engineer to design, maintain, and optimize Kubernetes clusters. Responsibilities include enhancing the reliability and scalability of cloud infrastructure, managing CI/CD pipelines, and collaborating with development teams. Ideal candidates will have experience in Infrastructure as Code tools and cloud platforms. This position offers 100% remote work with flexible hours.

Qualifications

  • Experience with 2+ Infrastructure as Code tools like Terraform or Pulumi.
  • Proficiency in monitoring tools such as Prometheus and Grafana.
  • Hands-on experience with cloud platforms like AWS, Azure, or GCP.
  • Knowledge of CI/CD processes using GitHub Actions, GitLab CI, or Azure DevOps.
  • Understanding of network fundamentals and security best practices.
  • Strong analytical and troubleshooting abilities.
  • Fluent English for remote asynchronous communication.
  • Ability to work independently in an agile manner.

Responsibilities

  • Design and maintain Kubernetes clusters across multiple environments.
  • Build automation for cluster deployment and management.
  • Monitor and troubleshoot clusters for high availability.
  • Implement security best practices for Kubernetes infrastructure.
  • Participate in incident response to reduce Mean Time To Recovery (MTTR).
  • Enhance reliability and scalability of Kubernetes infrastructure.
  • Manage CI/CD pipelines and DevOps tooling.
  • Collaborate with development teams.

Skills

Infrastructure as Code (IaC) tools
Proficiency with Prometheus
Cloud Platforms (AWS, Azure, GCP)
CI/CD processes
Networking fundamentals
Problem-solving
Fluent English
Self-motivated

Tools

Terraform
Pulumi
Grafana
GitHub Actions
GitLab CI
Azure DevOps

Job description

  • Design and maintain Kubernetes clusters across multiple environments (development, staging, production)
  • Build automation for cluster deployment, configuration, and management
  • Monitor and troubleshoot clusters to ensure high availability and optimal performance
  • Implement security best practices for Kubernetes and underlying infrastructure
  • Participate in incident response and work to reduce Mean Time To Recovery (MTTR)
  • Enhance the reliability and scalability of our Kubernetes infrastructure
  • Manage CI/CD pipelines and DevOps tooling
  • Collaborate with development teams on deployment strategies and best practices
Requirements
  • Infrastructure as Code - Experience with 2+ IaC tools (Terraform, Pulumi, etc.)
  • Monitoring & Observability - Proficiency with Prometheus, Grafana, and related tools
  • Cloud Platforms - Hands-on experience with AWS, Azure, or GCP
  • CI/CD - Knowledge of GitHub Actions, GitLab CI, or Azure DevOps
  • Networking & Security - Understanding of network fundamentals and security best practices
  • Problem-solving - Strong analytical and troubleshooting abilities
  • Communication - Fluent English for remote asynchronous work
  • Self-motivated - Ability to work independently with an agile approach
Nice-to-haves
  • Experience with GitOps tools (Flux, ArgoCD)
  • Go programming knowledge or willingness to learn
  • Active open-source contributions
  • Experience developing Kubernetes operators or controllers
  • 100% remote work with flexible hours
  • Work with cutting-edge cloud-native technologies
  • Contribute to open-source projects
  • Collaborative, distributed team environment
  • Opportunity to shape the future of Kubernetes tooling
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Platform Site Reliability Engineer
Platform Site Reliability Engineer

Specter • San Francisco (CA)

On-site
USD 180,000 - 230,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Clearwater Analytics • Boise (ID)

On-site
USD 130,000 - 170,000
Senior Devops Engineer/lead
Senior Devops Engineer/lead

Scalence • Morristown (NJ)

Hybrid
USD 140,000 - 190,000
Principal Security Engineer
Principal Security Engineer

Mass Digital Health • Boston (MA)

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

NextGen | GTA: A Kelly Telecom Company • Mount Laurel Township (NJ)

On-site
USD 110,000 - 170,000
Site Reliability Engineer
Site Reliability Engineer

Hidden Jobs • United States

On-site
USD 150,000 - 190,000
Healthcare
Retirement matching
Paid family leave
+3
DevOps Platform Engineer
DevOps Platform Engineer

The Jupiter Group, Inc • United States

On-site
USD 120,000 - 160,000
Lead Devops Platform Engineer
Lead Devops Platform Engineer

Techblocks • New York (NY)

Hybrid
USD 140,000 - 200,000
Infrastructure/Cloud DevOps - SRE
Infrastructure/Cloud DevOps - SRE

Bayside Solutions • Cupertino (CA)

On-site
USD 150,000 - 230,000
Senior Kubernetes Engineer
Senior Kubernetes Engineer

Ashley Furniture Industries • Tampa (FL)

On-site
USD 150,000 - 190,000