- Support infrastructure reliability and uptime across development, staging, and production environments
- Design and implement scalable, resilient AWS systems supporting customer and device growth
- Help manage multi-regional or highly available infrastructure with clear SLAs, failover patterns, and capacity planning
- Identify and remediate infrastructure and deployment-pipeline security risks
- Build and maintain CI/CD systems and GitOps workflows
- Improve developer tooling, deployment pipelines, and environment management
- Manage Kubernetes and containerized workloads using infrastructure-as-code practices
- Build and enhance monitoring, logging, tracing, and alerting practices
- Participate in on-call rotations and incident response, including root cause analysis and post-incident improvements
- Contribute to infrastructure health, performance, and cost reporting and dashboards
- Contribute to the infrastructure and platform roadmap
- Build and maintain operational processes, runbooks, and documentation
- Support infrastructure cost management by identifying waste and aligning resource usage with business goals
- Support new product initiatives with scalable platform solutions and early DevOps input into architecture decisions
- Mentor and provide technical guidance to other engineers
- Collaborate with engineering, product, and the Lead DevOps Engineer to translate business needs into technical initiatives and priorities
- Report to the Director of Platform and Infrastructure
Requirements
- 4+ years of experience in DevOps, SRE, or infrastructure engineering
- Demonstrated track record of building or significantly maturing DevOps practices — not just operating within an established one
- Experience operating and scaling systems to millions of users or connected devices
- Strong experience with AWS and cloud-based infrastructure at production scale
- Experience with Kubernetes and containerized workloads in production environments
- Proficiency in infrastructure as code (Terraform, Helm) and CI/CD pipeline design (GitOps workflows preferred)
- Strong command of observability practices: metrics, logging, tracing, and alerting systems
- Experience supporting multi-regional or highly available production systems with defined SLAs
- Experience with infrastructure cost optimization — not just monitoring spend, but actively managing it
- Ability to participate in on-call rotations and incident response
- Ability to collaborate with engineering, product, security, and platform teams
- Strong communication skills for technical and non-technical stakeholders
- Genuine interest in Gabb’s mission
Core Competencies
Demonstrates expertise in building and managing scalable AWS systems, Kubernetes, and CI/CD pipelines while ensuring infrastructure reliability and security. Proficient in observability practices and infrastructure cost optimization, with a strong ability to collaborate across teams and mentor engineers.
Highest-signal resume keywords
- AWS Infrastructure Management
- Kubernetes and Containerization
- Infrastructure as Code (Terraform, Helm)
- CI/CD Pipeline Design (GitOps)
- Observability Practices (Metrics, Logging, Tracing)
ATS Optimization Keywords
Hard Skills
- DevOps Engineering
- Infrastructure Reliability
- Capacity Planning
- Security Risk Remediation
- Monitoring and Alerting
- Incident Response
- Cost Management
- Root Cause Analysis
- Technical Documentation
- Scalable System Design
Soft Skills
- Strong Communication Skills
- Mentoring and Technical Guidance
- Collaboration with Cross-Functional Teams
Industry Keywords
- Infrastructure Engineering
- Site Reliability Engineering (SRE)
- Multi-Regional Systems
- Production Scale Systems
- DevOps Practices
Tools & Technologies
- AWS
- Kubernetes
- Terraform
- Helm
- GitOps
- CI/CD Systems