- Build, maintain, and evolve foundational platforms supporting product development teams
- Design and implement platform solutions using Infrastructure as Code tools such as Terraform and Ansible
- Collaborate with development teams to ensure application performance, scalability, and reliability
- Design, implement, and manage AWS cloud infrastructure, including EKS, EC2, RDS, IAM, VPC networking, CloudWatch, and CloudTrail
- Implement logging, monitoring, and alerting with Prometheus, Grafana, Loki, Alertmanager, CloudWatch, and Log Analytics
- Design, build, and maintain CI/CD pipelines using GitHub Actions and manage SDLC documentation
- Design, implement, and maintain containerized solutions using Kubernetes and Docker, including EKS deployments
- Contribute to architectural discussions and technical decisions for the DXO Platform
- Develop automation and self-service capabilities using Ansible, Crossplane, and custom scripts
- Provide production support to ensure platform availability, reliability, and performance
- Implement security best practices and maintain compliance with organizational policies and industry standards
- Automate testing procedures and document defects to meet accessibility standards
Requirements
- High school diploma required
- 5+ years of experience in Platform Engineering, DevOps, Infrastructure Engineering, or Site Reliability Engineering roles
- Hands-on experience designing, building, and operating production systems on AWS
- Familiarity with Azure
- Deep proficiency in Linux
- Strong scripting skills in Python or similar languages
- Extensive experience with CI/CD methodologies and GitHub Actions or GitLab CI
- Experience with Ansible
- Understanding of containerization technologies such as Podman and Docker
- Understanding of container orchestration such as Kubernetes and OpenShift
- Experience with infrastructure-as-code principles and tools such as Terraform
- Experience with Prometheus, Grafana, ELK stack, CloudWatch, and Log Analytics
- Familiarity with Git, GitOps, and semantic versioning
- Familiarity with artifact management using JFrog Artifactory or Nexus
- Understanding of network protocols and security best practices
- Excellent communication skills for mentoring engineers, documentation, and technical presentations
- Certifications in AWS, Red Hat Ansible Automation Platform, or Kubernetes preferred
- Experience with JFrog Artifactory and JFrog Xray administration preferred
- Knowledge of OpenShift Operators and Kubernetes CRDs preferred
- Experience integrating ServiceNow or other ITSM platforms into CI/CD workflows preferred
- Working knowledge of Backstage preferred
- Terraform, Packer, or Crossplane expertise preferred
Core Competencies
Demonstrates expertise in designing and implementing AWS cloud infrastructure, utilizing Infrastructure as Code tools like Terraform and Ansible, and managing CI/CD pipelines with GitHub Actions. Proficient in containerization and orchestration technologies, ensuring platform reliability and compliance with security best practices.
Highest-signal resume keywords
- AWS Cloud Infrastructure Management
- Infrastructure As Code (Terraform, Ansible)
- CI/CD Pipeline Development (GitHub Actions)
- Containerization (Docker, Kubernetes)
- Production System Support
ATS Optimization Keywords
Hard Skills
- Platform Engineering
- DevOps
- Infrastructure Engineering
- Site Reliability Engineering
- Linux Proficiency
- Scripting (Python)
- CI/CD Methodologies
- Container Orchestration (Kubernetes, OpenShift)
- Monitoring Tools (Prometheus, Grafana)
- Network Protocols
Soft Skills
- Excellent Communication Skills
Certifications & Qualifications
- AWS Certification
- Red Hat Ansible Automation Platform
- Kubernetes Certification
Industry Keywords
- Infrastructure As Code
- CI/CD Workflows
- Security Best Practices
- Accessibility Standards
- ITSM Integration
Tools & Technologies
- AWS (EKS, EC2, RDS, IAM, VPC)
- GitHub Actions
- Docker
- Kubernetes
- Prometheus
- Grafana
- CloudWatch
- Log Analytics
- JFrog Artifactory
- Crossplane