Job summary
A technology solutions provider seeks a Site Reliability Engineer in Richmond, Virginia. The role involves developing and operating secure cloud infrastructure, ensuring the availability, performance, and compliance of services based on AWS and Kubernetes. The ideal candidate has expertise in container management, automation, operational monitoring, and is familiar with the full software development lifecycle. Strong collaboration with engineering and product teams is essential to enhance the efficiency of cloud services.
Container management technologies (Docker, Kubernetes)
AWS services (EKS, ECS, IAM, S3, RDS)
Automation/configuration management (Terraform)
CI tools (Jenkins)
Operational monitoring tools (Datadog, NewRelic, Splunk)
Linux tools and scripting
Large-scale distributed systems
Software development lifecycle