Job summary
A leading graph database company is searching for a Software Engineer in Site Reliability Engineering. The role emphasizes automation for reliability across numerous instances and collaboration with teams on SRE principles. Candidates should have deep experience in backend development, preferably in Go, and be adept at troubleshooting large-scale systems. Familiarity with cloud-based environments and observability tools is essential. This position offers an opportunity to contribute to meaningful improvements in incident response and system reliability.
Backend tools and automation in Go
SRE practices
Collaboration with teams
Troubleshooting distributed systems
Monitoring systems performance
Designing for reliability
Using observability tools
Managing applications on Kubernetes
Infrastructure management with Kustomize and Terraform
CI/CD workflows with GitHub Actions