Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Artech Infosystems Private Limited seeks an experienced DevOps/SRE engineer with 9+ years of hands-on experience. You will manage Kubernetes operations at scale, architect multi-cloud infrastructure, and build automation for cross-team use.
The role requires strong skills in Python/Go/Bash, cloud architecture decisions, and advanced debugging of distributed systems. You will mentor engineers and influence technical direction across teams.
9 + years of DevOps, SRE, or Infrastructure Engineering experience
Deep, hands-on Kubernetes operations expertise — cluster upgrades, node group management, and complex troubleshooting across large-scale, multi-cluster environments
Strong programming/ ing skills (Python, Go, Bash, or similar) — proven track record building, scaling, and maintaining automation and internal tooling used by multiple teams
Advanced skills in reading and interpreting logs, metrics, and system state to diagnose complex, cross-system infrastructure issues
Extensive experience with cloud infrastructure (AWS, GCP, or Azure), including architecture decisions and cost/performance tradeoffs
Expert-level debugging skills across distributed systems — able to trace failures from symptom through to root cause in highly complex environments
Excellent written and verbal communication — able to produce clear status updates, bug reports, technical documentation, and influence technical direction across teams
Demonstrated experience mentoring engineers and leading technical initiatives
Preferred
Deep experience with infrastructure-as-code tools (Terraform, Helm, Ansible), including designing reusable modules/patterns for org-wide use
Strong familiarity with CI/CD systems (Jenkins, GitHub Actions, Spinnaker, or similar), including pipeline architecture Hands-on experience with GitOps workflows (ArgoCD, Flux) at scale
Advanced experience with observability stacks (Prometheus, Grafana, Datadog), including designing alerting/SLO frameworks
Proven experience operating Kubernetes at scale across multiple clusters, regions, or environments, including capacity planning and disaster recovery
Experience contributing to or leading architectural decisions for infrastructure platforms