Get more replies from employers
Send a job-specific resume in minutes.
AgileEngine is seeking a Senior Site Reliability Engineer to ensure core system administration and operational stability across on-premise and SaaS-hosted environments. The role emphasizes Kubernetes cluster management, monitoring, and observability with Snowflake and OpenTelemetry.
You will participate in on-call rotations, incident response, and root-cause analysis, and automate infrastructure tasks using Python, Bash, or Go to keep systems healthy across multiple platforms.
We are looking for a Senior Site Reliability Engineer to provide core system administration and operational stability for enterprise on-premise and SaaS-hosted systems, with a strong focus on Kubernetes cluster management, monitoring, and observability using Snowflake and OpenTelemetry. You will participate in on-call rotations, incident response, and root-cause analysis, automate infrastructure tasks using Python, Bash, or Go, and ensure system health across ESM and ECP platform environments. Experience with service mesh architectures is highly valued for ECP-focused roles.
What you will do
Must haves
Nice to haves
Perks and Benefits
Accelerate your professional journey with mentorship, TechTalks, and personalized growth roadmaps
We match your ever-growing skills, talent, and contributions with competitive USD-based compensation and budgets for education, fitness, and team activities
Join projects with modern solutions development and top-tier clients that include Fortune 500 enterprises and leading product brands
Tailor your schedule for an optimal work-life balance, by having the options of working from home and going to the office – whatever makes you the happiest and most productive.