Role Summary
This role adds senior DevOps capacity dedicated to the Analytics platform. The engineer sits within the DevOps team inheriting our engineering standards, change control, security posture, and on-call model — while working day-to-day as an embedded partner to Analytics on their Kubernetes workloads, pipelines, and platform needs.
The intent is simple: Analytics gets dedicated, senior engineering capacity for their platform without inheriting the operational burden of running it. Patching, CVE remediation, secrets handling, ingress, change control, backups, and upgrades remain a DevOps responsibility, executed to a single firm-wide standard.
The India location extends platform coverage beyond US business hours — materially improving response time on business-critical services, including the RMR Dispatcher supporting the trading desk.
Key Responsibilities
- Serve as the dedicated DevOps engineer for Analytics workloads: Kubernetes (Rancher), pipelines, deployment, and platform reliability
- Own and improve reliability of Analytics-supporting services — including the RMR Dispatcher — with observability, alerting, and structured incident response
- Build and maintain CI/CD pipelines (Bitbucket Pipelines, Jenkins) for Analytics applications
- Manage infrastructure as code (Terraform / Terragrunt) and Kubernetes configuration (Helm, Kustomize) through version-controlled, peer-reviewed change
- Operate and troubleshoot RabbitMQ and Redis in support of Analytics workloads
- Build and maintain observability — Grafana / Prometheus dashboards and alerting — across Analytics-supporting infrastructure
- Provide senior escalation coverage during India hours, extending our response window on business-critical services
- Partner directly with Analytics engineers on capacity, resource isolation, and workload placement within the shared Rancher platform
- Document runbooks and cross-train with the US DevOps team, reducing key-person dependency on both sides
Required Skills & Experience
- 7+ years in DevOps, SRE, or platform engineering, with senior-level autonomy
- Kubernetes — production experience: workload placement, resource quotas, affinity/taints, troubleshooting, RBAC. Rancher experience a strong plus
- Linux (RHEL) — strong administration and troubleshooting
- Infrastructure as code — Terraform / Terragrunt in production
- CI/CD — designing and maintaining pipelines (Bitbucket Pipelines, Jenkins, or equivalent)
- Helm / Kustomize — packaging and configuration management
- Observability — Grafana / Prometheus: instrumentation, dashboards, alert design
- Messaging / caching — RabbitMQ and Redis operational experience
- Clear written communication; able to document, elevate, and hand off across time zones
Preferred / "Plus" Skills
- Experience supporting analytics, quant, or research platforms
- VMware / virtualized infrastructure exposure
- Secrets management and identity integration (CyberArk, Active Directory, SSSD)
- Prior experience as an embedded or dedicated platform engineer for a business-facing team
- Financial services experience; familiarity with regulated change-control environments