Experience: 3.00+ years
Salary: INR 2000000-2500000 / year (based on experience)
Expected Notice Period: 15 days
Shift: GMT+05:30, Asia/Kolkata (IST)
Opportunity Type: Remote
Placement Type: Full Time Permanent position (Payroll and Compliance to be managed by Winmore)
Note: This is a requirement for one of Uplers' client - Winmore
Must have skills required:
- AI Tools, AI/ML understanding, Security, Kubernetes, production, AWS, CI/CD, Infrastructure as Code (IaC), Scripting, Observability
Winmore is looking for:
Own and improve the reliability, security, and delivery speed of cloud-native services on AWS. You’ll run production systems across Kubernetes/ECS, build CI/CD, automate infrastructure, and drive incident response. We prefer candidates comfortable in fast-moving, startup-style environments.
Key Responsibilities
- Operate and scale production Kubernetes (EKS) and ECS workloads and AWS infrastructure (EC2, S3, RDS, VPC, IAM, Lambda).
- Manage PostgreSQL operations: access controls, backups/restore, replication, performance tuning, and upgrades.
- Build and maintain CI/CD pipelines using Jenkins or CircleCI (safe releases, rollbacks, environment promotion).
- Automate infrastructure and operational workflows using Terraform / CloudFormation plus Python/Bash.
- Implement monitoring/alerting using Datadog, CloudWatch, Grafana, SigNoz; troubleshoot incidents and lead RCA/postmortems.
- Enforce security best practices across infra and delivery: least privilege IAM, secrets management, private networking, vulnerability management, compliance-aligned controls.
Requirements (Must Have)
- 3–5 years relevant DevOps/SRE/Platform experience.
- Production exposure (required): hands‑on operating and supporting live systems, on‑call/incident response, debugging under pressure.
- Kubernetes + ECS: deploy/operate containerized apps on EKS and ECS; Helm charts; resource tuning; production debugging.
- AWS: EC2, S3, RDS, IAM, VPC, Lambda; strong AWS CLI automation. PostgreSQL: admin fundamentals (roles/permissions, backups, replication, query optimisation).
- CI/CD: Jenkins or CircleCI; artifact/versioning; deployment strategies and rollback.
- IaC + Scripting: Terraform/CloudFormation; Python/Bash automation. Observability: metrics/logs/traces; alert hygiene; operational troubleshooting.
- Security exposure (required): IAM least privilege, secrets handling, secure networking, vulnerability management basics, working with compliance/security requirements.
Preferred
- Experience in startup/dynamic teams: high ownership, ambiguity, fast iteration, pragmatic trade-offs.
- MLOps / LLM flows: model serving patterns, vector DB basics, prompt/guardrail concepts, CI/CD for ML artifacts.
- SOC2-style operational controls / audit readiness exposure.
Qualifications
- Bachelor’s degree in CS/Engineering (preferred).
Soft Skills
- Strong problem‑solving and clear communication during incidents.
- Ownership mindset; proactive learner; able to operate independently.
- Collaborative across engineering, product, and security stakeholders.