Career Category
Information Systems
Job Description
ABOUT AMGEN
At Amgen, our mission to serve patients drives everything we do. We discover, develop, manufacture, and deliver innovative medicines, and we use technology, data, and collaboration to help teams across the enterprise operate with greater speed, quality, and impact.
Role Description
The Sr Associate Software Engineer - Cloud Platform & DevOps will join a Global Supply Chain Data & Analytics team that supports a portfolio of cloud-native technology solutions used by supply chain strategy, network, planning, and operations functions. The team builds and sustains enterprise applications and services that support data-driven operational and strategic decisions. In this role, you will independently own defined platform and production-reliability components across cloud infrastructure, Kubernetes, CI/CD, release automation, observability, identity, security, resilience, and operational support. You will work closely with application, data, architecture, quality, cybersecurity, and enterprise cloud teams to ensure solutions are secure, scalable, supportable, and production ready.
Roles & Responsibilities
- Operate and support cloud-native application environments on AWS, with a focus on Amazon EKS and Kubernetes-based workloads.
- Manage Kubernetes deployments, pods, services, ingress, namespaces, configuration, secrets, health checks, autoscaling, resource limits, and environment-specific settings.
- Build, maintain, and improve CI/CD pipelines for source validation, automated testing, security scanning, container-image creation, deployment, environment promotion, and rollback.
- Manage Docker images and container registries such as Amazon ECR, including image lifecycle, vulnerability remediation, and release traceability.
- Automate infrastructure and application configuration using Terraform, CloudFormation, Helm, or comparable infrastructure-as-code and configuration-management tools.
- Support development, test, staging, and production environments and maintain controlled, repeatable promotion paths between them.
- Implement and maintain application and platform observability, including logs, metrics, traces, dashboards, alerting, health indicators, and operational telemetry.
- Monitor availability, latency, errors, pod and node health, API performance, database connectivity, resource consumption, capacity, and cloud cost.
- Manage IAM roles, service identities, role-based access, secrets, credentials, certificates, network controls, and least-privilege access in alignment with enterprise standards.
- Coordinate vulnerability remediation, platform and dependency upgrades, patching, certificate and secret rotation, and resolution of security-scan findings.
- Lead incident response, service restoration, root-cause analysis, corrective actions, and continuous reliability improvements for platform-related issues.
- Define and test backup, restore, disaster-recovery, high-availability, and rollback procedures based on system criticality and enterprise requirements.
- Support integration and connectivity between cloud-hosted applications, databases, enterprise data platforms, APIs, and approved AI services.
- Partner with software engineers to diagnose production issues across React applications, backend services, APIs, databases, and distributed cloud components.
- Create and maintain architecture documentation, deployment procedures, support runbooks, operational standards, and knowledge-transfer materials.
- Support production releases, hyper care, and an appropriate release or production-support rotation in collaboration with global teams.
Basic Qualifications and Experience
Bachelor's or Master's degree in a relevant field (e.g., Computer Science, Engineering or equivalent). 5-9 years of work experience in cloud engineering, platform engineering, DevOps, SRE, or software engineering, subject to final HR leveling. Experience in the global pharmaceutical industry preferred.
Functional Skills
- Cloud platform engineering: independent operation and improvement of secure, scalable, production cloud environments and containerized workloads.
- Delivery automation: CI/CD, infrastructure as code, environment management, release orchestration, testing, security scanning, and rollback.
- Reliability and observability: service health, metrics, logs, traces, alerting, incident response, root-cause analysis, capacity, and resilience.
- Security and governance: IAM, service identities, least privilege, secrets, certificates, vulnerability management, auditability, and enterprise control alignment.
Must-Have Skills
- Strong hands‑on experience with AWS and Kubernetes, preferably Amazon EKS.
- Experience deploying, operating, and troubleshooting containerized applications using Docker and Kubernetes.
- Strong CI/CD experience using GitLab CI/CD, GitHub Actions, Jenkins, Azure DevOps, or comparable tooling.
- Experience with infrastructure as code using Terraform, CloudFormation, or an equivalent framework.