Описание:
B. Braun develops and manufactures medical devices, pharmaceuticals, and solutions for healthcare professionals and patients. Its Group Digital Department is building the digital foundation for clinical products and healthcare platforms, while the Center of Competence Platform Engineering & DevOps builds and operates an Internal Developer Platform for product engineering teams.
Задачи:
- Own and evolve the product’s CI/CD infrastructure, transitioning from Jenkins pipelines to GitHub Actions while maintaining continuity
- Introduce production-grade observability using OpenTelemetry and the centralized monitoring stack
- Manage the product’s Azure infrastructure across a multi-tenant, per-customer deployment model
- Execute a controlled, coexistence-based development toolchain migration aligned with DevSecOps standards
- Define and monitor SLOs for critical product services and contribute reusable patterns to the Internal Developer Platform
- Operate and maintain Jenkins pipelines for build, test, and deployment workflows
- Design and implement GitHub Actions pipelines with full traceability and version control
- Manage parallel Jenkins and GitHub Actions toolchains during the transition
- Integrate code quality, vulnerability scanning, and SBOM management into pipelines
- Maintain traceability and versioning of pipeline changes within the GitOps model
- Contribute reusable GitHub Actions workflow templates and composite actions to the Internal Developer Platform
- Cloud Manage and operate the product’s Azure environment and multi-tenant deployment model
- Manage Azure-hosted Windows infrastructure, including IIS and Windows Server, alongside cloud-native components
- Implement and maintain infrastructure automation using Crossplane XRDs, Helm, and Azure-native tooling
- Support provisioning of development, test, staging, and production environments in line with the product’s DTAP model
- Troubleshoot and resolve production infrastructure, deployment, and runtime issues
- Deploy and operate AKS workloads, including deployments, services, ingress, namespace RBAC, and day-to-day troubleshooting
- Deploy OpenTelemetry agent instrumentation alongside the product runtime, collect logs, metrics, and traces without modifying application source code, and forward telemetry to the centralized observability stack
- Build and maintain Grafana dashboards and Azure Monitor alert rules covering the four golden signals for critical services
- Define SLIs from collected telemetry and translate them into SLOs and error budgets with the product team
- Identify and address observability gaps
- Lead the GitHub migration, preserving repository history and metadata, coordinating Quality Management sign-off, and executing a controlled cutover
- Align branching strategies, environment models, and release processes with DevSecOps standards while maintaining existing release cycles
- Migrate documentation tooling while preserving traceability across requirements, code, and deployments
- Assess and document the integration engine topology, its connections to hospital systems, its dependencies, and potential migration paths
- Establish SLOs and SLIs for critical services, track error budgets, and provide data-driven input for prioritisation
- Create and maintain runbooks for common operational scenarios to support incident response and autonomous product team operations
- Participate in incident response and facilitate blameless postmortems for platform-level issues
- Contribute patterns, templates, and learnings from the product team assignment to the Internal Developer Platform
- Participate in CoC design reviews, code reviews, and technical standard-setting
- Share operational experience with CoC colleagues to improve shared platform capabilities
Требования:
- 5+ Years in DevOps, Platform Engineering, or a related discipline, with the ability to operate independently
- Production-grade Microsoft Azure operations across App Services, SQL, Key Vault, Azure Monitor, Entra ID, AKS, and Virtual Networks
- Hands-on deployment, operation, and troubleshooting of production Kubernetes workloads on AKS
- Experience with Jenkins and GitHub Actions, including operating the current toolchain while building its replacement
- Experience planning and executing controlled repository migrations while preserving history and traceability
- Experience with agent-based OpenTelemetry instrumentation on running systems, telemetry pipeline configuration, and log-to-OTel mapping
- SRE fundamentals, including defining SLOs and SLIs from scratch, error budget awareness, and reliability monitoring
- Fluent English
- Будет плюсом: IIS, .NET Framework and Windows Server experience; parallel legacy and modern toolchain migrations without service disruption; integration engine topology mapping and documentation; Spanish for local collaboration in Barcelona; experience in regulated or compliance-aware environments such as healthcare, pharma, or finance; ArgoCD, Crossplane, SBOM management, software composition analysis, JFrog Artifactory, EU MDR, ISO 13485, and IEC 62304 awareness
Условия:
- Flexible hybrid schedule: 3 days remote and 2 days on-site per week
- Adaptable working hours Monday to Thursday, with flexible start and finish times
- Reduced working hours on Fridays during summer (June to September)
- Compensation aligned with the Chemical Industry Collective Agreement
- 10% Annual variable bonus linked to company performance
- Flexible compensation plan including transport, childcare vouchers, and private health insurance
- Daily meal allowance for office days
- Work-from-home allowance
- Company-sponsored retirement plan
- Life and accident insurance (AXA)
- Company-funded training and certifications based on role and project needs
- Well-being platform with coaching, psychological support, and emotional health resources, free for employees and available at reduced cost for family members
- Full IT setup provided, with a choice between Windows or Mac
- Competitive salary