Senior Observability Engineer

Initial Therapeutics, Inc.

Poland

On-site

PLN 180,000 - 360,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Healthcare
Well-being resources
Family building benefits
Generous paid time off
Global recharge days

Job summary

Moderna’s Warsaw hub seeks an experienced Senior Platform Engineer to own the enterprise observability platform, drive governance, cost optimization, and scalable architectures across applications and cloud services.

Lead with OpenTelemetry, Prometheus, Grafana, and AI‑enabled monitoring; build pipelines for metrics, traces, logs; partner with security and incident teams.

You will mentor teams, develop self‑service observability, and shape strategic roadmaps while upholding enterprise standards.

Qualifications

  • 7+ years in site reliability engineering, observability engineering, or platform engineering.
  • Hands-on experience designing and operating modern observability platforms.
  • Experience with OpenTelemetry, Prometheus, Grafana, and similar tools.
  • Strong telemetry knowledge across metrics, logs, traces, and SLO/SLI.
  • Experience in cloud and hybrid environments; AWS or Azure.
  • Automation and IaC with Python, Terraform, Ansible, or Bash.
  • Excellent communication and stakeholder management.

Responsibilities

  • Own the enterprise observability platform, driving cost optimization, capacity planning, telemetry governance, and sustainable platform growth.
  • Manage and evolve Moderna's observability platform using OpenTelemetry, Prometheus, Grafana, VictoriaMetrics, and other modern observability solutions.
  • Lead platform governance, agent lifecycle management, architecture standards, roadmap development, and platform best practices.
  • Collaborate with technology vendors and open-source communities to influence product roadmaps and maximize platform value.
  • Design and build scalable, resilient, and cost‑efficient observability architectures across apps, databases, hosts, containers, cloud services, networks, and AI/LLM workloads.
  • Develop and optimize telemetry pipelines for metrics, traces, and logs across hybrid cloud and on‑prem environments.
  • Establish enterprise standards for monitoring, alerting, SLOs, SLIs, and proactive incident detection.
  • Enable self‑service observability capabilities that accelerate troubleshooting and platform reliability.
  • Design and implement observability for AI agents and LLM-based workloads, including prompts, latency, errors, and costs.
  • Define enterprise standards for AI observability and model performance monitoring.
  • Implement AI observability solutions using Langfuse or similar technologies for analytics and alerting.
  • Leverage AI/LMM technologies for anomaly detection and operational insights.
  • Lead enterprise logging as a core pillar with Grafana Loki and open‑source tech.
  • Define logging standards: ingestion, parsing, retention, and query performance.
  • Partner with Security to ensure compliance and forensic capabilities.
  • Integrate observability with incident platforms like PagerDuty for better response.
  • Optimize on‑call processes and ensure actionable alerts.
  • Provide real‑time telemetry during incidents and contribute to root cause analysis.
  • Develop automation with Python, Terraform, Ansible, CI/CD, and IaC practices.
  • Implement self‑healing capabilities and automated remediation.
  • Integrate with ServiceNow, Jira, and other enterprise systems.
  • Develop dashboards and executive reports on adoption, MTTA/MTTR, and cost efficiency.
  • Produce docs, runbooks, and training to support adoption and consistency.
  • Participate in post‑incident reviews and drive continuous improvement.

Skills

SRE / Observability
Platform Engineering
Automation
Python
Terraform
CI/CD
Communication

Tools

OpenTelemetry
Prometheus
Grafana
VictoriaMetrics
Elastic
Datadog
Dynatrace

Job description

If you’re interested in this role, please apply in English and include an English version of your Resume/CV.

The Role:

Joining Moderna means advancing mRNA science to transform medicine. Work with exceptional global teams on a broad pipeline and build a career that makes a real difference for patients.

Moderna is strengthening its international business services hub in Warsaw, supporting our growing global operations. We welcome professionals ready to help advance our mission and shape the future of mRNA medicines.

This is an opportunity to define the future of enterprise observability by building a modern, AI‑enabled observability platform that delivers scalable, resilient, and cost‑efficient monitoring across Moderna's technology landscape. As a senior individual contributor, you will lead platform strategy, governance, and engineering while enabling intelligent operational insights across applications, infrastructure, cloud services, networks, containers, databases, and AI‑powered systems. Working with open standards, automation, and emerging AI technologies, you will help improve reliability, operational excellence, and business outcomes across the region.

Here’s What You’ll Do
  • Own the enterprise observability platform, driving cost optimization, capacity planning, telemetry governance, and sustainable platform growth.
  • Manage and evolve Moderna's observability platform using technologies including OpenTelemetry, Prometheus, Grafana, VictoriaMetrics, and other modern observability solutions.
  • Lead platform governance, agent lifecycle management, architecture standards, roadmap development, and platform best practices.
  • Collaborate with technology vendors and open-source communities to influence product roadmaps and maximize platform value.
  • Design and build scalable, resilient, and cost‑efficient observability architectures supporting applications, databases, hosts, containers, cloud services, networks, distributed systems, and AI/LLM‑based workloads.
  • Develop and optimize telemetry pipelines for metrics, traces, and logs across hybrid cloud and on‑premises environments.
  • Establish enterprise standards for monitoring, alerting, Service Level Objectives (SLOs), Service Level Indicators (SLIs), and proactive incident detection.
  • Enable self‑service observability capabilities that accelerate troubleshooting, operational visibility, and platform reliability.
  • Design and implement observability capabilities for AI agents, LLM‑powered applications, and agentic workflows, including monitoring prompts, responses, execution flows, latency, errors, taken consumption, and operational costs.
  • Define enterprise standards for AI observability, monitoring model performance, user interactions, reliability, and cost efficiency.
  • Implement AI observability solutions using Langfuse or similar technologies, enabling prompt analytics, optimization, intelligent alerting, debugging, and failure pattern detection.
  • Leverage AI and LLM technologies to support anomaly detection, operational insights, and root cause analysis.
  • Lead the enterprise logging strategy as a core pillar of the observability platform.
  • Design, build, and scale cost‑efficient logging solutions using Grafana Loki and other modern open‑source technologies.
  • Define enterprise logging standards, including centralized log ingestion, parsing, querying, retention policies, storage optimization, and query performance.
  • Partner with Security teams to ensure compliance, audit readiness, and forensic capabilities.
  • Integrate observability capabilities with incident management platforms such as PagerDuty to improve operational responsiveness.
  • Optimize on‑call processes by ensuring alerts are meaningful, actionable, and routed effectively while supporting rapid incident resolution.
  • Provide real‑time telemetry during incidents and contribute to root cause analysis activities.
  • Develop automation using Python, Terraform, Ansible, CI/CD pipelines, and infrastructure‑as‑code practices.
  • Implement self‑healing capabilities and automated remediation to improve operational resilience.
  • Integrate the observability platform with enterprise technologies including ServiceNow, Jira, and other operational systems.
  • Develop dashboards and executive reporting that provide visibility into platform adoption, telemetry coverage, MTTA, MTTR, alert quality, reliability, operational performance, and cost efficiency.
  • Produce documentation, runbooks, knowledge articles, and technical training to support platform adoption and engineering consistency.
  • Participate in post‑incident reviews and drive continuous improvement initiatives that strengthen operational excellence and foster a culture of observability and data‑driven decision‑making.
The key Moderna Mindsets you'll need to succeed in the role:

We act with dynamic range, driving strategy and execution at the same time at every step.

We obsess over learning. We don’t have to be the smartest, we have to learn the fastest.

Here’s What You’ll Need (Minimum Qualifications)
  • 7+ years of experience in site reliability engineering, observability engineering, platform engineering, or related technical disciplines.
  • Strong hands‑on experience designing, implementing, and operating modern observability platforms.
  • Experience with observability technologies such as OpenTelemetry, Prometheus, Grafana, VictoriaMetrics, Elastic, Datadog, Dynatrace, or similar solutions.
  • Strong understanding of metrics, logs, traces, telemetry pipelines, and SLO/SLI frameworks.
  • Experience supporting applications, infrastructure, containers, cloud services, and distributed systems.
  • Experience integrating observability platforms with incident management and operational workflows.
  • Hands‑on experience with automation and infrastructure‑as‑code technologies such as Python, Terraform, Ansible, or Bash.
  • Experience supporting cloud‑native and hybrid environments (AWS and/or Azure).
  • Strong analytical, troubleshooting, and problem‑solving skills.
  • Strong communication and stakeholder management skills.
Here’s What You’ll Bring to the Table (Preferred Qualifications)
  • Experience working in biotech, pharmaceutical, healthcare, or other regulated environments (e.g., GxP, HIPAA).
  • Experience with AI observability tools such as Langfuse, Arize, Phoenix, LangSmith, or similar platforms.
  • Experience monitoring AI agents, LLM applications, retrieval systems, or agentic workflows.
  • Experience with enterprise logging platforms and large‑scale log management strategies.
  • Experience integrating observability platforms with PagerDuty, ServiceNow, Jira, or similar operational platforms.
  • Relevant certifications in AWS, Azure, Kubernetes, observability, or related technologies.
Pay & Benefits

At Moderna, we believe that when you feel your best, you can do your best work. That’s why our benefits and well‑being resources are designed to support you—at work, at home, and everywhere in between.

  • Competitive healthcare, plus voluntary benefit programs to support your unique needs
  • A holistic approach to well‑being with access to fitness, mindfulness, and mental health support
  • Family building benefits, including fertility, adoption, and surrogacy support
  • Generous paid time off, including vacation, bank holidays, volunteer days, sabbatical, global recharge days, and a discretionary year‑end shutdown
  • Savings and investments to help you plan for the future
  • Location‑specific perks and extras

The benefits offered may vary depending on the nature of your employment with Moderna and the country where you work.

About Moderna

Since our founding in 2010, we have aspired to build the leading mRNA technology platform, the infrastructure to reimagine how medicines are created and delivered, and a world‑class team. We believe in giving our people a platform to change medicine and an opportunity to change the world.

By living our mission, values, and mindsets every day, our people are the driving force behind our scientific progress and our culture. Together, we are creating a culture of belonging and building an organization that cares deeply for our patients, our employees, the environment, and our communities.

We are proud to have been recognized as a Science Magazine Top Biopharma Employer, a Fast Company Best Workplace for Innovators, and a Great Place to Work in the U.S.

As we build our company, we have always believed an in‑person culture is critical to our success. Moderna champions the significant benefits of in‑office collaboration by embracing a 70/30 work model. This 70% in‑office structure helps to foster a culture rich in innovation, teamwork, and direct mentorship. Join us in shaping a world where every interaction is an opportunity to learn, contribute, and make a meaningful impact.

Moderna is a smoke‑free, alcohol‑free, and drug‑free work environment.

Moderna is a place where everyone can grow. If you meet the Basic Qualifications for the role and you would be excited to contribute to our mission every day, please apply!

Moderna is committed to equal opportunity in employment and non‑discrimination for all employees and qualified applicants without regard to a person's race, color, sex, gender identity or expression, age, religion, national origin, ancestry or citizenship, ethnicity, disability, military or protected veteran status, genetic information, sexual orientation, marital or familial status, or any other personal characteristic protected under applicable law. We consider qualified applicants regardless of criminal histories, consistent with legal requirements.

We’re focused on attracting, retaining, developing, and advancing our employees. By cultivating a workplace that values diverse experiences, backgrounds, and ideas, we create an environment where every employee can contribute their best.

Moderna is committed to offering reasonable accommodation or adjustments to qualified job applicants with disabilities. Any applicant requiring an accommodation or adjustment in connection with the hiring process and/or to perform the essential functions of the position for which the applicant has applied should contact the Accommodations and Adjustments team at leavesandaccommodations@modernatx.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Observability Engineer
Senior Observability Engineer

Moderna • Warszawa

Hybrid
PLN 300,000 - 520,000
Healthcare
Well-being programs
Family building benefits
+3
Senior Observability Engineer
Senior Observability Engineer

BioSpace • Warszawa

On-site
PLN 240,000 - 420,000
Competitive healthcare
Global learning & development
Flexible time off
Software Engineer
Software Engineer

Initial Therapeutics, Inc. • Poland

On-site
PLN 130,000 - 170,000
Best-in-class healthcare
Generous paid time off
Family building benefits
Software Engineer, DevSecOps
Software Engineer, DevSecOps

Moderna • Warszawa

On-site
PLN 240,000 - 320,000
Competitive healthcare
Well-being resources
Family building benefits
+2
Software Engineer, DevSecOps
Software Engineer, DevSecOps

Initial Therapeutics, Inc. • Poland

On-site
PLN 180,000 - 320,000
Healthcare
Well‑being programs
Generous paid time off
Senior Software Engineer, Developer Platform Engineering
Senior Software Engineer, Developer Platform Engineering

Initial Therapeutics, Inc. • Poland

On-site
PLN 180,000 - 240,000
Healthcare benefits
Well-being resources
Family building benefits
+3
Principal Data Platform Engineer
Principal Data Platform Engineer

BioSpace • Warszawa

On-site
PLN 260,000 - 420,000
Healthcare
Well-being resources
Paid time off
+1
Associate Director, Regulatory Data Management & Intelligence
Associate Director, Regulatory Data Management & Intelligence

Initial Therapeutics, Inc. • Poland

Hybrid
PLN 400,000 - 700,000
Comprehensive benefits
Hybrid work model
Global career opportunities
Power Platform & AI Engineer
Power Platform & AI Engineer

Initial Therapeutics, Inc. • Poland

On-site
PLN 180,000 - 280,000
Competitive healthcare
Well‑being resources
Family building benefits
+2
Power Platform & AI Engineer
Power Platform & AI Engineer

BioSpace • Warszawa

On-site
PLN 180,000 - 240,000
Healthcare benefits
Well-being resources
Paid time off