Staff Observability Engineer (Remote)

Tealium Inc.

Polska

Hybrid

PLN 290,000 - 375,000

Full time

13 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Remote-first working
New-hire equity grants
Stipends for home office
Volunteer time off

Job summary

Tealium Inc. is seeking a Senior or Staff Observability/SRE Engineer to implement a robust reliability strategy across customer-facing products, AI features, platform services, and internal systems.

You will own telemetry pipelines, cost controls, and cross-functional collaboration with SRE, MLOps, data engineering, security, and product teams. You will design AI observability for model usage, latency, and cost attribution, leverage OpenTelemetry and tools like Datadog, Prometheus, and Grafana,

Qualifications

  • 6+ years in SRE, Observability, or Platform Engineering for 24x7 systems.
  • Deep experience with OpenTelemetry and telemetry stacks (Datadog, Prometheus, Grafana).
  • Design observability for distributed APIs and event-driven architectures.
  • Hands-on with AI/ML/GenAI instrumenting prompts, retrieval, tool use, and evaluation.
  • Proficiency with Amazon Bedrock or similar platforms to track model usage, latency, token metrics, and cost attribution.
  • Familiarity with agentic workflows, prompt engineering, vector databases (Neptune), RAG architectures, and tools like LangChain, LlamaIndex, or SageMaker.
  • Proficiency in Java, Python, or Go.
  • Strong AWS expertise, including cloud networking, IAM, security, and quotas.
  • Solid IaC, CI/CD, and container experience (Terraform, Kubernetes, Argo CD, Jenkins, GitHub Actions).
  • Experience with data modeling, telemetry pipelines, retention, and responsible AI.

Responsibilities

  • Lead end-to-end observability design and telemetry pipelines across Tealium products and services.
  • Define SLIs/SEOs, error budgets, and production-readiness requirements.
  • Architect AI observability for Bedrock, tracking model usage, latency, and costs.
  • Establish end-to-end traceability with privacy and security controls.
  • Develop dashboards, alerts, and runbooks to enable rapid incident diagnosis.
  • Define signals for response quality, groundedness, guardrails, and task completion across AI systems.
  • Participate in on-call rotation and drive capacity planning.

Skills

Site Reliability
Observability
Platform Engineering
AI observability
AWS
Java
Python
Go
Cross-functional leadership

Tools

OpenTelemetry
Datadog
Sumo Logic
Prometheus
Grafana
LangChain
LlamaIndex
SageMaker
Neptune
Terraform
Kubernetes
Argo CD
Jenkins
GitHub Actions

Job description

WHO WE ARE

Tealium is the trusted leader in real-time Customer Data Platforms (CDP), helping organizations unify their customer data to deliver more personalized, privacy-conscious experiences. As the demand for connected, intelligent customer engagement grows, Tealium’s leadership in CDP is translating directly into leadership in enabling enterprise AI strategies. By providing clean, consented, and actionable data, Tealium empowers its customers to accelerate the adoption of AI and machine learning, fueling smarter personalization, predictive insights, and business outcomes at scale. More than 800 leading global brands trust Tealium to power their customer data strategies and deliver real-time, personalized experiences at scale.

Team Tealium has team members present in nearly 20 countries worldwide, serving customers across more than 30 countries. We win together with respect and appreciation for the talents required of all positions and the people who contribute to each of these. We are intentional about our WOWs (Ways of Work) culture, our investment in our team members, and how we care and connect. With an extraordinary portfolio of investors (including Georgian, Silver Lake Waterman, Battery, and others) and deep industry experience, Tealium has the financial backing, profitability, and expertise to continue to outpace competitors and lead the way in innovation. Today, Tealium holds over 50 patents, and a few of the recent industry recognitions include: A Leader in the 2025 Gartner® Magic Quadrant™ for Customer Data Platforms 2025, TrustRadius Award Winner: Buyer’s Choice 2024, Invoca Partner Collaboration Award 2024, G2 Leader in Tag Management & Enterprise Data Governance, Tealium Customer Data Hub achieved the Top Rated Award by TrustRadius (2024), Named on Destination CRM’s 2024 Top 100 Technologies List for Sales, Named on the 2024 Best and Brightest in the Nation list, BuiltIn’s 2024 Best Place to Work.

WHAT WE ARE LOOKING FOR

We are seeking a Senior or Staff Observability/SRE Engineer to help implement Tealium’s observability strategy across customer-facing products, AI features, platform services, and internal systems. This role requires a strong ownership mentality, strategic thinking, and data-driven decision-making to establish robust reliability practices, telemetry pipelines, and cost controls across conventional and AI-powered systems. You will focus heavily on AI observability, ensuring model usage, agentic workflows, retrieval systems, and AI-assisted features produce trustworthy, timely, and economically sustainable outcomes. Working cross-functionally with SRE, MLOps, data engineering, security, and product teams, you will leverage modern tools and AI workflows to drive scalable, high-impact reliability solutions across the enterprise.

YOUR DAY TO DAY
  • Lead end-to-end observability design, telemetry schemas, and OpenTelemetry pipeline architectures across Tealium products, platform services, data pipelines, and internal tools.
  • Partner cross-functionally with engineering and product teams to define service-level indicators, objectives, error budgets, and production-readiness requirements, embodying the Win Together and Trusted Outcomes WOWs.
  • Architect comprehensive AI observability solutions for Amazon Bedrock and agentic workflows, tracking model selection, latency, retries, token usage, tool execution, and cost attribution.
  • Establish end-to-end traceability across user requests, prompts, model calls, retrieved context, downstream services, and final responses while maintaining data privacy and security controls.
  • Develop automated dashboards, alerts, SLOs, runbooks, and operational views to drive rapid incident diagnosis, performance optimization, and economic sustainability.
  • Innovate continuously by defining measurable signals for response quality, groundedness, guardrail outcomes, and task completion across AI systems.
  • Participate in on-call rotation (approximately 20% of the time) and drive proactive failure injection, production testing, and capacity planning.
WHAT YOU BRING TO TEALIUM
  • 6+ years in Site Reliability, Observability, or Platform Engineering supporting 24x7x365 production systems.
  • Deep experience with OpenTelemetry and platforms like Datadog, Sumo Logic, Prometheus, or Grafana for telemetry, logging, tracing, and alerting.
  • Experience designing observability for distributed systems, APIs, event-driven architectures, and asynchronous workflows.
  • Hands-on experience instrumenting AI/ML or GenAI systems, including model invocation, prompts, retrieval, tool use, and evaluation.
  • Proficiency with Amazon Bedrock or similar platforms to track model usage, latency, throttling, token metrics, and cost attribution.
  • Familiarity with agentic workflows, prompt engineering, vector databases (e.g., Neptune), RAG architectures, and frameworks like LangChain, LlamaIndex, or SageMaker.
  • Proficiency in Java, Python, or Go.
  • Strong AWS expertise, with an understanding of cloud networking, IAM, security, and service quotas.
  • Solid IaC, CI/CD, and container experience (Terraform, Kubernetes, Argo CD, Jenkins, or GitHub Actions).
  • Experience with data modeling, telemetry pipelines, cardinality management, retention, responsible AI, and data privacy controls.
  • Strong communication, mentoring, and cross-functional leadership skills across engineering, product, and non-technical stakeholders.
WAGE TRANSPARENCY

In several countries worldwide, including regions within the EMEA and APJ, employers are required or strongly encouraged to include salary ranges in job postings. While requirements vary by location, transparency is a core value at Tealium. We’re committed to providing clear and consistent compensation information to all applicants, regardless of location. This position offers a base salary range of 290,000 - 375,000 PLN (Polish zloty) annually. The final offer is determined by job-related skills, experience, and qualifications. The role may also be eligible for a performance-based bonus and equity options.

WHY YOU WANT TO WORK HERE
  • Tealium WOWs (Ways of Work).
  • Our award‑winning culture is how we think, act and connect together at Tealium Mosaic.
  • Our commitment to diversity, equity and inclusion is grounded in our mosaic of diverse perspectives and shared belonging as we live in work across the US and in nearly 20 countries.
  • Tealium Cares, to promote caring in our communities.
  • 15 hours of paid work time for volunteer activities and programs is offered annually.
  • Tealium Connects (remote‑first working), enabling many of us to choose where we do our best work and offering new‑hire stipends to assist with purchasing things we need to support a successful home‑office environment.
  • Tealium Ownership, share in the success of Tealium by becoming an owner of Tealium beginning with new‑hire equity grants.
  • Tealium Time, paid‑time‑off policy to offer flexibility to take time when needed and robust leave programs, including extended paid parental leave and company holidays.
  • Healium, health and wellness programs to help us be our best selves in the experiences of health, physical, mental, social, and even financial well‑being and wellness.
  • Tealium LIFT (Learning is Facilitated at Tealium), offering a myriad of professional development opportunities with over 6,000 courses available on demand to best‑in‑class manager and leadership development programs.
  • Health and Related Benefits Programs, offering market competitive benefits programs.
  • It is our continuing philosophy to recruit and employ the best qualified individuals without regard to race, color, sex, religion, national origin, disability, age, sexual orientation, gender identity, and/or any other protected characteristic.
  • Tealium does not tolerate unlawful discrimination of any kind and strives to be an inclusive and respectful workplace.

The highly relevant and differentiated positioning of Tealium’s solutions makes this a unique and rewarding career opportunity. *Offerings vary by level and location. Tealium connects customer data across web, mobile, offline, and IoT so businesses can better connect with their customers. Tealium’s turnkey integration ecosystem supports more than 1,300 built‑in connections, empowering brands to create a complete and real‑time customer data infrastructure. More than 1,000 leading businesses throughout the world trust Tealium to power their customer data strategies.

Our Vision: To create a world where businesses can intelligently engage and delight their customers through real‑time unified data. Our Mission: To create a best‑of‑breed, secure, and compliant global platform that helps companies improve the value and actionability of their customer data.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Product Manager (Remote - Poland)
Sr. Product Manager (Remote - Poland)

Tealium • Poland

On-site
PLN 230,000 - 290,000
Senior AI Observability Engineer
Senior AI Observability Engineer

Tealium Inc. • Polska

Hybrid
PLN 290,000 - 375,000
Remote-first working
New-hire equity grants
Stipends for home office
+1
Senior C++ Engineer with Observability
Senior C++ Engineer with Observability

EPAM Systems • Poland

On-site
PLN 180,000 - 260,000
Hybrid work within Poland
Work abroad up to 60 days/yr
Relocation opportunities
+2
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Kraków

On-site
PLN 250,000 - 350,000
Health insurance
Multisport
Shopping vouchers
+1
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Łódź

Hybrid
PLN 200,000 - 320,000
Hybrid by design
Remote work within Poland
Relocation opportunities
+2
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Województwo pomorskie

Hybrid
PLN 240,000 - 360,000
Hybrid by design
Remote within Poland
Relocation opportunities
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Warszawa

Hybrid
PLN 260,000 - 380,000
Hybrid by design
Remote within Poland
Relocation opportunities
+2
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Katowice

Hybrid
PLN 180,000 - 300,000
Health insurance
Multisport
Shopping vouchers
+2
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Poznań

Hybrid
PLN 180,000 - 260,000
Hybrid work model
Employee Stock Purchase Plan
Relocation opportunities
Senior Engineer - C++ Go or Rust
Senior Engineer - C++ Go or Rust

EPAM Systems • Wrocław

Hybrid
PLN 200,000 - 270,000
Health insurance
Multisport
Shopping vouchers
+3