Sr. Site Reliability Engineer, tvScientific

Pinterest

San Francisco (CA)

On-site

USD 139,764 - 287,749

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Competitive salary

Job summary

Pinterest is seeking a Senior Site Reliability Engineer to operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows.

You will drive reliability and automation, handle incident response, and advance observability while collaborating with security and platform teams. Strong Kubernetes and AWS production experience are required, plus a bachelor’s degree or equivalent.

Qualifications

  • 4+ years in SRE/DevOps/Platform Eng/Cloud Infra
  • Strong production AWS experience
  • Kubernetes expertise and multi-tenancy
  • GitOps with ArgoCD
  • Terraform/Terragrunt for infra provisioning
  • Scripting in Bash/Python
  • CI/CD pipelines (GitHub Actions)
  • Observability, monitoring, and alerting
  • Strong collaboration across teams
  • Bachelor's degree or equivalent
  • Bias for security and platform guardrails

Responsibilities

  • Ensure reliability, availability, and performance of production infra and services
  • Operate and scale Kubernetes platforms for multi-tenant workloads
  • Manage GitOps workflows with ArgoCD and Helm
  • Drive infrastructure provisioning with Terraform/Terragrunt
  • Build and maintain CI/CD pipelines (GitHub Actions)
  • Lead incident response, RCA, and post-incident improvements
  • Reduce toil with automation and tooling
  • Advance observability with logs, metrics, traces, dashboards, alerts
  • Support secure secrets, IAM-aware operations, and guardrails
  • Collaborate with application, security, and platform teams

Skills

AWS in production
Kubernetes operations
GitOps
ArgoCD
Helm
Terraform/Terragrunt
Bash/Python scripting
CI/CD
Incident response
Observability
Collaboration across teams
Ownership mindset

Education

Bachelor's degree in CS/Engineering or equivalent

Tools

ArgoCD
Helm
Terraform
Terragrunt
GitHub Actions

Job description

About Pinterest:

Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we're on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product.

About tvScientific

tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.

We are seeking a Senior Site ReliabilityEngineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows. This role will be instrumental in advancing the reliability, scalability, automation, observability, and operational maturity of our infrastructure and delivery ecosystem.

The ideal candidate is a highly hands‑on engineer with strong production experience and a proven ability to build and support resilient platforms using infrastructure as code, automation, and modern Kubernetes operational practices.

What you'll do:
  • Ensuring the reliability, availability, and performance of production infrastructure and platform services
  • Operating and scaling Kubernetes platforms, including governance and support for multi‑tenant workloads
  • Managing GitOps-based deployment workflows using ArgoCD and Helm
  • Driving infrastructure provisioning and change management through Terraform/Terragrunt
  • Building and supporting CI/CD automation and deployment workflows using GitHub Actions
  • Leading incident response efforts, root cause analysis, and post‑incident improvement initiatives
  • Reducing operational toil through scripting, tooling, and process automation
  • Advancing observability practices across logs, metrics, traces, dashboards, and alerting
  • Supporting secure secrets integration, IAM-aware operations, and platform guardrails
  • Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes
What we're looking for:
  • 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
  • Strong hands‑on experience operating AWS in production environments
  • Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration
  • Proven experience with Kubernetes multi‑tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns
  • Experience implementing and operating ArgoCD within a GitOps delivery model
  • Strong hands‑on experience with Helm
  • Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management
  • Solid scripting and automation skills using Bash and/or Python
  • Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions
  • Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems
  • Experience with monitoring, alerting, and observability in production environments
  • Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow‑through after outages
  • Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams
  • Bachelor's degree in computer science, engineering, a related field or equivalent experience
  • Demonstrated ability to use AI to improve speed and quality in your day‑to‑day workflow for relevant outputs
  • Strong track record of critical evaluation and verification of AI‑assisted work (e.g., testing, source‑checking, data validation, peer review)
  • High integrity and ownership: you protect sensitive data, avoid over‑reliance on AI, and remain accountable for final decisions and deliverables.
In‑Office Requirement Statement:

We recognize that the ideal environment for work is situational and may differ across departments. What this looks like day‑to‑day can vary based on the needs of each organization or role.

Relocation Statement:

This position is not eligible for relocation assistance. Visit ourPinFlexpage to learn more about our working model.

US based applicants only

We are sharing the base salary range for this position. The position is also eligible for equity. Final salary is based on a number of factors including location, travel, relevant prior experience, or particular skills and expertise.

$139,764 — $287,749 USD

Our Commitment to Inclusion:

Pinterest is an equal opportunity employer and makes employment decisions on the basis of merit. We want to have the best qualified people in every job. All qualified applicants will receive consideration for employment without regard to race, color, ancestry, national origin, religion or religious creed, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, age, marital status, status as a protected veteran, physical or mental disability, medical condition, genetic information or characteristics (or those of a family member) or any other consideration made unlawful by applicable federal, state or local laws. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, tvScientific
Senior Site Reliability Engineer, tvScientific

Pinterest • San Francisco (CA)

On-site
USD 140,000 - 288,000
Sr. Site Reliability Engineer, tvScientific
Sr. Site Reliability Engineer, tvScientific

Pinterest • United States

On-site
USD 139,000 - 288,000
Software Engineer II, tvScientific
Software Engineer II, tvScientific

Pinterest • California (MO)

On-site
USD 130,000 - 180,000
Sr. Production Engineer, Solutions Engineering
Sr. Production Engineer, Solutions Engineering

Pinterest • California (MO)

Hybrid
USD 139,000 - 288,000
Senior Software Engineer, Machine Learning, tvScientific
Senior Software Engineer, Machine Learning, tvScientific

Pinterest • San Francisco (CA)

On-site
USD 156,000 - 320,000
Equity
Sr. Data Scientist, tvScientific
Sr. Data Scientist, tvScientific

Pinterest • San Francisco (CA)

On-site
USD 139,000 - 288,000
Sr. Data Scientist, tvScientific
Sr. Data Scientist, tvScientific

Pinterest • San Francisco (CA)

On-site
USD 139,000 - 288,000
Senior Client Account Manager, CTV & Performance
Senior Client Account Manager, CTV & Performance

Pinterest • San Francisco (CA)

On-site
Manager I, Client Account Manager, tvScientific (Affiliate/CPA)
Manager I, Client Account Manager, tvScientific (Affiliate/CPA)

Pinterest • San Francisco (CA)

On-site
USD 91,000 - 190,000
Group Product Manager II, Creative Tech - tvScientific
Group Product Manager II, Creative Tech - tvScientific

Pinterest • San Francisco (CA)

Hybrid
USD 195,000 - 403,000