Sr. Site Reliability Engineer I

DoubleVerify

Northern, New York (KY, NY)

Hybrid

USD 104,000 - 178,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

DoubleVerify is seeking a Sr. Site Reliability Engineer I to join our Hybrid NYC HQ. You’ll drive automation and observability across DV’s digital media measurement platforms, spanning GCP, AWS, and on-prem environments.

You’ll implement IaC, build monitoring, and lead projects to improve reliability and scale. You will partner with engineering and operations to reduce incidents, optimize performance, and enable self-service tooling.

Qualifications

  • 4+ years in Site Reliability Engineering, DevOps, or related roles.
  • Strong scripting/programming in Python, Bash, or Go for automation.
  • Experience with cloud platforms and container orchestration like Kubernetes.
  • Expertise with monitoring/observability tools such as Prometheus, Grafana, Splunk, or Nagios.

Responsibilities

  • Build and maintain reliability, scalability, and performance of DV's platforms.
  • Implement observability dashboards, metrics, and alerting.
  • Reduce MTTR through automation and proactive monitoring.
  • Develop IaC with Terraform, Helm, and configuration tools.
  • Lead technical projects from planning through deployment.

Skills

Python
Linux/Unix
Bash
Go
Automation
Observability

Education

Bachelor's degree in CS/Engineering or related field

Tools

Kubernetes
Prometheus
Grafana
Splunk
Nagios
Terraform
Ansible
Helm
GitLab

Job description

## Sr. Site Reliability Engineer IApply: Hybrid: NYC Global HQ: Full time: Posted 12 Days Ago: JR00000779# ****Who We Are****DV is the leader in digital performance solutions, helping our advertiser and agency partners Verify the quality of their digital campaigns, Optimise to improve performance and Prove that they’re achieving their business outcomes, through unbiased 3rd party data and analytics. DV’s mission is to be the definitive source of transparency and data-driven insights into the quality and effectiveness of digital advertising for the world’s largest brands, agencies, publishers, and digital ad platforms. Since 2008, DV has helped hundreds of Fortune 500 companies gain the most from their media spend by delivering best-in-class solutions across the digital advertising ecosystem, helping to build a better industry. Learn more at www.doubleverify.com.## # ****About the Role****### **Role Overview**You will join the Site Reliability Engineering (SRE) team within DoubleVerify's Technology organization. The team is responsible for building and maintaining the reliability, scalability, and performance of DV's digital media measurement platforms, which run across GCP, AWS, and on-premises environments. As a Sr. Site Reliability Engineer I, you'll be a hands-on technical contributor driving automation, observability, and operational excellence across the platforms that power DV's product suite.### ****What You’ll Do***** Build and maintain the reliability, scalability, and performance of DV's digital media measurement platforms.* Leverage AI-assisted development tools to accelerate automation development and problem resolution.* Build custom integrations and MCP servers for monitoring platforms to enable programmatic access and AI-driven analysis.* Implement observability best practices, including metrics collection, dashboarding, and alerting strategies that support proactive reliability improvements.* Monitor and maintain high-availability infrastructure and services across GCP, AWS, and on-premises environments.* Respond to incidents and drive them to resolution, managing Sev1/Sev2 situations.* Reduce MTTR (mean time to resolution) for critical incidents through automation, improved observability, and proactive monitoring.* Build and deploy automations to eliminate operational toil and improve efficiency across deployment workflows, validation scripts, and self-service capabilities.* Implement Infrastructure-as-Code using Terraform, Helm charts, Python scripts, and configuration management tools to ensure repeatable, version-controlled infrastructure deployments.* Develop production automations for routine operational tasks, reducing manual intervention and accelerating task completion.* Create and maintain documentation, runbooks, and SOPs in Confluence to ensure consistent incident response across the team.* Participate in on-call rotations and post-incident reviews to minimize downtime and prevent recurrence.* Lead technical projects from planning through deployment, ensuring proper stakeholder communication and team enablement.# ****About You****### ****Required Experience & Skills***** 4+ years in Site Reliability Engineering, DevOps, or related operational roles, with proven experience in Linux/Unix systems administration.* Proficiency in scripting and programming languages such as Python, Bash, or Go for automation and tool development.* Strong experience with cloud platforms and container orchestration tools like Kubernetes.* Expertise in monitoring and observability tools such as Prometheus, Grafana, Splunk, or Nagios.* Hands-on experience with Infrastructure-as-Code tools like Terraform, Ansible, or Helm.* Proven ability to develop and track SLIs, SLOs, and SLAs to drive reliability improvements.### ### ****Technical Knowledge***** Deep understanding of networking, DNS, load balancing, and CDN technologies.* Familiarity with databases (SQL, NoSQL, Vertica, MongoDB, Snowflake) and data pipeline technologies.* Knowledge of CI/CD pipelines, GitLab, and deployment automation.* Experience with workflow automation platforms is a strong plus.### ### ****Soft Skills & Mindset***** Exceptional communication skills with the ability to collaborate across teams and explain technical concepts clearly.* Proactive problem-solving approach with a focus on automation and continuous improvement.* Ownership mentality — you take full responsibility for complex challenges and reliably deliver outcomes.* Trailblazing spirit — innovative use of AI, automation, and new technologies to solve problems and drive improvements.* Passion for mentorship and knowledge sharing, elevating the capabilities of the entire team.### ### **Preferred Qualifications*** Bachelor's or Master's degree in Computer Science, Engineering, or a related field.* Industry certifications such as AWS Certified DevOps Engineer, Google Professional Cloud DevOps Engineer, Certified Kubernetes Administrator (CKA), or Terraform/Grafana certifications.* Experience with AI-assisted development using tools like ChatGPT, Cursor, Glean, or Copilot.* Familiarity with security best practices in cloud and containerized environments. The successful candidate’s starting salary will be determined by a number of non-discriminatory factors, including qualifications for the role, level, skills, experience, location, and internal equity relative to peers at DV.The estimated salary range for this role, based on the qualifications set forth in the job description, is between $104,000 $178,000. This role will also be eligible for bonus/commission (as applicable), equity, and benefits.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer I
Sr. Site Reliability Engineer I

DoubleVerify Inc. • New York (NY)

On-site
USD 104,000 - 178,000
Sr. Software Engineer II
Sr. Software Engineer II

DoubleVerify • Northern (KY), New York (NY)

Hybrid
USD 107,000 - 212,000
Sr. DevOps Engineer II, AI Platforms
Sr. DevOps Engineer II, AI Platforms

DoubleVerify • Northern (KY), New York (NY)

Hybrid
USD 153,000 - 260,000
Bonus / equity
Benefits
Hybrid work
Sr. Software Engineer II
Sr. Software Engineer II

DoubleVerify Inc. • New York (NY)

Hybrid
USD 107,000 - 212,000
Sr. DevOps Engineer II, AI Platforms
Sr. DevOps Engineer II, AI Platforms

DoubleVerify Inc. • New York (NY)

Hybrid
USD 153,000 - 260,000
Sr. Analytics Data Platform Engineer
Sr. Analytics Data Platform Engineer

DoubleVerify • Northern (KY), New York (NY)

Hybrid
USD 107,000 - 212,000
Sr. Analytics Data Platform Engineer
Sr. Analytics Data Platform Engineer

DoubleVerify Inc. • New York (NY)

On-site
USD 107,000 - 212,000
Corporate Systems Engineer, Collaboration Platforms
Corporate Systems Engineer, Collaboration Platforms

DoubleVerify • Northern (KY), New York (NY)

Hybrid
USD 94,000 - 160,000
Business Development Manager
Business Development Manager

DoubleVerify • Northern (KY), New York (NY)

Hybrid
USD 79,000 - 129,000
Bonus/Commission
Equity
Benefits
Sr. Application Security Manager
Sr. Application Security Manager

DoubleVerify Inc. • New York (NY)

Hybrid
USD 153,000 - 260,000
Bonus/Commission
Equity
Benefits