Lead Site Reliability Engineer, Data- FreeWheel

Comcast

Reston (VA)

On-site

USD 152,000 - 228,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

FreeWheel, a Comcast company, seeks an experienced Data SRE to ensure reliability, scalability, and performance of data systems. You will work with data engineers and operations to automate daily tasks, monitor pipelines, and resolve issues across storage and analytics platforms.

Responsibilities include designing monitoring, building automation tools, optimizing data processing, and ensuring high availability.

Qualifications

  • At least 10+ years of experience as an SRE, DevOps, or Data Operations Engineer.
  • Experience with cloud platforms (AWS, GCP, Azure).
  • Familiarity with modern data architectures and technologies including Kafka, Hadoop, Spark, Cassandra, HDFS, and AWS S3.
  • Extensive experience in database management including NoSQL, MySQL, and PostgreSQL.
  • Proficiency with Ansible, Terraform, Kubernetes, and Docker.
  • Programming skills in Python, Go, Java, or Scala.
  • Experience with Prometheus, Grafana, ELK Stack, or similar tools.
  • Strong troubleshooting and debugging skills.
  • Excellent communication skills with technical and non-technical stakeholders.
  • Education: Bachelor’s degree or higher in Computer Science, Software Engineering, or a related field.

Responsibilities

  • Design and implement monitoring and alerting systems to ensure the stability, reliability, and performance of data platforms.
  • Quickly respond to and resolve issues impacting data pipelines or storage layers.
  • Develop and maintain automation tools and scripts for deployment, monitoring, backup, recovery, and disaster recovery of data systems.
  • Analyze and optimize the performance of data storage, query performance, and data flows.
  • Ensure efficient processing of large-scale datasets.
  • Reduce latency and improve processing speed.
  • Respond quickly to data platform failures.
  • Perform troubleshooting and coordinate cross-team efforts to resolve issues.
  • Ensure high availability and reliability of data platforms.
  • Work with data engineering teams to analyze and forecast capacity requirements.
  • Ensure systems can accommodate data growth and scale infrastructure accordingly.
  • Document the architecture, configurations, and operational procedures for data platforms.
  • Share operational knowledge across the team and provide relevant training.
  • Ensure data platforms meet security standards and compliance requirements.
  • Prevent data breaches, unauthorized access, and misuse of data.
  • Collaborate with data science, product, and development teams.
  • Support data product design and implementation efforts.
  • Resolve reliability-related issues and improve platform stability.

Skills

AWS
Kafka
Python

Education

Bachelor's degree

Tools

Kubernetes
Docker
Terraform
Ansible
Prometheus
Grafana
ELK Stack
Hadoop
Spark
Cassandra
MySQL
PostgreSQL
NoSQL
AWS S3
Kafka
Snowflake

Job description

FreeWheel, a Comcast company, provides comprehensive ad platforms for publishers, advertisers, and media buyers. Powered by premium video content, robust data, and advanced technology, we’re making it easier for buyers and sellers to transact across all screens, data types, and sales channels. As a global company, we have offices in nine countries and can insert advertisements around the world.

Job Summary

FreeWheel is seeking an experienced Data SRE to join the FreeWheel Data SRE team. As a member of the Global Operation team, you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage our data infrastructure, optimize system reliability, automate daily operations, and resolve technical issues that impact our data pipelines and backend data platforms.

Job Description
Key Responsibilities
System Monitoring and Optimization
  • Design and implement monitoring and alerting systems to ensure the stability, reliability, and performance of data platforms.
  • Quickly respond to and resolve issues impacting data pipelines or storage layers.
Automation and Tool Development
  • Develop and maintain automation tools and scripts for deployment, monitoring, backup, recovery, and disaster recovery of data systems.
Performance Optimization
  • Analyze and optimize the performance of data storage, query performance, and data flows.
  • Ensure efficient processing of large-scale datasets.
  • Reduce latency and improve processing speed.
Incident Response and Troubleshooting
  • Respond quickly to data platform failures.
  • Perform troubleshooting and coordinate cross-team efforts to resolve issues.
  • Ensure high availability and reliability of data platforms.
Capacity Planning and Scaling
  • Work with data engineering teams to analyze and forecast capacity requirements.
  • Ensure systems can accommodate data growth and scale infrastructure accordingly.
Documentation and Knowledge Sharing
  • Document the architecture, configurations, and operational procedures for data platforms.
  • Share operational knowledge across the team and provide relevant training.
Security and Compliance
  • Ensure data platforms meet security standards and compliance requirements.
  • Prevent data breaches, unauthorized access, and misuse of data.
Cross-Team Collaboration
  • Collaborate with data science, product, and development teams.
  • Support data product design and implementation efforts.
  • Resolve reliability-related issues and improve platform stability.
Qualifications
  • At least 10+ years of experience as an SRE, DevOps, or Data Operations Engineer.
  • Experience with cloud platforms (AWS, GCP, Azure).
  • Familiarity with modern data architectures and technologies including Kafka, Hadoop, Spark, Cassandra, HDFS, and AWS S3.
  • Extensive experience in database management including NoSQL, MySQL, and PostgreSQL.
  • Proficiency with Ansible, Terraform, Kubernetes, and Docker.
  • Programming skills in Python, Go, Java, or Scala.
  • Experience with Prometheus, Grafana, ELK Stack, or similar tools.
  • Strong troubleshooting and debugging skills.
  • Excellent communication skills with technical and non-technical stakeholders.
  • Education: Bachelor’s degree or higher in Computer Science, Software Engineering, or a related field.
Additional Preferred Skills
  • Experience with Aerospike, Kafka, Snowflake, and other big data technologies.
  • Familiarity with containerization, microservices architecture, and Kubernetes.
  • Experience designing and maintaining large-scale distributed systems.
  • Experience in data quality management, data governance, or ETL pipelines.

Disclaimer: This information has been designed to indicate the general nature and level of work performed by employees in this role. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications.

Comcast is an equal opportunity workplace. We will consider all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender identity, national origin, disability, veteran status, genetic information, or any other basis protected by applicable law.

Skills

Amazon Web Services (AWS); Apache Kafka; Python (Programming Language)

Salary

Virginia Pay Range: $152,020.01 - $228,030.02

Comcast intends to offer the selected candidate base pay within this range, dependent on job-related, non-discriminatory factors such as experience. The application window is 30 days from the date job is posted, unless the number of applicants requires it to close sooner or later.

Base pay is one part of the Total Rewards that Comcast provides to compensate and recognize employees for their work. Most sales positions are eligible for a Commission under the terms of an applicable plan, while most non-sales positions are eligible for a Bonus. Additionally, Comcast provides best-in-class Benefits to eligible employees. We believe that benefits should connect you to the support you need when it matters most, and should help you care for those who matter most. That’s why we provide an array of options, expert guidance and always-on tools, that are personalized to meet the needs of your reality - to help support you physically, financially and emotionally through the big milestones and in your everyday life. Please visit the compensation and benefits summary on our careers site for more details.

Education

Bachelor's Degree

While possessing the stated degree is preferred, Comcast also may consider applicants who hold some combination of coursework and experience, or who have extensive related professional experience.

Relevant Work Experience

10 Years +

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer, Data - FreeWheel
Site Reliability Engineer, Data - FreeWheel

Comcast • West Chicago (IL)

On-site
USD 100,000 - 164,000
Site Reliability Engineer, Data - FreeWheel
Site Reliability Engineer, Data - FreeWheel

Comcast (CC) of Willow Grove • Reston (VA)

On-site
USD 109,000 - 164,000
Site Reliability Engineer, Data - FreeWheel
Site Reliability Engineer, Data - FreeWheel

Comcast Advertising • Chicago (IL), Northern (KY)

Hybrid
USD 100,000 - 150,000
SMB Senior Account Executive, Comcast Business
SMB Senior Account Executive, Comcast Business

Comcast Advertising • Olympia (WA)

On-site
USD 134,000 - 201,000
Site Reliability Engineer, Data - FreeWheel
Site Reliability Engineer, Data - FreeWheel

Comcast • Chicago (IL)

On-site
USD 120,000 - 180,000
Data Engineer 3 - Reston, VA - Freewheel
Data Engineer 3 - Reston, VA - Freewheel

Comcast • Reston (VA), Northern (KY)

On-site
USD 104,000 - 170,000
Site Reliability Engineer, Streaming HUB - FreeWheel
Site Reliability Engineer, Streaming HUB - FreeWheel

Comcast • Reston (VA)

On-site
USD 109,000 - 164,000
Principal Software Engineer - Distributed Systems - FreeWheel
Principal Software Engineer - Distributed Systems - FreeWheel

FreeWheel • Chicago (IL)

On-site
USD 152,829 - 229,243
Sr. C++ Backend Engineer - Hybrid 2 Days Reston, VA - FreeWheel
Sr. C++ Backend Engineer - Hybrid 2 Days Reston, VA - FreeWheel

Comcast • Reston (VA)

On-site
USD 142,000 - 213,000
Sr. Data Engineer - Freewheel
Sr. Data Engineer - Freewheel

Comcast • New York (NY)

On-site
USD 148,000 - 222,000