TAS Tower Lead

Tata Consultancy Services

Irvine (CA)

On-site

USD 110,000 - 140,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Tata Consultancy Services in Irvine, CA seeks an experienced TAS Tower Lead/L3 Production Support Engineer to provide advanced technical and operational support for data and analytics applications. You will resolve incidents, optimize pipelines, and support CI/CD with Harness in a containerized, cloud-based environment.

The role requires hands-on expertise in Python, R, Alteryx, SQL/PLSQL, AWS, Airflow, Kubernetes, and production excellence.

Qualifications

  • Hands-on experience with Python, R, Alteryx, SQL/PLSQL and AWS.
  • Strong production incident resolution and root-cause analysis skills.
  • Experience with Airflow, Kubernetes and Harness in production environments.
  • Ability to design, troubleshoot, and optimize data pipelines and analytics workflows.
  • Experience with BAU support, preventive fixes, RCA documentation, and enhancements.

Responsibilities

  • Provide L3 production support for critical data, analytics, and application platforms, ensuring availability, stability, and performance.
  • Investigate and resolve complex production incidents across applications, data pipelines, databases, infrastructure, and integrations.
  • Develop, troubleshoot, and optimize applications and automation scripts using Python and R.
  • Design, maintain, and monitor data workflows developed with Alteryx.
  • Write and optimize SQL/PLSQL queries, stored procedures, and data validation scripts.
  • Support data storage, file processing, archival, and recovery activities using AWS S3.
  • Monitor and troubleshoot notifications and event-driven integrations using AWS SNS.
  • Manage container images and repositories in AWS ECR.
  • Monitor, troubleshoot, and optimize Apache Airflow DAGs, scheduling, and performance.
  • Support containerized applications on Kubernetes including pods, deployments, services, and logs.
  • Use Harness to support CI/CD, deployments, release validation, rollback, and production monitoring.
  • Perform RCA for critical incidents and coordinate actions with engineering and infra teams.
  • Identify recurring issues and implement preventive and permanent fixes.
  • Handle BAU activities: health checks, batch monitoring, data validation, service requests, access requests, and operational reporting.
  • Coordinate with development, database, cloud, infra, DevOps, and business teams during high-priority incidents.
  • Participate in incident, problem, change, and release management processes with defined SLAs.
  • Perform impact analysis, design, coding, testing, deployment, and post-production validation for enhancements.
  • Support planned releases, infra changes, platform upgrades, patching, and maintenance.
  • Develop automation solutions to reduce manual effort and improve support efficiency.
  • Create and maintain runbooks, troubleshooting guides, and knowledge-base articles.
  • Monitor logs, alerts, workflows, cloud components, and Kubernetes workloads to pre-empt failures.
  • Participate in on-call rotations and provide technical assistance during incidents and releases.
  • Track incidents through closure with status updates and RCA/ preventive actions.
  • Drive continuous service improvement through automation, performance, monitoring, and process optimization initiatives.

Skills

Python
R
Alteryx
SQL/PLSQL
AWS
Airflow
Kubernetes
Harness
Production Support
RCA
BAU
Enhancements

Tools

AWS S3
AWS SNS
AWS ECR
Airflow
Kubernetes
Harness

Job description

Job Description
TAS Tower Lead
Must Have Technical/Functional Skills

Python, R, Alteryx, SQL/PLSQL, AWS S3, SNS, ECR, Airflow, Kubernetes, Harness, production support, RCA, preventive fixes, BAU and enhancements

Roles & Responsibilities

We are seeking an experienced L3 Production Support Engineer to provide advanced technical and operational support for business-critical data and analytics applications. The candidate should have strong hands-on experience with Python, R, Alteryx, SQL/PLSQL, AWS, Airflow, Kubernetes, and Harness, along with proven expertise in production incident resolution, root cause analysis, preventive fixes, BAU support, and application enhancements.

  • Provide L3 production support for critical data, analytics, and application platforms, ensuring system availability, stability, and performance.
  • Investigate and resolve complex production incidents involving applications, data pipelines, database procedures, infrastructure components, and system integrations.
  • Develop, troubleshoot, and optimize applications, automation scripts, and analytical workflows using Python and R.
  • Design, maintain, monitor, and support data preparation and analytics workflows developed using Alteryx.
  • Write and optimize complex SQL and PL/SQL queries, stored procedures, functions, packages, and scripts for production troubleshooting and data validation.
  • Support application data storage, file processing, archival, and recovery activities using AWS S3.
  • Monitor and troubleshoot application notifications and event-driven integrations implemented using AWS SNS.
  • Manage, maintain, and troubleshoot container images and repositories hosted in AWS Elastic Container Registry (ECR).
  • Monitor, support, and troubleshoot batch and data-processing workflows orchestrated through Apache Airflow, including DAG failures, scheduling issues, dependencies, and performance bottlenecks.
  • Support containerized applications deployed on Kubernetes, including troubleshooting pods, deployments, services, configurations, resource utilization, and application logs.
  • Use Harness to support CI/CD pipelines, application deployments, release validation, rollback activities, and production deployment monitoring.
  • Perform detailed Root Cause Analysis (RCA) for critical and recurring incidents, document findings, and coordinate corrective actions with engineering and infrastructure teams.
  • Identify recurring production issues and implement preventive and permanent fixes to improve platform reliability and reduce incident volume.
  • Handle Business-as-Usual (BAU) activities, including daily health checks, batch monitoring, job reruns, data validation, service requests, access-related requests, and operational reporting. Analyze data issues and perform reconciliation, profiling, validation, and correction activities to ensure data accuracy, completeness, and consistency.
  • Coordinate with application development, database, cloud, infrastructure, DevOps, and business teams during high-priority production incidents.
  • Participate in incident, problem, change, and release management processes and ensure compliance with defined operational procedures and SLAs.
  • Perform impact analysis, technical design, coding, testing, deployment, and post-production validation for minor and medium-sized application enhancements.
  • Support planned releases, infrastructure changes, platform upgrades, patching, and production maintenance activities.
  • Develop automation solutions using Python, SQL/PLSQL, and platform utilities to reduce manual operational effort and improve support efficiency.
  • Create and maintain technical documentation, operational runbooks, troubleshooting guides, support procedures, and knowledge-base articles.
  • Monitor application logs, system alerts, scheduled workflows, cloud components, and Kubernetes workloads to proactively identify potential failures.
  • Participate in on-call support rotations and provide technical assistance during critical incidents, production releases, and planned maintenance activities.
  • Track incidents through closure, provide timely status updates, and communicate business impact, recovery actions, root causes, and preventive measures to stakeholders.
  • Drive continuous service improvement by identifying automation opportunities, performance enhancements, monitoring improvements, and process optimization initiatives.

Salary Range- $110,000-$140,000 a year

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DFS Tower Lead
DFS Tower Lead

Tata Consultancy Services • Irvine (CA)

On-site
USD 110,000 - 140,000
Mid Level AWS Production Support Engineer
Mid Level AWS Production Support Engineer

System One • Birmingham (AL)

On-site
USD 80,000 - 120,000
Production Support
Production Support

Tata Consultancy Services • Phoenix (AZ)

On-site
USD 90,000 - 100,000
L2 Production support
L2 Production support

Tata Consultancy Services • Clearwater (FL)

On-site
USD 90,000 - 110,000
Production Support Engineer
Production Support Engineer

ASM Tech Solutions • Town of Florida (NY)

Hybrid
USD 90,000 - 120,000
Manager - Production Support - Application / Infrastructure / Middleware
Manager - Production Support - Application / Infrastructure / Middleware

Request Technology, LLC • Chicago (IL)

Hybrid
USD 140,000 - 200,000
Application Support Engineer
Application Support Engineer

Programmers.io • Sunnyvale (CA)

On-site
USD 120,000 - 150,000
Level 3 Support Engineer
Level 3 Support Engineer

Tata Consultancy Services • New York (NY)

On-site
USD 100,000 - 115,000
Sr Technical Lead-App Development
Sr Technical Lead-App Development

Birlasoft • Northern (KY)

Hybrid
USD 110,000 - 150,000
Technology Support III
Technology Support III

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 90,000 - 120,000