Big Data Support Engineer Lead - Vice President

Citibank (Switzerland) AG

Irving (TX)

On-site

Confidential

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A global financial institution seeks an experienced candidate for an Incident Management and SRE role in Irving, Texas. This position requires leading service delivery teams and managing production services, while ensuring automation and proper monitoring tools are implemented. The ideal candidate has over 10 years of IT experience, with a preference for backgrounds in finance and applications development. Strong communication and analytical skills are critical, and candidates must be able to work under pressure and manage competing priorities.

Qualifications

  • 10+ years of overall IT experience, with 2+ years in cloud technologies.
  • Experience in financial services and applications development preferred.
  • Strong knowledge of monitoring and observability tools like Splunk and Grafana.

Responsibilities

  • Lead incident management, ensuring swift resolution of issues.
  • Manage service delivery functions including service risk assessments.
  • Collaborate with teams to develop and test knowledge objects.

Skills

Incident Management
Automation
Monitoring/Observability
Communication Skills
Problem-solving
Data Engineering
Cloud Technologies

Education

Bachelor’s or Master’s degree in engineering or computer science

Tools

App Dynamics
Splunk
Grafana
Ansible

Job description

## For additional information, please review .* ## Has a strong understanding and experience in leading all aspects of Incident Management, Problem Management, Service Improvements, Monitoring and Observability instrumentation, SRE(Site Reliability engineering) Frameworks and adoption, Disaster recovery and resiliency, and automation of production services.* ## Leads the production monitoring, Implementation of Observability using AppD, Splunk, Grafana & strong knowledge of monitoring tools used in the industry.* ## Collaborates with development team, Architecture teams and Infrastructure teams and leads service improvement plans.* ## Supports the delivery of the L2 Service Delivery and SRE (Site reliability engineering) objectives for the business/region.* ## Leads the team and contributes towards achievement of service performance against targets for the organization.* ## Strong bias towards automation and using SRE Framework.* ## Responsible for execution of day-to-day service delivery functions including the following: + ## Incident Management: Performs incident triage, root cause analysis, and collect and validate business impact. + ## Service Management: Collaborates with Technology Organization and manages service risk/maturity assessments and drives the Service Improvement Plans. + ## Knowledge Management: Develops/tests knowledge objects to support increased L0, L1, and L2 resolution. + ## Change Management: Review and approve changes. + ## Capacity Management: Review capacity across service components. + ## Continuity Management: Schedule and facilitate COB testing, maintain recovery plans. + ## Configuration Management: Build/update service configuration. + ## Third Party Asset Management: Manage 3rd party asset management (licensing compliance/optimization) + ## Service Readiness: This activity encompasses review of major releases & new application install from very early stage of the project/program, to ensure Risks are documented and remediated before production Go Live. + ## Service Risks: Ability to identify, document and Manage Service Risks within Applications and effectively manage the resolutions of Risks. + ## Monitoring: Collaborate and engage with various teams to enable monitoring / observability of production services.* ## At least 10+ years of hands on Overall IT experience of which 2 or more years in one or more of the Cloud technologies running services on Open Shift, AWS or Google Cloud.* ## Has a good understanding of Data Engineering function and Role and tools and technologies used in one or more technologies including Ab Initio, Big Data, Master Data Management (MDM) and Hybrid cloud.* ## Strong knowledge of using CICD tools for automated code deployments.* ## Strong knowledge of SOAP Rest API’s and Micro services.* ## Knowledge of creating Observability Dashboards using Splunk, App Dynamics, ELK and Grafana.* ## Working knowledge of Ansible scripts for Automation.* ## Track record of successfully triaging issues and driving them to resolution.* ## Ability to work under pressure and manage deadlines or unexpected changes in expectations or requirements* ## Expectation for the role is to be available "on call" or "shift basis" for off hours Production support.* ## Can handle multiple, competing priorities simultaneously* ## Ability to work with Offshore and Onsite Production Support Teams across multiple organizations.* ## Excellent Oral and Written Communication Skills.* ## Good knowledge of Disaster Recovery process across Data centers.* ## Strong analytical skills, strong problem-solving skills and ability to logically break down tasks into smaller manageable parts.* ## Strong individual with the ability to communicate and negotiate at all levels and ability to influence people.* ## Effective meeting management, team management and organizational skills.* ## Effective Presentation skills and creating Visual Presentations using Microsoft PowerPoint.* ## Ability to interact with individuals at all organizational levels.* ## Bachelor’s or Master’s degree in engineering or computer science.* ## Prior experience in financial services preferred.* ## Prior experience in applications development preferred.* ## Prior Experience in Data Warehouse and Business Intelligence Applications.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr Data Engineer
Sr Data Engineer

ViziRecruiter,LLC. • Merrimack (NH)

On-site
USD 110,000 - 170,000
Data &Analytics Data Operations Lead
Data &Analytics Data Operations Lead

TechDigital Group • Alpharetta (GA)

On-site
USD 85,000 - 120,000
Data Practice Architect/Data Solution Lead
Data Practice Architect/Data Solution Lead

Hitachi Automotive Systems Americas, Inc. • Dallas (TX)

Hybrid
USD 130,000 - 160,000
Analytics Lead
Analytics Lead

Compunnel, Inc. • San Diego (CA)

Hybrid
USD 90,000 - 130,000
Splunk Administrator (Site Reliability Engineer)
Splunk Administrator (Site Reliability Engineer)

SchoolsFirst FCU • Tustin (CA)

Hybrid
USD 63,000 - 103,000
Big Data Infrastructure Administrator
Big Data Infrastructure Administrator

Right Talent Right Now • Nashua (NH)

On-site
USD 100,000 - 130,000
Business Information Mgmt Specialist (US)
Business Information Mgmt Specialist (US)

TD Bank • New York (NY)

On-site
USD 80,000 - 110,000
Assurance santé
Opportunités de formation
Flexibilité du travail
Big Data Infrastructure Administrator
Big Data Infrastructure Administrator

Right Talent Right Now • Garland (TX)

On-site
USD 100,000 - 130,000
Big Data Infrastructure Administrator
Big Data Infrastructure Administrator

Right Talent Right Now • Fort Worth (TX)

On-site
USD 90,000 - 120,000
Big Data Infrastructure Administrator
Big Data Infrastructure Administrator

Right Talent Right Now • Dallas (TX)

On-site
USD 100,000 - 120,000