Senior Associate-Technology Operations Engineering

American Express

Town of Florida (NY)

On-site

USD 140,000 - 210,000

Full time

24 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

American Express in United States seeks a seasoned software engineering professional to drive reliability, performance, and modernization of enterprise platforms across distributed environments. You will partner with engineering, product, and operations to design, build, automate, and support highly available solutions powering customer capabilities.

You will apply SRE principles, develop automation, and contribute to incident response, capacity planning, and security posture.

Qualifications

  • Bachelor's degree in Computer Science, Computer Engineering, and/or comparable experience.

Responsibilities

  • Provide service reliability, incident response, and capacity planning.
  • Design, develop, test, and deploy scalable software solutions across distributed platforms.
  • Collaborate with engineering, product, and operations teams to automate and support enterprise platforms.
  • Apply SRE principles to improve system stability and reduce operational risk.

Skills

Java
Python
SQL
Distributed systems
REST API
GRPC
NodeJS
Bash
Spring
Git
Kubernetes
Docker
Kafka
GraphQL

Education

Bachelor's degree in Computer Science or related field

Tools

Jenkins
Grafana
Docker
Kubernetes
Git
Cloud platforms (AWS/GCP/Azure)

Job description

Job Description

This role is responsible for driving the reliability, resiliency, performance, and modernization of critical American Express platforms across Distributed environments. You will leverage deep technical expertise in software engineering, runtime engineering, production support, and platform operations to quickly assess and remediate complex availability, performance, and operational issues.

Job Description

This role is responsible for driving the reliability, resiliency, performance, and modernization of critical American Express platforms across Distributed environments. You will leverage deep technical expertise in software engineering, runtime engineering, production support, and platform operations to quickly assess and remediate complex availability, performance, and operational issues.

As part of our technology team, you will partner with engineering, product, infrastructure, and operations teams to design, build, automate, and support highly available enterprise platforms. You will help accelerate modernization initiatives, improve operational excellence, and deliver secure, scalable, and resilient solutions that power critical customer and business capabilities.

Responsibilities
Software Engineering & Platform Development
  • Serve as a hands‑on engineer with experience in supporting complex enterprise applications, platforms, and operational tooling across Distributed environments.
  • This role requires and must have 5+ years of distributed experience and knowledge.
  • Design, develop, prototype, code, test, and implement scalable software solutions using technologies such as Java, Python, SQL, and related frameworks.
  • Act as a technical contributor in change management reviews, root cause analysis, and troubleshooting of complex technical issues.
  • Design and implement automation solutions, and engineering practices that improve platform resiliency, operational efficiency, and security.
  • Use best practices in incident management, problem management and change management as this role focuses on application production support.
Runtime Engineering, Reliability & Operations
  • Contribute to the technical roadmap for runtime systems, ensuring platform reliability, scalability, availability, recoverability, and performance.
  • Establish, monitor, and continuously improve key performance indicators (KPIs), service level objectives (SLOs), and operational metrics (MTTR, MTBF) for platform health and resiliency.
  • Perform diagnosis and resolution of production incidents, batch failures, application outages, performance bottlenecks, and infrastructure issues across Mainframe and Distributed platforms.
  • Apply Site Reliability Engineering (SRE) principles and operational excellence practices to improve system stability and reduce operational risk.
  • Support disaster recovery, high availability, workload management, capacity planning, and business continuity initiatives.
Distributed Platform Engineering
  • Must have experience in Support and optimize enterprise platforms across distributed technologies including Java, Python, C, SQL, NodeJS, Bash, JS/HTML/CSS, Spanner, BigTable, BigQuery, Spring, GraphQL, OOP, MVC, Algos, Git, CoPilot, Jenkins, XLR, Elastic, Jira, AI, ML, Docker, Kubernetes, Kafka, Rest API, GRPC, Grafana.
  • Desirable Support and optimize Distributed Platform technologies including cloud infrastructure, Linux/Unix, containers, APIs, Java‑based services, distributed databases, and modern application platforms.
  • Implement and support Cloud/Distributed architectures, modernization initiatives, API enablement, and enterprise integration capabilities.
  • Knowledge of cloud platforms AWS, GCP or general Cloud fundamentals.
Data, Integration & Automation
  • Develop and support enterprise integration solutions utilizing APIs, MQ, Connect:Direct, event‑driven architectures, and batch and real‑time data integration patterns.
  • Utilize relational and NoSQL databases including DB2, PostgreSQL, Redis, and Couchbase to support critical business applications.
  • Automate operational processes, deployments, monitoring, reporting, and remediation activities using Python, Bash, Ansible, Jenkins, and related technologies.
  • Collaborate with engineering teams to adopt scalable automation and self‑service capabilities for deployment, monitoring, and operational support.
Observability & Continuous Improvement
  • Implement and utilize monitoring, observability, logging, and analytics solutions using tools such as Splunk, OMEGAMON, RMF/SMF, Sysview, MainView, Dynatrace, AppDynamics, ELK, or equivalent technologies.
  • Analyze operational trends, identify opportunities for optimization, and formulate strategic recommendations to improve platform health and engineering effectiveness.
  • Contribute to continuous improvement initiatives focused on reliability, performance, security, operational maturity, and customer experience.
DevOps, Security & Governance
  • Understand CI/CD pipelines and DevOps practices using tools such as Git, Jenkins, Maven, DBB, Endevor, Changeman, ISPW, UrbanCode Deploy, or equivalent platforms.
  • Apply enterprise security controls, compliance requirements, audit standards, and access management practices, including RACF, ACF2, Top Secret, and cloud security principles.
  • Ensure solutions meet non‑functional requirements (NFRs) including availability, scalability, performance, security, recoverability, and maintainability.
Qualifications
  • Bachelor's degree in Computer Science, Computer Engineering, and/or comparable experience
  • Work experience in software engineering, app support or infrastructure operations or runtime engineering.
  • A working understanding of cloud infrastructure, distributed systems, and containerization technologies, with experience in supporting critical business applications being a plus.
  • Familiarity with monitoring and logging tools, and incident management best practices, to ensure reliability and performance of applications in a production environment.
  • Solid programming and scripting skills, with hands on experience to automate operational tasks using tools such as Python
  • Knowledge of scripting languages (e.g., PowerShell, Python) for automation tasks
  • Experience in technology operations work
  • Hands on experience with relational and NoSQL databases such as DB2, Redis, Postgres, Couchbase etc.
  • Experience in cloud platforms such as AWS, Azure, or Google Cloud, Public Cloud certification is a plus

Employment eligibility to work with American Express in the United States is required as the company will not pursue visa sponsorship for these positions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Associate - Tech Ops Engineering
Senior Associate - Tech Ops Engineering

American Express • Phoenix (AZ)

On-site
USD 90,000 - 120,000
Senior Platform Reliability & Automation Engineer
Senior Platform Reliability & Automation Engineer

American Express • Town of Florida (NY)

On-site
USD 140,000 - 210,000
Senior Software Engineer, Digital Banking & Payments
Senior Software Engineer, Digital Banking & Payments

MainframeMaster • Northern (KY)

Hybrid
USD 120,000 - 180,000
Senior Staff Software Engineer
Senior Staff Software Engineer

Jobtailor • Atlanta (GA)

On-site
USD 150,000 - 190,000
Director-Infra Engineering
Director-Infra Engineering

American Express • Phoenix (AZ)

On-site
USD 190,000 - 260,000
Sr Software Engineer II
Sr Software Engineer II

American Express • New York (NY)

On-site
USD 180,000 - 230,000
Senior Service Assurance Engineer (Operations Engineer)
Senior Service Assurance Engineer (Operations Engineer)

American Express Services Europe Limited • Phoenix (AZ)

Hybrid
USD 110,000 - 190,000
6% Company Match on retirement savings plan
Free financial coaching
Comprehensive health insurance
+3
Senior Staff Engineer - Global Commercial Services
Senior Staff Engineer - Global Commercial Services

American Express Services Europe Limited • New York (NY)

Hybrid
USD 170,000 - 255,000
6% Company Match on retirement savings plan
Comprehensive medical and dental benefits
Flexible working model
+2
Data Center Operations Technician - Associate (night shift)
Data Center Operations Technician - Associate (night shift)

American Express • Phoenix (AZ)

On-site
USD 60,000 - 90,000
Site Reliability Engineer I
Site Reliability Engineer I

American Express • Town of Florida (NY)

On-site
USD 90,000 - 130,000