Infrastructure Engineer III

JPMorgan Chase & Co.

Plano (TX)

On-site

USD 120,000 - 170,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPMorgan Chase & Co. is seeking an Infrastructure Engineer III to drive storage and data-management excellence across on-premises and cloud environments. You will own end-to-end administration, resilience, and automation of the storage estate, supporting global services with a structured shift rotation.

This role emphasizes AI-enabled observability, incident leadership, and collaboration with cross-functional teams to meet reliability targets and auditability standards.

Qualifications

  • Formal training or certification in infrastructure engineering with 3+ years of applied experience.
  • Experience using AI capabilities in infrastructure with validation and security awareness.
  • Ability to review AI recommendations and ensure resiliency and auditability.
  • Strong storage fundamentals including RAID/erasure coding, replication, and latency.
  • Hands-on with major storage ecosystems (e.g., NetApp, Dell EMC, Pure Storage, Ceph).
  • Scripting proficiency in Python, Go, or Bash; solid Linux basics.
  • Experience with observability stacks (Prometheus, Grafana, Splunk, Datadog).
  • Proven incident management and on-call experience.
  • Practical AI/DataOps skills for anomaly detection and CI/CD integration.

Responsibilities

  • Apply storage infra expertise to projects at moderate scope with end-to-end monitoring and resilience.
  • Leverage AI capabilities to accelerate analysis, capacity documentation, and automation.
  • Improve delivery by identifying recurring issues and ensuring traceability and security.
  • Operate and optimize block, file, and object storage across on-prem and cloud.
  • Lead incident response for storage outages and perform root cause analyses.
  • Build automation for provisioning, patching, upgrades, replication, and backups.
  • Create runbooks and on-call procedures to support team knowledge continuity.
  • Partner with multiple teams to meet workload reliability targets.
  • Implement AI-driven observability and safe LLM-powered workflows.

Skills

Infrastructure engineering
AI in infrastructure
Python/Go/Bash scripting
Linux fundamentals
Observability stacks
Incident management
Storage fundamentals
SAN/NAS/storage

Education

Professional infrastructure certification

Tools

Prometheus/Grafana
Elastic/OpenSearch
Splunk
Datadog/OpenTelemetry

Job description

Job Description

You belong to the top echelon of talent in your field. At JPMorganChase, infrastructure isn't just a foundation — it's a competitive advantage. This is your opportunity to bring deep storage expertise to a team that operates at global scale, where your contributions directly impact the stability and performance of critical financial services.

As an Infrastructure Engineer III at JPMorganChase within Enterprise Technology (Infrastructure Platforms), you apply strong knowledge of software, applications, and technical processes within the infrastructure engineering discipline. In this role, you will hold end-to-end accountability for the administration, stability, and resilience of the storage technology estate — spanning reactive incident management through to proactive automation and toil reduction. You will operate on a structured shift rotation, including weekend day coverage, to ensure continuity of service and operational excellence across the infrastructure landscape.

Job responsibilities
  • Apply technical knowledge and problem-solving methodologies to storage infrastructure projects of moderate scope, ensuring end-to-end monitoring, performance, and resilience of storage services running at scale
  • Use enterprise-authorized AI capabilities to accelerate infrastructure analysis, monitoring, and capacity documentation, validating outputs and handling operational data according to sensitivity and security requirements
  • Apply reuse-first, AI-assisted practices within delivery and automation routines to identify recurring issues, improve remediation workflows, and ensure changes are traceable, auditable, and aligned to resiliency and security expectations
  • Operate and enhance block, file, and object storage platforms across on-premises and cloud environments, including performance tuning, capacity planning, lifecycle management, and resiliency testing such as failover and disaster recovery validation
  • Lead incident response for storage outages and performance degradations, drive root cause analyses, and implement preventative actions to reduce recurrence
  • Build and maintain automation for provisioning, patching, upgrades, replication, backup and restore, and compliance checks to reduce toil and improve operational consistency
  • Create and maintain runbooks, escalation paths, and standardized operational procedures to support on-call readiness and team knowledge continuity
  • Partner with infrastructure, network, operating system, database, and application teams to meet workload requirements and reliability targets
  • Implement AI-driven observability and AIOps capabilities — including telemetry correlation, anomaly and regression detection, and large language model-assisted incident and runbook workflows — with a focus on accuracy, auditability, and safe rollout
  • Own and continuously improve service level objectives, service level indicators, error budgets, and on-call readiness for storage services
Required qualifications, capabilities, and skills
  • Formal training or certification on infrastructure engineering concepts and 3+ years applied experience
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support infrastructure engineering workflows, with strong validation habits and awareness of data sensitivity
  • Ability to review and validate AI-assisted recommendations before implementation, escalating when uncertain and ensuring outcomes align to resiliency, security, and auditability expectations
  • Strong knowledge of storage fundamentals including RAID and erasure coding, replication, snapshots, tiering and caching, IOPS and latency, multipathing, SAN and NAS, and object storage semantics
  • Hands‑on experience with at least one major storage ecosystem such as NetApp, Dell EMC PowerStore or Isilon, Pure Storage, Hitachi, Ceph, IBM, or cloud‑native storage services
  • Strong scripting or programming proficiency in one or more of Python, Go, or Bash
  • Solid Linux fundamentals including system performance, networking basics, and kernel and storage‑stack concepts
  • Experience with observability stacks such as Prometheus and Grafana, Elastic or OpenSearch, Splunk, Datadog, or OpenTelemetry
  • Proven incident management skills and ability to operate effectively within an on‑call rotation
  • Practical skills in AI and data operations including anomaly detection, forecasting, correlation, classification, feature extraction, and integrating AI into production tooling and continuous integration and delivery pipelines with safe large language model use, guardrails, and human‑in‑the‑loop review
Preferred qualifications, capabilities, and skills
  • Experience with Kubernetes storage using the Container Storage Interface, stateful workloads, and container platform operations
  • Proficiency with Infrastructure as Code tools such as Terraform or CloudFormation, and configuration management tools such as Ansible, Chef, or Puppet
  • Familiarity with streaming and queue tooling for telemetry and event pipelines such as Kafka
  • Experience with IT service management and event management platforms such as ServiceNow
  • Knowledge of backup and disaster recovery products and strategy design, including recovery point objective and recovery time objective tradeoffs
  • Experience with security controls for data platforms including key management services, hardware security modules, secrets management, and key rotation
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure Engineer III
Infrastructure Engineer III

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 110,000 - 160,000
Infrastructure Engineer III
Infrastructure Engineer III

JPMorganChase • Plano (TX)

On-site
USD 120,000 - 180,000
Lead Infrastructure Engineer - Storage
Lead Infrastructure Engineer - Storage

JPMorgan Chase & Co. • Columbus (OH)

On-site
USD 140,000 - 190,000
Infrastructure Engineer III
Infrastructure Engineer III

JPMorganChase • San Francisco (CA)

On-site
USD 150,000 - 205,000
Lead Infrastructure Engineer - Storage
Lead Infrastructure Engineer - Storage

JPMorganChase • Houston (TX)

Hybrid
USD 130,000 - 190,000
Sr. Lead Infrastructure Engineer - Storage SRA
Sr. Lead Infrastructure Engineer - Storage SRA

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 140,000 - 190,000
Lead Infrastructure Engineer - Storage
Lead Infrastructure Engineer - Storage

Fairygodboss • Houston (TX)

On-site
USD 140,000 - 190,000
Infrastructure Engineer III - Site Reliability
Infrastructure Engineer III - Site Reliability

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 110,000 - 150,000
Storage Infra Engineer III — AI-Driven Resilience Expert
Storage Infra Engineer III — AI-Driven Resilience Expert

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 120,000 - 170,000
Lead Infrastructure Engineer - Storage
Lead Infrastructure Engineer - Storage

JPMorganChase • Columbus (OH)

On-site
USD 140,000 - 190,000