Sr Technical Consultant - Azure Cloud, DevOps, Sql & Site Reliability Engineering(SRE)

JDA Software

Bengaluru

On-site

INR 1,800,000 - 3,200,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Night shift rotations
Global collaboration

Job summary

Blue Yonder seeks an experienced Cloud Platform Engineer to provide hands-on Azure, Kubernetes, and observability expertise. You will own incident triage, drive automation, and optimize platform operations across compute, storage, and networking.

The role includes on-call rotations and night-shift blocks as part of global coverage. You will collaborate with teams to ensure reliable provisioning, scalable workloads, and secure, efficient runtimes with a focus on automation and post-incident

Qualifications

  • Minimum 5-10 years of relevant work experience in Azure cloud infrastructure, DevOps, site reliability engineering or platform operations.
  • Hands‑on production experience with Microsoft Azure services and operational troubleshooting.
  • Strong Kubernetes troubleshooting skills covering workloads, events, services, networking, scheduling, scaling and resource management.

Responsibilities

  • Review and act on incidents, service requests, infrastructure requests and provisioning failures logged by implementation teams and platform users.
  • Own L2/L3 cloud-platform issues from initial triage through recovery, validation, communication, root-cause analysis and closure.
  • Troubleshoot Kubernetes pods, deployments, replicas, services, events, health checks, node pools, scheduling and resource constraints.
  • Diagnose container image-pull, startup, shutdown, registry-authentication, rollout and workload-reconciliation failures.
  • Investigate CPU, memory, GPU, quota, region, placement and capacity issues affecting platform workloads.
  • Troubleshoot workload and event-driven autoscaling using Kubernetes metrics and technologies such as KEDA.
  • Support Azure resource provisioning, provider operations, resource lifecycle workflows and reconciliation between desired and actual state.
  • Diagnose ACR, private endpoint, firewall, allowlist, CIDR, DNS, TLS and runtime-connectivity problems.
  • Trace logs, metrics and distributed telemetry from the workload through collectors, exporters and observability ingestion pipelines.
  • Support Elastic/Logstash/Kibana ingestion, index mappings, access, dashboards, alerts and environment filters.
  • Investigate operational issues involving MongoDB/Atlas, SQL, Redis, NFS and related managed data or storage services.
  • Support certificates, service principals, API keys, image-pull credentials and infrastructure credential rotation.
  • Support regional releases, environment configuration, disaster-recovery workflows and post-deployment validation.
  • Develop automation to improve platform reliability, reduce manual provisioning and shorten incident recovery time.
  • Maintain runbooks, dashboards, alerts, known-error records and diagnostic procedures for use by platform support teams.
  • Participate in capacity planning, incident reviews, change reviews, release readiness and an agreed production on-call rotation.

Skills

Azure Cloud infrastructure
Kubernetes
Observability
Automation
Scripting

Education

Bachelor's degree in Computer Science or related field

Tools

Azure Container Registry
Elastic Stack
Logstash
KEDA
Terraform
Helm

Job description

Overview

The BYX Platform provides shared cloud provisioning, compute, storage, observability and runtime capabilities for enterprise applications and implementation teams. We are seeking an astute Cloud Platform with a strong technical foundation and hands‑on experience in Microsoft Azure, Kubernetes, containers, observability, networking and production automation. Our current technical environment: Cloud and Compute: Microsoft Azure, Kubernetes, node pools, container workloads, Azure Container Registry, KEDA, Dask, autoscaling, CPU, memory and GPU capacity. Observability: Logs, metrics and traces; collectors and exporters; Elastic/Elasticsearch, Logstash, Kibana, dashboards, alerts and correlation identifiers. Data and Storage: MongoDB/Atlas, SQL, Redis, NFS, managed storage, replicas, regional capacity and service quotas. Networking and Security: DNS, CIDR, private endpoints, firewall rules, allowlists, TLS/certificates, SPNs, API keys, tokens, image-pull credentials and Git-hosted secrets. Operations and Automation: Linux, Git, CI/CD deployment workflows, infrastructure automation and scripting using Bash, Python or PowerShell.

What you’ll do
  • Review and act on incidents, service requests, infrastructure requests and provisioning failures logged by implementation teams and platform users.
  • Own L2/L3 cloud-platform issues from initial triage through recovery, validation, communication, root-cause analysis and closure.
  • Troubleshoot Kubernetes pods, deployments, replicas, services, events, health checks, node pools, scheduling and resource constraints.
  • Diagnose container image-pull, startup, shutdown, registry-authentication, rollout and workload-reconciliation failures.
  • Investigate CPU, memory, GPU, quota, region, placement and capacity issues affecting platform workloads.
  • Troubleshoot workload and event-driven autoscaling using Kubernetes metrics and technologies such as KEDA.
  • Support Azure resource provisioning, provider operations, resource lifecycle workflows and reconciliation between desired and actual state.
  • Diagnose ACR, private endpoint, firewall, allowlist, CIDR, DNS, TLS and runtime-connectivity problems.
  • Trace logs, metrics and distributed telemetry from the workload through collectors, exporters and observability ingestion pipelines.
  • Support Elastic/Logstash/Kibana ingestion, index mappings, access, dashboards, alerts and environment filters.
  • Investigate operational issues involving MongoDB/Atlas, SQL, Redis, NFS and related managed data or storage services.
  • Support certificates, service principals, API keys, image-pull credentials and infrastructure credential rotation.
  • Support regional releases, environment configuration, disaster-recovery workflows and post-deployment validation.
  • Develop automation to improve platform reliability, reduce manual provisioning and shorten incident recovery time.
  • Maintain runbooks, dashboards, alerts, known-error records and diagnostic procedures for use by platform support teams.
  • Participate in capacity planning, incident reviews, change reviews, release readiness and an agreed production on-call rotation.
What we are looking for
  • Minimum 5-10 years of relevant work experience in Azure cloud infrastructure, DevOps, site reliability engineering or platform operations.
  • This role includes Rotational Shifts (Night shifts of 2 months in Year).
  • Hands‑on production experience with Microsoft Azure services and operational troubleshooting.
  • Strong Kubernetes troubleshooting skills covering workloads, events, services, networking, scheduling, scaling and resource management.
  • Experience with a container registry; Azure Container Registry experience is highly relevant.
  • Practical observability experience across logs, metrics, traces, dashboards and alerts.
  • Experience with Elastic Stack components—Elasticsearch, Logstash and Kibana—or a closely comparable platform.
  • Working knowledge of DNS, TLS/certificates, CIDR, firewalls, proxies, load balancing and private networking.
  • Strong Linux administration, application log analysis and production incident‑troubleshooting skills.
  • Scripting experience with Bash, Python or PowerShell for diagnostics and operational automation.
  • Experience troubleshooting CI/CD deployments, configuration changes and failed rollouts.
  • Working knowledge of Git, service identities and secrets‑management fundamentals.
  • Working knowledge of MongoDB/Atlas, Redis, SQL or NFS operations is preferred.
  • Experience with infrastructure as code such as Terraform, Bicep or ARM templates and Kubernetes packaging such as Helm is preferred.
  • Experience with incident response, Root Cause Analysis, post‑incident reviews and controlled production changes.
  • Strong collaboration and communication skills with the ability to work across application, security, networking and vendor teams.
  • Willingness to participate in a scheduled production on‑call rotation and planned out‑of‑hours changes when required.

Our Values If you want to know the heart of a company, take a look at their values. Ours unite us. They are what drive our success – and the success of our customers. Does your heart beat like ours? Find out here:

Core Values All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status.

Who are we? We are a proven, passionate bunch of disruptors. Our work is all about tapping into your potential so we can deliver the best solutions and customer experiences on the planet. Collaboration, respect, and a great work‑life balance earned us the title of "Best Place to Work- Employees' Choice" by Glassdoor. Our people are smart, creative, rock stars with over 400 patents and 10,000 people years of domain expertise.

What do we do? Blue Yonder is the world leader in digital supply chain and omni-channel commerce fulfillment. Our intelligent, end-to-end platform enables retailers, manufacturers and logistics providers to seamlessly predict, pivot and fulfill customer demand. With Blue Yonder, you can make more automated, profitable business decisions that deliver greater growth and re‑imagined customer experiences.

Blue Yonder - Fulfill your Potential. blueyonder.com “Blue Yonder” is a trademark or registered trademark of Blue Yonder, Inc. Any trade, product or service name referenced in this document using the name “Blue Yonder” is a trademark and/or property of Blue Yonder, Inc. Blue Yonder, Inc. 15059 N Scottsdale Rd, Ste 400 Scottsdale, AZ 85254

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer I - Kubernetes,ArgoCD
Staff Software Engineer I - Kubernetes,ArgoCD

JDA Software • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Associate Technical Consultant - Cloud - Devops,Kubernetes,Terraform,Scripting
Associate Technical Consultant - Cloud - Devops,Kubernetes,Terraform,Scripting

JDA Software • Coimbatore District

On-site
INR 900,000 - 1,300,000
Enterprise Technical Architect Int cloud - Sre,Performance Engineering,Azure
Enterprise Technical Architect Int cloud - Sre,Performance Engineering,Azure

JDA Software • Bengaluru

On-site
INR 2,500,000 - 4,200,000
Software Engineer – Python, SQL, Microservices, Kafka/Rabbit MQ, Azure
Software Engineer – Python, SQL, Microservices, Kafka/Rabbit MQ, Azure

JDA Software • Bengaluru

On-site
INR 1,200,000 - 2,100,000
Sr DevOps Engineer I - AWS,Azure,Kubernetes,Terraform
Sr DevOps Engineer I - AWS,Azure,Kubernetes,Terraform

JDA Software • Hyderabad

On-site
INR 2,600,000 - 4,200,000
Staff Software Engineer I - NodeJS, Microservices, Cloud & Typescript/JavaScript
Staff Software Engineer I - NodeJS, Microservices, Cloud & Typescript/JavaScript

JDA Software • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Support Engineer 2 (WMS/Warehouse Management )
Support Engineer 2 (WMS/Warehouse Management )

JDA Software • Coimbatore District

On-site
INR 1,200,000 - 1,800,000
Staff Software Engineer I - Devops architecture,IaC, Kubernetes,Python
Staff Software Engineer I - Devops architecture,IaC, Kubernetes,Python

JDA Software • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Technical Consultant - Platform Full Stack & AI Applications Engineer
Technical Consultant - Platform Full Stack & AI Applications Engineer

JDA Software • Coimbatore District

On-site
INR 900,000 - 1,500,000
Support Engineer 2 (BY WMS/Blueyonder Warehouse Management )
Support Engineer 2 (BY WMS/Blueyonder Warehouse Management )

JDA Software • Coimbatore District

On-site
INR 1,200,000 - 1,800,000