ESSENTIAL DUTIES & RESPONSIBILITIES
- Lead the infrastructure engineering team across shifts, covering the Wintel, Messaging, SCCM and Intune, Linux, Cloud Platform, and Storage and Backup roles, keeping each discipline represented on every shift within the 24x7 operating model.
- Set performance expectations, conduct evaluations, and build technical depth and cross-skilling so that no discipline depends on a single engineer; support recruitment, onboarding and a current skills matrix.
- Run a documented handover at every shift boundary with the other IT Infrastructure Managers, transferring ownership of open incidents, in-flight changes, active priority incidents and pending actions explicitly.
- Report operations to the Sr. Manager IT Infrastructure and US IT Infrastructure leadership on availability, incidents, change, patch compliance, capacity, cloud cost and risk, and coordinate infrastructure effort during priority and major incidents.
On-Premises Compute, Virtualization and Storage Operations
- Oversee day-to-day operation of the Nutanix and hyperconverged infrastructure estate, VMware virtualization, and on-premises compute and server hardware.
- Oversee Dell Isilon and HPE storage platforms, including provisioning, capacity, performance, replication and tiering.
- Maintain firmware, hypervisor and storage software currency, and plan capacity and end-of-life replacement ahead of need.
- Manage OEM and partner vendors against contracted support, and provide input to capital and operating budget planning for hardware refresh and support renewals.
Cloud Platform Operations – AWS and Azure
- Support oversight of day-to-day operation of AWS and Microsoft Azure workloads, covering compute, storage, virtual networking, load balancing, and platform monitoring and alerting.
- Maintain cloud identity and access integration with Active Directory and Entra ID, including AWS IAM roles and policies, Azure RBAC, federation and privileged access controls.
- Maintain landing zone, subscription and account conventions, resource tagging, naming standards and guardrails in line with standards agreed with US infrastructure leadership.
- Manage cloud cost and consumption, including rightsizing, commitment planning and orphaned resource cleanup, and support migration, modernization and exit activity along with hybrid connectivity, DNS and certificate services.
Microsoft 365, Identity and Endpoint Management
- Manage operation of Microsoft 365 services, including messaging, and the hybrid identity estate across Active Directory and Entra ID, with conditional access, multi-factor authentication and privileged identity management.
- Manage license assignment and entitlement coordinate asset and license records with Service Asset and Configuration Management.
- Manage endpoint management through SCCM and Intune, including image and configuration baselines, application packaging and deployment, and co-management.
- Maintain endpoint compliance and protection coverage with Information coordinate tenant and service changes.
Patching, Security Compliance, Resilience and Audit Readiness
- Execute the defined patching cycle to published standards and maintenance windows across servers, endpoints, hypervisors, storage and cloud workloads, with every exception carrying documented risk acceptance and a remediation date.
- Maintain configuration and hardening baselines, remediate drift, and partner with Information Security on vulnerability remediation, certificate lifecycle and privileged access.
- Maintain backup and disaster recovery, including immutable or offsite copies, documented recovery time and recovery point objectives, and periodic restore and DR testing with identified gaps closed.
- Operate change through the approved process with CAB submission, tested rollback and post-implementation review, and maintain monitoring coverage, configuration item accuracy in ServiceNow and automation of routine work.
- Support internal and external audit with evidence, control walkthroughs and corrective action closure, and maintain runbooks, SOPs and the infrastructure risk register.
JOB COMPETENCIES (Skills & Abilities)
- Team leadership and shift management: Leads a multi-discipline engineering team across a 24x7 roster, sets expectations, develops technical depth and cross-skilling, and keeps coverage intact through absence and surge.
- On-premises technical breadth: Practical command of Nutanix and hyperconverged infrastructure, VMware, on-premises compute and server hardware, Dell Isilon and HPE storage, Windows Server and Linux.
- Cloud platform command: Working command of AWS and Microsoft Azure operations, including compute, storage, virtual networking, IAM and RBAC, native backup and resilience services, tagging and guardrails, and monitoring across both platforms.
- Hybrid identity and collaboration: Microsoft 365, Active Directory and Entra ID in a hybrid configuration, including conditional access, multi-factor authentication and privileged identity management.
- Endpoint management: SCCM and Intune configuration baselines, application deployment, compliance policy and co-management.
- Automation and infrastructure as code: Uses scripting and declarative tooling to remove repetitive operational work and reduce configuration variance, rather than scaling by headcount.
- Patch and compliance rigor: Treats the patching cycle as a commitment rather than a best effort, governs exceptions with documented risk acceptance, and sustains compliance month to month rather than recovering it before an audit.
- Resilience and recovery discipline: Designs for recovery against stated objectives and proves it by test, rather than treating backup success reports as evidence of recoverability.
- Change discipline: Operates change through the approved process with assessed risk, tested rollback and post-implementation review, and does not permit undocumented change.
- Monitoring and observability judgement: Builds coverage that detects failure early and tunes alerting so that what reaches an engineer is actionable.
- Cost and capacity management: Manages cloud consumption and on-premises capacity against budget and growth, with variances explained rather than discovered.
- Security partnership: Works with Information Security on vulnerability remediation, hardening, endpoint protection coverage, certificate lifecycle and privileged access.
- Documentation and audit discipline: Maintains runbooks, SOPs and audit-ready evidence, and is comfortable supporting control walkthroughs with auditors.
- Handover discipline: Structures shift so that ownership of open work transfers explicitly and context is not lost at a shift or regional boundary.
- Stakeholder engagement: Works effectively with leadership, security, applications and business stakeholders across time zones, and manages expectations with both technical and senior audiences.
- Composure under pressure: Coordinates infrastructure effort during priority and major incidents, keeps stakeholders informed and holds situational control.
- Structured problem-solving: Applies disciplined root cause analysis, converts recurring incidents into permanent engineering fixes, and resists symptomatic repair.
- Communication: Clear, concise written and verbal communication in English across executive, technical and end-user audiences, including operational reporting to global leadership.
- Flexibility: Works effectively within a 24x7 operating model, including shift rotation, maintenance windows outside standard hours and overlap with US business hours
MINIMUM QUALIFICATIONS (Knowledge & Experience)
- Engineering, Computer Science, or a related field. (Required)
- 10+ years of experience in IT infrastructure engineering and operations across virtualisation, storage, identity and cloud. (Required)
- 3+ years in a formal supervisory or management role leading infrastructure engineers, including shift-based teams. (Required)
- Demonstrated experience operating Nutanix and hyperconverged infrastructure, and VMware virtualization, at enterprise scale. (Required)
- Demonstrated experience operating enterprise storage and backup, including Dell Isilon and HPE storage platforms, with capacity, performance and restore accountability. (Required)
- Demonstrated experience administering Windows Server and Linux estates, and on-premises compute and server hardware from major OEMs. (Required)
- Demonstrated experience operating production workloads in AWS, including compute, storage, virtual networking, IAM and native backup services. (Required)
- Demonstrated experience operating production workloads in Microsoft Azure, including IaaS, virtual networking, RBAC, Azure Backup and Azure Site Recovery. (Required)
- Experience with hybrid cloud connectivity and identity integration, including VPN or direct interconnect, certificate services and federation with Active Directory and Entra ID. (Required)
- Experience managing cloud cost and consumption, including rightsizing, reserved or savings-plan commitments, tagging governance and budget reporting. (Required)
- Demonstrated experience with Microsoft 365, Active Directory and Entra ID in a hybrid identity environment, including conditional access and privileged identity management. (Required)
- Demonstrated experience with endpoint management through SCCM and Intune, including configuration baselines and application deployment. (Required)
- Demonstrated experience owning a patch management cycle across servers, endpoints and cloud workloads, including exception governance and compliance reporting against defined standards. (Required)
- Experience with configuration hardening baselines and vulnerability remediation in partnership with a security function. (Required)
- Experience owning backup and disaster recovery, including documented RTO and RPO for business-critical services and periodic recovery testing. (Required)
- Experience with infrastructure monitoring and observability tooling across on-premises and cloud estates. (Required)
- Working experience with automation and infrastructure as code, such as PowerShell, Bash, Ansible or Terraform. (Required)
- Experience operating infrastructure change through a formal change management process, including Change Advisory Board and post-implementation review. (Required)
- Experience operating within a 24x7 or shift-based support model, including structured between shifts and regions. (Required)
- Experience working with geographically distributed leadership and teams across multiple time zones, including follow-the-sun models. (Required)
- Experience supporting internal or external audit, including evidence preparation, control walkthroughs and corrective action closure. (Required)
- Working knowledge of enterprise networking fundamentals, including DNS, DHCP, load balancing and firewall change requests. (Required)
- Working knowledge of ServiceNow for incident, change and configuration management. (Preferred)
- Familiarity with IT infrastructure operations delivered from an India-based capability center or offshore delivery organizations. (Preferred)
- ITIL 4 Foundation certification. (Preferred)
- Cloud certification such as AWS Certified Solutions Architect or SysOps Administrator, or Microsoft Azure Administrator (AZ-104) or Azure Solutions Architect (AZ-305). (Preferred)
- Platform certification such as Nutanix NCP, VMware VCP, Microsoft 365 administration, or Red Hat RHCSA. (Preferred)
- Experience in regulated or audit-driven environments, including exposure to SOX ITGC, NIST, ISO/IEC 27001 or equivalent control frameworks. (Preferred)
Disclaimer: This job description indicates in general terms, the type and level of work performed as well as the typical responsibilities of employees in this classification and it may be changed by management at any time. Other duties may also apply. Nothing in this job description changes the at-will employment relationship existing between the Company and its employees.
Why Work at GIA Services Pvt. Ltd.?
Working at GIA Services means you can express your passion for education, gemstones, business, and make a difference. GIA Services seeks skilled professionals that will work as a team to fulfill its mission of building new capabilities, modernizing processes, and enabling future scalability.
Pre-Employment Requirements
- Applicants and new hires are subject to background checks as per the company policy, this will include reference checks, address, education verification, police verification and work history verification.
Equal Opportunity
GIA Services is an equal opportunity employer and does not discriminate against persons on the basis of race, religion, national origin, sexual orientation, gender, marital status, age, disability, or veteran’s status.
Due to the high volume of resumes we receive daily, we regret that we cannot personally contact every applicant. Candidates whose experience closely matches our qualification requirements will be contacted by GIA Services Human Resources department and invited for an interview. Thank you for your interest.