Stand out for this role — generate a tailored resume and cover letter in about a minute.
Umanist Staffing LLC is seeking a Platform Engineer for Pune to support a 24×7 AI platform operations environment. You will work on platform monitoring, incident troubleshooting, recovery, escalation and operational support for Kubernetes/OpenShift based systems.
The role requires a Bachelor’s degree in CS/IT/Engineering and 5–10 years of hands-on experience, with a willingness to work in a 24×7 shift. Pune-based candidates are preferred.
Experience: 5–10 Years
Location: Pune
Notice Period: Immediate to 45 Days(685)
Shift: 24×7 Operations
Qualification: Bachelor’s degree in CS, IT, Engineering or related field
We are hiring a Platform Engineer to support a 24×7 AI platform operations environment. The role involves platform monitoring, incident troubleshooting, recovery, escalation and operational support across Kubernetes/OpenShift-based platforms.
Monitor platform/application dashboards, alerts and operational systems.
Perform L1 incident detection, troubleshooting and recovery.
Troubleshoot basic Linux, networking, DNS, ports and connectivity issues.
Check Kubernetes/OpenShift nodes, pods, deployments and services.
Review logs and collect diagnostics for escalation.
Execute approved runbooks, SOPs and GitOps-based recovery procedures.
Restart/redeploy workloads and verify service recovery.
Create incident tickets, provide status updates and coordinate with L2/L3 teams.
Maintain proper shift handovers and operational documentation.
Platform Monitoring
Incident Troubleshooting & Recovery
Kubernetes & Red Hat OpenShift
Experience with Linux command-line
Experience withNetworking: IP, DNS, ports & connectivity
Experience withContainers knowledge
Experience in IT Operations / Infrastructure / Cloud / Application Support
Ability to follow SOPs, runbooks and technical procedures
Strong troubleshooting and analytical skills
Excellent communication & stakeholder management
Willingness to work in a 24×7 shift environment
Minimum 2 years of stability in an organization preferred
Git / GitOps
Grafana / Prometheus
Bash / Python scripting
ITIL / Incident Management
Cloud / Data Centre infrastructure
AI/ML or GPU infrastructure
SRE / Platform Engineering exposure
Strong communication skills are mandatory.
Pune/local candidates preferred.
Candidates should be comfortable working in a structured L1 operations and incident management environment.
2 Technical Rounds Client Round HR