Linux Automation Engineer

Netweb Technologies India Ltd.

Faridabad District

On-site

INR 1,400,000 - 2,100,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Netweb Technologies India Ltd. is looking for an experienced Linux Automation Engineer to design, develop, and operate enterprise-scale Linux infrastructure automation, HPC clusters, and monitoring systems.

You will automate provisioning, OS deployment, configuration management, patching, and lifecycle management, while driving best practices across DevOps and software development teams. You will build scalable backend systems, REST APIs, centralized monitoring, and management dashboards with a

Qualifications

  • Master-level knowledge of Linux Internals (RHEL/ Rocky Linux/ Ubuntu) and high-performance networking.
  • Direct experience deploying and managing Slurm/PBS Pro with parallel storage (Lustre/ GPFS).
  • Proficiency in Python or Go and Bash scripting for automation.
  • Enterprise-level Terraform and Ansible/AWX implementations.

Responsibilities

  • Design, develop, and manage enterprise-scale Linux infrastructure automation solutions.
  • Lead end-to-end infrastructure automation using IaC and CI/CD pipelines.

Skills

Linux Internals
HPC Stack
Python
Go
Bash
Terraform
Ansible/AWX
InfiniBand
REST APIs
Monitoring

Tools

OpenStack
VMware NSX-T
KVM
Prometheus
Grafana
ELK

Job description

Experienced Linux Automation Engineer responsible for designing, developing, and managing enterprise-scale Linux infrastructure automation solutions, HPC cluster management platforms, and monitoring systems. The role focuses on automating server provisioning, operating system deployment, configuration management, patching, infrastructure lifecycle management, and cluster orchestration to improve operational efficiency and scalability. Based on the job description, the position also involves leading technical initiatives, collaborating with cross-functional teams, and driving best practices in software development, DevOps, and infrastructure automation.

The position requires strong expertise in Linux administration, Python development, infrastructure automation tools, and observability platforms. The engineer is expected to build and enhance scalable backend systems, REST APIs, centralized monitoring solutions, and management dashboards while ensuring high availability, security, and maintainability of infrastructure environments. Responsibilities also include implementing CI/CD pipelines, integrating enterprise platforms, and supporting HPC environments through cluster provisioning, resource monitoring, scheduler integration, and performance analytics.

Key Responsibilities:
  • HPC & Parallel Storage Design: Architect, tune, and maintain high-performance, multi-petabyte scale Parallel File Systems (PFS) like Lustre, IBM Spectrum Scale (GPFS), or BeeGFS. Optimize data-in-transit pipelines over low-latency InfiniBand or RoCE fabrics.
  • Private Cloud Infrastructure: Design, configure, and bootstrap custom on-premises private clouds using OpenStack, VMware NSX-T, or KVM-based hypervisors (Proxmox/RHEV).
  • Custom Tooling & Scripting: Act as a developer within operations. Write advanced, production-grade automation scripts and internal CLI tools in Python or Go to interface directly with cloud APIs and manage resources.
  • End-to-End Infrastructure Automation: Champion Infrastructure as Code (IaC) by creating modular Terraform plans and massive Ansible Playbooks/Roles to automate compute provisioning, job schedulers, and node validation.
  • Hands-on DevOps Pipelines: Construct and manage multi-stage Jenkins or GitLab CI pipelines to test and securely deploy system configuration changes across development, staging, and cluster nodes.
1. Technical Leadership:
  • Lead and mentor a team of Full Stack Developers, Backend Developers, and Automation Engineers.
  • Define technical architecture, coding standards, and development best practices.
  • Conduct code reviews, design reviews, and technical feasibility assessments.
  • Drive product roadmap discussions and technical decision-making.
  • Collaborate with Product, Infrastructure, Validation, and Support teams.
2. Platform & Product Development:
  • Architect and develop infrastructure management platforms, including:
  • Infrastructure Lifecycle Management Systems
  • Asset & Inventory Management Solutions
  • Capacity Planning & Analytics Platforms
  • Firmware & Driver Compliance
  • Automated Validation & Benchmarking
  • Develop integrations with Linux-based infrastructure and enterprise platforms.
4. Monitoring & Observability:
  • Integrate monitoring tools such as Prometheus, Grafana, Open Telemetry, ELK, Redfish, SNMP, and IPMI.
  • Build predictive monitoring, alerting, and analytics capabilities.
Develop and enhance cluster management capabilities including:
  • Node Discovery & Registration
  • Job Monitoring
  • Health & Performance Analytics
  • Cluster Lifecycle Management
6. Software Engineering:
  • Design scalable backend architectures and REST APIs.
  • Guide development of web-based dashboards and management portals.
  • Implement CI/CD pipelines and DevOps best practices.
  • Ensure high availability, scalability, security, and maintainability of developed solutions.
Required Technical Skills
  • Linux Internals: Master-level knowledge of RHEL, Rocky Linux, or Ubuntu Server (advanced kernel tuning, memory management, and high-performance network configurations).
  • HPC Stack: Direct experience deploying and managing Slurm/PBS Pro alongside a verified track record handling parallel storage layers (Lustre, GPFS).
  • Languages: Advanced proficiency in Python or Go, and robust Bash shell scripting.
  • Automation Frameworks: Enterprise-level implementation of Terraform and Ansible/AWX.
  • Fabrics & Interconnects: Comprehensive understanding of InfiniBand routing, Subnet Managers, and network topology tuning for high-performance computing.
  • RHEL, Rocky Linux, Ubuntu, Alma Linux
  • Performance Tuning & Troubleshooting
  • Shell Scripting
Monitoring & Observability
  • Zabbix/Nagios
Front-End Awareness
  • ReactJS / Angular (working knowledge to guide development teams)
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. HPC ENGINEER
Sr. HPC ENGINEER

Cognizant • Hyderabad

On-site
INR 800,000 - 1,200,000
Senior Cloud Infrastructure Engineer
Senior Cloud Infrastructure Engineer

Innovalus Technologies • Mumbai

On-site
INR 400,000 - 650,000
System Engineer- Linux/Storage
System Engineer- Linux/Storage

GSPANN Technologies, Inc. • Hyderabad, Pune District

On-site
INR 1,500,000 - 2,600,000
Software - Platform Automation Architect
Software - Platform Automation Architect

Dayal Group • Meerut

On-site
INR 1,200,000 - 1,800,000
Lead HPC Engineer
Lead HPC Engineer

Clovertex • Hyderabad

On-site
INR 2,000,000 - 3,000,000
Devops Engineer
Devops Engineer

Yotta Infrastructure • Navi Mumbai, Delhi

On-site
INR 1,200,000 - 1,800,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Colruyt Group India • Hyderabad

Hybrid
INR 3,000,000 - 4,000,000
Linux System Administrator
Linux System Administrator

SISL Global • Chennai District

On-site
INR 800,000 - 1,200,000
Senior Devops Engineer
Senior Devops Engineer

Bounteous • Gurugram District, Bengaluru

On-site
INR 1,800,000 - 2,400,000
Senior Staff Engineer, DevOps
Senior Staff Engineer, DevOps

Sierra Wireless • India

On-site
INR 2,500,000 - 4,000,000