Senior DevOps Engineer (AI, GPU & Cloud Infrastructure)

Wisewit Solutions Private Limited

Hinoba-an

On-site

PHP 1,200,000 - 1,800,000

Full time

9 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Wisewit Solutions Private Limited seeks a Senior DevOps Engineer with 5–7 years of hands-on experience to lead end-to-end infrastructure, automation pipelines, and deployment frameworks for AI/ML, real-time voice, and data-intensive apps.

You will own CI/CD pipelines, container orchestration, and data/messaging clusters, while collaborating with AI engineers and project teams to ensure reliability, scalability, and security across cloud and on-prem environments.

Qualifications

  • 5–7 years hands-on DevOps, Cloud, and Infrastructure Engineering experience.

Responsibilities

  • Cloud & hardware infrastructure management for GPU/CPU resources.

Skills

Linux administration
Networking
Shell/Python scripting
Docker
Kubernetes
CI/CD pipelines
Kafka
ELK/OpenSearch
PostgreSQL
MongoDB
Hardware estimation
Traffic engineering

Tools

Terraform
Ansible
Prometheus
Grafana
Jenkins
GitHub Actions
GitLab CI/CD
Azure DevOps
NVIDIA GPU tooling

Job description

We are seeking an experienced, proactive, and self-driven Senior DevOps Engineer (5–7 Years Experience) to manage end-to-end cloud and on-premise infrastructure, automation pipelines, and deployment frameworks for complex AI/ML, real-time voice, and data-intensive applications.In this role, you will bridge the gap between AI engineering, core software development, and infrastructure operations. You will take full ownership of designing, building, and maintaining robust CI/CD pipelines, container orchestration environments, data/messaging clusters (Kafka, ELK, MongoDB), and specialized GPU/CPU hardware configurations.

  • Experience Required: 5 to 7 Years
  • Reporting & Collaboration: Direct collaboration with Engineering Leads, AI/ML Engineers, and Project Managers/Scrum Leads.
Key Responsibilities
1. Cloud & Hardware Infrastructure Management (GPU & CPU)
  • Resource Optimization: Oversee end-to-end hardware resource management, from GPU/CPU procurement estimations based on traffic engineering to server topology design and optimization.
  • AI/ML Workload Support: Configure, tune, and maintain compute environments required to run Large Language Models (LLMs), Computer Vision (CV) solutions, and Real-Time Voice Bots using NVIDIA GPU, CUDA, and inference server platforms.
  • Cloud Operations: Architect, manage, and scale cloud infrastructure across major platforms (AWS, Azure, or GCP, OCI).
2. Infrastructure as Code (IaC) & Container Orchestration
  • Automation: Implement and maintain Infrastructure as Code using Terraform, Ansible, or equivalent automation frameworks.
  • Containerization: Architect multi-container environments using Docker and manage production-grade cluster orchestration with Kubernetes (including Helm charts).
  • Network & Security: Manage cloud networking, load balancing, DNS, SSL/TLS, reverse proxies (Nginx), IAM, access controls, and security configurations.
3. Data Platforms & Messaging Infrastructure
  • Streaming & Search: Set up, configure, and maintain high-throughput streaming and logging environments (Kafka, ELK/OpenSearch), managing topics, partitioning, and multi-node cluster setups.
  • Database Management: Manage and scale database clusters, specifically MongoDB, ensuring high availability, backups, and disaster recovery.
4. CI/CD & Enterprise Integrations
  • Pipeline Automation: Design, implement, and standardise automated CI/CD pipelines (via GitHub Actions, GitLab CI/CD, Jenkins, or Azure DevOps) to replace manual deployment processes.
  • Integrations: Assist in establishing API/system integrations with client enterprise platforms, such as ERP systems and accounting software (e.g., Tally).
5. Delivery, Reliability & Operational Leadership
  • Ownership & Sprint Tracking: Drive daily priority management, task allocation, and tracking for infrastructure tasks; bridge technical teams and project managers/scrum leads.
  • Monitoring & Alerting: Set up end-to-end monitoring and logging (Prometheus, Grafana, ELK) to ensure optimum system availability, performance, and log tracing.
  • Disaster Recovery: Establish and enforce best practices for backup, failover, system availability, and disaster recovery.
Required Technical Skills & Qualifications
  • Experience: 5–7 years of hands-on experience in DevOps, Cloud, and Infrastructure Engineering.
  • OS & Scripting: Strong expertise in Linux system administration, networking, and Shell/Python scripting.
  • Containers & Orchestration: Advanced experience with Docker and Kubernetes management.
  • CI/CD: Expertise in setting up enterprise-level continuous integration and deployment pipelines.
  • Data & Messaging Systems: Direct experience in managing Kafka clusters, ELK Stack, PostgreSQL and MongoDB databases.
  • Hardware & Traffic Engineering: Proven track record in capacity planning, traffic management, and hardware estimation for data-heavy workloads.
Preferred & AI-Specific Skills
  • GPU Infrastructure: Direct experience managing NVIDIA GPU environments, CUDA setups, vector databases, and model serving/inference environments for LLM or AI/ML workloads.
  • Cloud & IaC: Deep knowledge of at least one major cloud platform (AWS, Azure, GCP) and IaC tools (Terraform, Ansible).
  • Enterprise Integrations: Experience connecting client-side ERP systems or third-party enterprise platforms.
  • Agile Leadership: Strong sprint management capability and experience operating in fast-paced or startup engineering environments.
What We Look For
  • Ownership Mindset: Takes full responsibility for infrastructure stability, performance, and deployments across dev, staging, and production environments.
  • Problem-Solving Skills: Strong capability to troubleshoot complex deployment, networking, or hardware-level issues independently.
  • Communication & Collaboration: Clear visibility into sprint tasks, effective cross-team communication, and documentation skills.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Devops Engineer
Senior Devops Engineer

CodeRound • Hinoba-an

On-site
PHP 800,000 - 1,200,000
Senior Platform Engineer
Senior Platform Engineer

Our Clients • Pasay

On-site
PHP 1,200,000 - 1,800,000
Lead DevOps Engineer
Lead DevOps Engineer

Skit • Hinoba-an

On-site
PHP 1,500,000 - 2,600,000
Senior AWS DevOps & Platform Engineer
Senior AWS DevOps & Platform Engineer

Neara • Hinoba-an

On-site
INR 1,800,000 - 2,400,000
Sr. DevOps Engineer
Sr. DevOps Engineer

NCS Philippines • Philippines

On-site
PHP 1,200,000 - 1,800,000
Senior DevOps/Lead DevOps
Senior DevOps/Lead DevOps

Goodfit • Hinoba-an

On-site
PHP 1,800,000 - 3,200,000
AI/ DevOps Engineer
AI/ DevOps Engineer

Universal Access and Systems Solutions Inc. • Angeles

On-site
PHP 600,000 - 1,000,000
Senior DevOps Engineer
Senior DevOps Engineer

Cloud Bridge • Philippines

On-site
PHP 1,200,000 - 2,400,000
Senior DevOps Engineer (Hybrid Infrastructure & Cloud)
Senior DevOps Engineer (Hybrid Infrastructure & Cloud)

YAHSHUA Outsourcing Worldwide, Inc • Philippines

On-site
PHP 1,800,000 - 2,400,000
Lead AI/ML Engineer
Lead AI/ML Engineer

HRTX • Philippines

On-site
PHP 1,800,000 - 3,000,000