DevOps / Site Reliability Engineer

Triomics

India

On-site

INR 900,000 - 1,400,000

Full time

22 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health insurance
Meal benefit (Zomato)

Job summary

Triomics is seeking a DevOps/SRE to design, build, and maintain scalable, secure infrastructure in production. You will drive automation, cloud management, monitoring, and incident response to ensure high availability and performance.

You will implement CI/CD pipelines, manage Kubernetes and Docker environments, and automate provisioning with Terraform and Helm while writing Python and Bash scripts. Collaboration with engineering teams and thorough documentation are essential to success.

Qualifications

  • Experience in DevOps or SRE role with cloud deployments.
  • Strong automation and scripting skills are required.
  • Ability to design and maintain secure, scalable infrastructure.

Responsibilities

  • Design, implement, and manage cloud-based infrastructure and deployment pipelines.
  • Build and maintain CI/CD pipelines for reliable software delivery.
  • Manage and optimize containerized environments using Kubernetes and Docker.
  • Automate provisioning with Terraform and Helm.
  • Develop automation scripts using Python and Bash.
  • Monitor system health, performance, and reliability.
  • Respond to incidents and perform root-cause analysis.
  • Ensure security and hardening of infrastructure.
  • Collaborate with development teams to improve reliability and deployment processes.
  • Maintain documentation for infrastructure and operations processes.

Skills

Kubernetes
Docker
Jenkins
Terraform
Helm
Python
Bash scripting
Linux administration
Networking concepts
Security practices
Monitoring/incident response

Tools

Kubernetes
Docker
Jenkins
Terraform
Helm

Job description

We are looking for a DevOps/Site Reliability Engineer to help design, build, and maintain scalable, secure, and reliable infrastructure. In this role, you will work closely with engineering teams to streamline deployments, improve system reliability, and ensure our platforms run efficiently in production. You will be responsible for building automation, managing cloud infrastructure, monitoring systems, and responding to incidents to maintain high availability and performance.

Key Responsibilities:
  • Design, implement, and manage cloud-based infrastructure and deployment pipelines.
  • Build and maintain CI/CD pipelines to enable reliable and efficient software delivery.
  • Manage and optimize containerized environments using Kubernetes and Docker.
  • Automate infrastructure provisioning and configuration using Terraform and Helm.
  • Develop and maintain automation scripts using Python and Bash.
  • Monitor system health, performance, and reliability using logging and monitoring tools.
  • Troubleshoot production issues and participate in incident response and root cause analysis.
  • Ensure infrastructure security, network configuration, and system hardening best practices.
  • Collaborate with development teams to improve reliability, scalability, and deployment processes.
  • Maintain clear documentation for infrastructure, processes, and operational procedures.
Requirements:
  • 1+ years of experience in DevOps, Site Reliability Engineering, or a related role.
  • Hands-on experience with at least one cloud platform (AWS, Azure, or GCP).
  • Strong experience with Kubernetes, Docker, Jenkins, Terraform, and Helm .
  • Proficiency in Python and Bash scripting .
  • Solid understanding of Linux system administration, networking concepts, and security practices.
  • Experience with monitoring, logging, and incident response systems.
  • Strong communication skills and the ability to document technical processes effectively.
  • Software development experience is a plus.
Nice to Have:
  • Experience deploying and scaling AI/ML workloads in production environments.
  • Familiarity with single-tenant deployment models.
  • Experience with MLOps/AIOps platforms such as SageMaker or Kubeflow.
  • Knowledge of chaos engineering and disaster recovery strategies.
  • Experience with cloud cost optimization strategies.
  • Relevant cloud certifications (AWS, Azure, or GCP).
Why Join Us?
  • Impact at scale - The AI you build directly accelerates cancer research and improves patient outcomes worldwide.
  • Cutting-edge problems - You will work on some of the hardest and most interesting LLM engineering challenges in a highly regulated industry.
  • World-class team - Collaborate with experts across AI, engineering, product, and oncology with best-in-industry compensation.
  • Culture that ships - We are a team that works hard and plays hard (company-sponsored workations in Bali, Sri Lanka, Goa, and more).
  • Lunch provided at the office - one less daily decision, one happier employee.
  • Flexible working hours - we care about output, not clock-ins.
  • Health insurance - comprehensive coverage for you and your family.
  • Zomato meal benefit - breakfast and dinner can be ordered when you come in early or leave late, because effort deserves fuel.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Triomics • Varanasi

On-site
INR 350,000 - 600,000
Lunch Provided at Office
Flexible Working Hours
Health Insurance
+1
SDE2 – Software Development Engineer II (Backend)
SDE2 – Software Development Engineer II (Backend)

Triomics • India

On-site
INR 2,800,000 - 4,200,000
Comprehensive health insurance for you
Zomato meal benefits
Lunch provided at the office
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SourcingXPress • Hyderabad

On-site
INR 3,000,000 - 5,000,000
SDE2 – Software Development Engineer II (Full-Stack)
SDE2 – Software Development Engineer II (Full-Stack)

Triomics • India

On-site
INR 1,800,000 - 3,200,000
Lunch provided
Meal benefits
Flexible hours
+1
Senior Infrastructure Engineer
Senior Infrastructure Engineer

SourcingXPress • Hyderabad

On-site
INR 2,000,000 - 3,500,000
High ownership
Rapid learning opportunities
Career advancement potential
Lead DevOps AI Engineer
Lead DevOps AI Engineer

Trine Infotech • India

Remote
INR 2,500,000 - 6,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Clarus Advisers • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Senior DevOps Engineer
Senior DevOps Engineer

Sitero LLC • Bengaluru

On-site
INR 3,000,000 - 5,000,000
Hybrid work model
Learning & development budget
Health insurance
+1
AI_DevOps Engineer
AI_DevOps Engineer

RIA Advisory • Pune District

On-site
INR 1,000,000 - 1,500,000
DevOps Engineer
DevOps Engineer

Recrew AI • Bengaluru

On-site
INR 1,500,000 - 2,000,000
Competitive compensation
Access to cutting-edge GPU compute infrastructure
Opportunity to work alongside AI researchers