Site Reliability Engineer

Triomics

Varanasi

On-site

INR 350,000 - 600,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Lunch Provided at Office
Flexible Working Hours
Health Insurance
Meal Benefit

Job summary

Triomics in India is seeking a DevOps / Site Reliability Engineer to design, build, and maintain scalable cloud infrastructure. You will partner with engineering to streamline deployments, boost reliability, and keep platforms running efficiently in production.

The role covers automation, CI/CD, Kubernetes and Docker management, Terraform/Helm, Python and Bash scripts, monitoring, incident response, and security hardening, with a culture of documentation and collaboration.

Qualifications

  • 1+ years of experience in DevOps, Site Reliability Engineering, or related role.
  • Hands-on experience with at least one cloud platform (AWS, Azure, or GCP).
  • Strong experience with Kubernetes, Docker, Jenkins, Terraform, and Helm.
  • Proficiency in Python and Bash scripting.
  • Solid understanding of Linux system administration, networking concepts, and security practices.
  • Experience with monitoring, logging, and incident response systems.
  • Strong communication skills and the ability to document technical processes effectively.
  • Software development experience is a plus.

Responsibilities

  • Design, implement, and manage cloud-based infrastructure and deployment pipelines.
  • Build and maintain CI/CD pipelines to enable reliable and efficient software delivery.
  • Manage and optimize containerized environments using Kubernetes and Docker.
  • Automate infrastructure provisioning and configuration using Terraform and Helm.
  • Develop and maintain automation scripts using Python and Bash.
  • Monitor system health, performance, and reliability using logging and monitoring tools.
  • Troubleshoot production issues and participate in incident response and root cause analysis.
  • Ensure infrastructure security, network configuration, and system hardening best practices.
  • Collaborate with development teams to improve reliability, scalability, and deployment processes.
  • Maintain clear documentation for infrastructure, processes, and operational procedures.

Skills

DevOps
Site Reliability
Python scripting
Bash scripting
CI/CD
Cloud fundamentals
Linux administration
Networking basics
Security practices
Documentation

Tools

Kubernetes
Docker
Jenkins
Terraform
Helm

Job description

DevOps / Site Reliability Engineer (SRE)

About Triomics: Triomics is building the agentic AI layer for oncology EHRs. Cancer hospitals spend billions on highly trained staff manually reading unstructured patient records such as pathology reports, clinical notes, genomic panels to power workflows like trial matching, registry curation, visit prep, and quality reporting. We replace that manual work with task‑driven AI agents that sit inside the EMR and process records with >95% accuracy, at scale, in real time. We have grown 10x in the last one year and are processing millions of documents monthly.

About The Role

We are looking for a DevOps / Site Reliability Engineer to design, build, and maintain scalable, secure, and reliable infrastructure. In this role, you will work closely with engineering teams to streamline deployments, improve system reliability, and ensure our platforms run efficiently in production. You will be responsible for building automation, managing cloud infrastructure, monitoring systems, and responding to incidents to maintain high availability and performance.

Key Responsibilities
  • Design, implement, and manage cloud‑based infrastructure and deployment pipelines.
  • Build and maintain CI/CD pipelines to enable reliable and efficient software delivery.
  • Manage and optimize containerized environments using Kubernetes and Docker.
  • Automate infrastructure provisioning and configuration using Terraform and Helm.
  • Develop and maintain automation scripts using Python and Bash.
  • Monitor system health, performance, and reliability using logging and monitoring tools.
  • Troubleshoot production issues and participate in incident response and root cause analysis.
  • Ensure infrastructure security, network configuration, and system hardening best practices.
  • Collaborate with development teams to improve reliability, scalability, and deployment processes.
  • Maintain clear documentation for infrastructure, processes, and operational procedures.
Requirements
  • 1+ years of experience in DevOps, Site Reliability Engineering, or related role.
  • Hands‑on experience with at least one cloud platform (AWS, Azure, or GCP).
  • Strong experience with Kubernetes, Docker, Jenkins, Terraform, and Helm.
  • Proficiency in Python and Bash scripting.
  • Solid understanding of Linux system administration, networking concepts, and security practices.
  • Experience with monitoring, logging, and incident response systems.
  • Strong communication skills and the ability to document technical processes effectively.
  • Software development experience is a plus.
Nice to Have
  • Experience deploying and scaling AI/ML workloads in production environments.
  • Familiarity with single‑tenant deployment models.
  • Experience with MLOps/AIOps platforms such as SageMaker or Kubeflow.
  • Knowledge of chaos engineering and disaster recovery strategies.
  • Experience with cloud cost optimization strategies.
  • Relevant cloud certifications (AWS, Azure, or GCP).
Why Join us?
  • Impact at scale - the AI you build directly accelerates cancer research and improves patient outcomes worldwide.
  • Cutting‑edge problems - you’ll work on some of the hardest and most interesting LLM‑engineering challenges in a highly regulated industry.
  • World‑class team - collaborate with experts across AI, engineering, product, and oncology with best‑in‑industry compensation.
  • Culture that ships - we’re a team that works hard and plays hard (company‑sponsored workations in Bali, Sri Lanka, Goa, and more).
Perks & Benefits
  • Lunch Provided at the Office – one less daily decision, one happier employee.
  • Flexible Working Hours – we care about output, not clock‑ins.
  • Health Insurance – comprehensive coverage for you and your family.
  • Zomato Meal Benefit – breakfast and dinner can be ordered when you come in early or leave late, because effort deserves fuel.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

MLOps & Data Engineer
MLOps & Data Engineer

Triomics • Varanasi

On-site
INR 1,800,000 - 3,200,000
Lunch provided at the office
Flexible working hours
Comprehensive health insurance for you
+1
MLOps & Data Engineer
MLOps & Data Engineer

Triomics • India

On-site
INR 1,200,000 - 2,000,000
Lunch provided at the office
Flexible working hours
Comprehensive health insurance for you
+1
Research Engineer, Applied ML
Research Engineer, Applied ML

Triomics • Varanasi

On-site
INR 900,000 - 1,500,000
Lunch provided at the office
Flexible working hours
Comprehensive health insurance for you
+1
Implementation Engineer
Implementation Engineer

Triomics • Varanasi

On-site
INR 900,000 - 1,700,000
Lunch provided at the office
Flexible working hours
Comprehensive health insurance for you
+1
ML Model Evaluation Engineer
ML Model Evaluation Engineer

Triomics • India

On-site
INR 1,200,000 - 2,400,000
Lunch provided
Flexible working hours
Comprehensive health insurance for you
ML Model Evaluation Engineer
ML Model Evaluation Engineer

Triomics • Varanasi

On-site
INR 1,200,000 - 2,400,000
Lunch provided at the office.
Flexible working hours.
Comprehensive health insurance for you
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SourcingXPress • Hyderabad

On-site
INR 3,000,000 - 5,000,000
MLOps & Data Engineer
MLOps & Data Engineer

Triomics • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Comprehensive health insurance
Flexible working hours
Zomato meal benefits
+1
Implementation Engineer
Implementation Engineer

Triomics • India

On-site
INR 600,000 - 1,200,000
Lunch provided at the office
Flexible working hours
Comprehensive health insurance for you
+1
Subject Matter Expert - Clinical Team
Subject Matter Expert - Clinical Team

Triomics • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Company-sponsored workations