Platform Engineer

UST

Bengaluru

On-site

INR 2,400,000 - 4,200,000

Full time

21 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

UST is seeking a Platform Engineer (Observability) in Bengaluru to design and operate enterprise monitoring platforms that ensure reliable service delivery. You will build scalable telemetry pipelines for metrics, logs, and traces and automate platform operations across cloud and on-prem environments.

The role requires 5+ years of experience in observability or SRE, with hands-on exposure to Dynatrace, Datadog, Grafana, Prometheus, OpenTelemetry, and related tooling.

Qualifications

  • 5+ years of experience in monitoring/observability, platform engineering, or SRE.
  • Hands-on with at least one enterprise monitoring/observability platform.

Responsibilities

  • Engineer and operate enterprise observability platforms across cloud and on-prem environments.
  • Design telemetry collection for servers, networks, storage, databases, containers, apps, and cloud services.
  • Manage capacity, performance, security, and upgrades of monitoring platforms.
  • Establish alerting standards including thresholds, anomaly detection, suppression, and routing.
  • Reduce alert fatigue with actionable alerts and dashboards.
  • Build dashboards for service health, performance, capacity, availability, and SLO/SLA reporting.
  • Automate monitoring deployment via IaC and CI/CD.
  • Integrate platforms with Jira, CMDB, Teams, Slack, PagerDuty, Opsgenie, etc.

Skills

Observability concepts
Telemetry pipelines
Cloud platforms
Containers
Infrastructure as Code
Automation
CI/CD practices
SRE / Incident management
Alerting & dashboards
OpenTelemetry

Education

Bachelor’s degree in Computer Science or related

Tools

Dynatrace
Datadog
New Relic
Splunk
Elastic
Prometheus
Grafana
Tempo
Loki
Terraform
Ansible

Job description

Job Role: Platform Engineer (Observability)

Experience Required: 5+ years

Mandatory Skills: Observability, Telemetry, Cloud, Containers, IaC, Automation, CI/CD

Who we are:

At UST, we help the world’s best organizations grow and succeed through transformation. Bringing together the right talent, tools, and ideas, we work with our client to co-create lasting change. Together, with over 30,000 employees in 30+ countries, we build for boundless impact—touching billions of lives in the process. Visit us at UST.com.

Summary:

UST is looking for a Monitoring Platform Engineer who will build, operate, and continuously improve enterprise monitoring and observability platforms that support reliable service delivery. You will design scalable telemetry pipelines for metrics, logs, and traces; improve alert quality; automate platform operations; and help teams detect, diagnose, and resolve service issues quickly.

The Opportunity:

Key Roles and Responsibilities:

  • Engineering and operating enterprise observability platforms such as Dynatrace, Datadog, New Relic, Splunk, Elastic, Prometheus, and Grafana across cloud and on-premises environments.
  • Designing and maintaining telemetry collection for servers, networks, storage, databases, containers, applications, and cloud services.
  • Managing the capacity, performance, availability, security, and lifecycle upgrades of monitoring platforms.
  • Establishing alerting standards covering severity, thresholds, anomaly detection, suppression, correlation, deduplication, and routing.
  • Reducing alert fatigue and creating actionable alerts supported by relevant dashboards, runbooks, and diagnostic context.
  • Building standardized dashboards for service health, performance, capacity, availability, and SLO/SLA reporting.
  • Partnering with service owners to define SLIs, SLOs, and error budgets where applicable.
  • Automating monitoring deployment, configuration, and onboarding through infrastructure-as-code, configuration-management, and CI/CD practices.
  • Integrating monitoring platforms with Jira, CMDB systems, Microsoft Teams, Slack, PagerDuty, Opsgenie, and other operational tools.
  • Creating reusable templates, policies, alerts, and dashboards that enable consistent self-service adoption.
  • Defining monitoring standards, telemetry requirements, tagging conventions, governance controls, documentation, and operational runbooks.
  • Supporting incident response and post-incident reviews to improve detection, reduce recurrence, and strengthen monitoring coverage.
  • Participating in a scheduled 24×7 on-call rotation supporting monitoring platforms and services.
  • Collaborating with Security and Compliance teams to meet access-control, retention, privacy, audit, and regulatory requirements.

What you need:

  • Three to seven or more years of experience in monitoring and observability, platform engineering, site reliability engineering, or IT operations.
  • Hands-on experience with at least one enterprise monitoring or observability platform.

Required Skills:

  • A strong understanding of metrics, logs, traces, distributed-system troubleshooting, and OpenTelemetry practices.
  • Experience administering and troubleshooting Linux and Windows systems in large-scale environments.
  • Working knowledge of networking concepts, including DNS, TCP/IP, load balancers, and firewalls.
  • Experience with cloud-monitoring services such as Azure Monitor and Log Analytics, AWS CloudWatch, or Google Cloud Operations.
  • Experience monitoring containers and Kubernetes using technologies such as Prometheus, OpenTelemetry, Grafana, Tempo, or Loki.
  • Automation and scripting experience with Python, PowerShell, or Bash, along with familiarity with CI/CD pipelines.
  • Experience with infrastructure-as-code or configuration-management tools such as Terraform, Bicep, CloudFormation, or Ansible.
  • Familiarity with ITSM, CMDB, event-management workflows, and SRE practices such as SLOs and error budgets.
  • Strong systems thinking and the ability to troubleshoot across application, platform, infrastructure, and network layers.

Desired Skills:

  • Excellent judgment in separating actionable signals from noise and tuning alerts accordingly.
  • Clear communication skills and the ability to collaborate with technical and non-technical stakeholders, especially during incidents.
  • A customer-focused approach to enablement, self-service, documentation, and platform adoption.
  • The ability to prioritize effectively, work in a fast-paced operational environment, and manage escalations calmly.
  • Relevant monitoring-platform, cloud-platform, or ITIL Foundation certifications would be advantageous.

Qualification:

  • A bachelor’s degree in Computer Science, Information Systems, or a related discipline, or equivalent professional experience.

What we believe:

We’re proud to embrace the same values that have shaped UST since the beginning. Since day one, we’ve been building enduring relationships and a culture of integrity. And today, it's those same values that are inspiring us to encourage innovation from everyone, to champion diversity and inclusion and to place people at the centre of everything we do.

Humility:

We will listen, learn, be empathetic and help selflessly in our interactions with everyone.

Humanity:

Through business, we will better the lives of those less fortunate than ourselves.

Integrity:

We honour our commitments and act with responsibility in all our relationships.

Equal Employment Opportunity Statement

UST is an Equal Opportunity Employer. We believe that no one should be discriminated against because of their differences, such as age, disability, ethnicity, gender, gender identity and expression, religion, or sexual orientation.

All employment decisions shall be made without regard to age, race, creed, colour, religion, sex, national origin, ancestry, disability status, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by federal, state, or local law.

UST reserves the right to periodically redefine your roles and responsibilities based on the requirements of the organization and/or your performance.

  • To support and promote the values of UST.
  • Comply with all Company policies and procedures
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead SRE
Lead SRE

UST • Bengaluru

On-site
INR 4,000,000 - 7,000,000
DevOps Specialist
DevOps Specialist

UST • Thiruvananthapuram

On-site
INR 1,800,000 - 3,200,000
Cloud Engineer
Cloud Engineer

UST • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Software Architect I
Software Architect I

UST • Bengaluru

Hybrid
INR 5,000,000 - 9,000,000
DevOps Engineer
DevOps Engineer

UST • Bengaluru

On-site
INR 2,500,000 - 4,200,000
DevOps Engineer
DevOps Engineer

UST • Thiruvananthapuram

On-site
INR 2,200,000 - 3,600,000
Product Engineer II
Product Engineer II

UST • Hyderabad

On-site
INR 1,500,000 - 2,300,000
Enterprise Observability Platform Engineer
Enterprise Observability Platform Engineer

Be a Catalyst • Gurugram District

On-site
INR 1,500,000 - 2,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Hyderabad

On-site
INR 4,200,000 - 7,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Chennai District

On-site
INR 3,500,000 - 7,000,000