Engineering Manager

Higlobe, Inc.

India

Remote

INR 17,341,000 - 23,121,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Flexible Paid Time Off
Equity Compensation & Employee Stock-P
Growth and Development Fund
Parental Leave

Job summary

GitLab is seeking an Engineering Manager for Production Engineering - Observability. You will lead a globally distributed team responsible for metrics, logging, alerting, and capacity planning that power GitLab.com and GitLab Dedicated.

In your first year, you’ll drive improvements to Prometheus pipelines, log ingestion/retention, and SLO-driven incident management. You will collaborate with Site Reliability Engineering, Product Engineering, and other Infrastructure Platforms teams to optimize

Qualifications

  • Experience leading observability, platform engineering, or SRE at scale in distributed environments.
  • Strong knowledge of metrics systems (Prometheus) and long-term storage.
  • Experience using SLOs, error budgets, and capacity forecasts for reliability decisions.

Responsibilities

  • Lead, hire, onboard, and develop a globally distributed Observability team.
  • Set priorities with cross-functional teams and deliver observability services iteratively.
  • Own reliability, scalability, and cost of metrics, logging, alerting, and capacity planning platforms.
  • Reduce noisy alerts and telemetry gaps using SLOs and self-service instrumentation.
  • Guide decisions on time-series storage, high-cardinality metrics, and distributed tracing.
  • Participate in incident response and ensure sustainable on-call rotation.
  • Leverage AI tools to support engineering workflows while maintaining accountability.
  • Coordinate with Site Reliability Engineering, Product Engineering, and GitLab Dedicated teams.

Skills

Distributed leadership
Observability expertise
SRE at scale
Stakeholder communication
AI for engineering

Tools

Prometheus
Elasticsearch
Cloud-native services
SLOs & error budgets

Job description

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.

The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our valuesand continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software.

*Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.

Engineering Manager, Production Engineering - Observability
An overview of this role

You'll lead the globally distributed Observability team. The team builds and operates the metrics, logging, alerting, and capacity planning platforms that GitLab engineers use to understand GitLab.com and GitLab Dedicated. You'll help determine how the team collects, stores, queries, and acts on telemetry, balancing reliable signals with scale and cost.
In your first year, you'll guide improvements to Prometheus-based metrics pipelines, log ingestion and retention, alerting driven by service-level objectives (SLOs), and capacity forecasting. You'll work with Site Reliability Engineering, Product Engineering, and other Infrastructure Platforms teams to make it easier for engineers to observe the services they own. You'll also take part in incident response and help keep the team's on-call work sustainable.

What you’ll do
  • Lead, hire, onboard, and develop a distributed engineering team working asynchronously.
  • Set priorities with Site Reliability Engineering, Product Engineering, and GitLab Dedicated teams, and help the team deliver observability services iteratively.
  • Own the reliability, scalability, and cost of the team's metrics, logging, alerting, and capacity planning platforms.
  • Reduce noisy or missing alerts and telemetry gaps, and use SLOs, error budgets, and self-service instrumentation to help engineers maintain the health of their services.
  • Guide technical decisions about time-series storage, high-cardinality metrics, log pipelines, and distributed tracing.
  • Participate in the Incident Manager On Call (IMOC) rotation, coordinating the response to high-severity incidents affecting GitLab.com.
  • Keep the team's on-call rotation sustainable through coverage across time zones, useful runbooks, better alerts, and follow-through on post-incident actions.
  • Use AI tools and agents to support engineering workflows and incident triage, reviewing their output while engineers retain responsibility for decisions.
What you’ll bring
  • Experience leading an observability, platform engineering, or site reliability engineering team operating at scale, including supporting people in a distributed, asynchronous environment.
  • Technical knowledge of metrics systems such as Prometheus and long-term storage, logging platforms such as Elasticsearch or cloud-native services, and alerting design.
  • Experience using SLOs, error budgets, and capacity forecasts to make reliability and investment decisions.
  • Experience operating a large software-as-a-service platform and investigating production issues such as telemetry gaps, ingestion limits, or noisy and missing alerts.
  • Experience participating in and improving production on-call rotations, including incident coordination and balancing operational load with project work.
  • The ability to explain technical tradeoffs to engineering partners and other stakeholders.
  • Experience using AI tools or agents in engineering or management work; you can describe how you would apply them to operational problems such as incident triage.
About the team

We're part of Production Engineering within Infrastructure Platforms and work asynchronously across regions. Our tools include Tamland, the capacity forecasting tool. We bring lessons from operating GitLab's production systems back into the platforms we build.

How GitLab Supports Full-Time Employees
  • Benefits to support your health, finances, and well-being
  • Flexible Paid Time Off
  • Team Member Resource Groups
  • Equity Compensation & Employee Stock Purchase Plan
  • Growth and Development Fund
  • Parental Leave

Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process.

Privacy Policy: Please review our Recruitment Privacy Policy. Your privacy is important to us.

GitLab is proud to be an equal opportunity workplace and is an affirmative action employer. GitLab’s policies and practices relating to recruitment, employment, career development and advancement, promotion, and retirement are based solely on merit, regardless of race, color, religion, ancestry, sex (including pregnancy, lactation, sexual orientation, gender identity, or gender expression), national origin, age, citizenship, marital status, mental or physical disability, genetic information (including family medical history), discharge status from the military, protected veteran status (which includes disabled veterans, recently separated veterans, active duty wartime or campaign badge veterans, and Armed Forces service medal veterans), or any other basis protected by law. GitLab will not tolerate discrimination or harassment based on any of these characteristics. See alsoGitLab’s EEO PolicyandEEO is the Law. If you have a disability or special need that requiresaccommodation, please let us know during therecruiting process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Engineering Manager, Gitlab Delivery: Upgrades
Engineering Manager, Gitlab Delivery: Upgrades

GitLab • Mumbai

On-site
INR 1,500,000 - 2,000,000
Flexible Paid Time Off
Equity Compensation
Home office support
Engineering Manager, Gitaly
Engineering Manager, Gitaly

GitLab • Bengaluru

On-site
INR 12,497,625 - 26,780,626
Flexible Paid Time Off
Equity Compensation
Home office support
Engineering Manager, Continuous Delivery
Engineering Manager, Continuous Delivery

GitLab • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Flexible Paid Time Off
Team Member Resource Groups
Equity Compensation
+3
Staff Software Engineer - NLP
Staff Software Engineer - NLP

Higlobe, Inc. • India

Remote
INR 4,000,000 - 7,000,000
Flexible PTO
Equity Compensation & Employee Stock-P
Growth and Development Fund
+2
Staff Backend Engineer, Gitlab Delivery: Upgrades
Staff Backend Engineer, Gitlab Delivery: Upgrades

GitLab • Delhi

On-site
INR 1,000,000 - 1,500,000
Flexible Paid Time Off
Team Member Resource Groups
Equity Compensation & Employee Stock Purchase Plan
+2
Director, Engineering, Platform Operations & Productivity
Director, Engineering, Platform Operations & Productivity

Higlobe, Inc. • Bengaluru

On-site
INR 4,000,000 - 8,000,000
Flexible Paid Time Off
Equity Compensation
Parental Leave
+2
Staff Frontend Engineer
Staff Frontend Engineer

Higlobe, Inc. • India

Remote
INR 13,500,000 - 19,286,000
Equity compensation
Flexible PTO
Growth & Development Fund
+3
Site Reliability Engineer, Environment Automation
Site Reliability Engineer, Environment Automation

GitLab • Mumbai

On-site
INR 1,800,000 - 3,200,000
Flexible Paid Time Off
Equity Compensation & Employee Stock Purchase Plan
Growth and Development Fund
+3
Engineering Manager, Continuous Delivery
Engineering Manager, Continuous Delivery

GitLab • Mumbai

On-site
INR 1,500,000 - 2,100,000
Flexible Paid Time Off
Equity Compensation & Employee Stock Purchase Plan
Growth and Development Fund
+2
Senior Frontend Engineer
Senior Frontend Engineer

Higlobe, Inc. • India

Remote
INR 1,800,000 - 2,400,000
Flexible Paid Time Off
Equity Compensation & Employee Stock
Growth and Development Fund
+1