Senior Site Reliability Engineer (In-Office Required)

Nebius

New York (NY)

On-site

USD 100,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401(k) plan
Parental leave

Job summary

Nebius in New York is searching for a DevOps Engineer to manage Kubernetes clusters and infrastructure as code. You'll work closely with a small engineering team to optimize CI/CD pipelines and maintain real-time data pipelines that handle billions of events daily.

The ideal candidate has 5-8 years of experience in production environments, with solid skills in API reliability and Kubernetes management. Benefits include 100% paid health insurance, a 401(k) plan, and more.

Qualifications

  • 5-8 years in a DevOps or SRE role, working in production environments.
  • Experience designing and operating large-scale, distributed systems.
  • Proven Kubernetes experience in a managed cloud environment.

Responsibilities

  • Manage Kubernetes clusters across multiple environments and regions.
  • Own infrastructure as code for all resources.
  • Maintain and improve CI/CD pipelines and GitOps-based deployments.

Skills

DevOps or SRE experience
API design knowledge
Kubernetes expertise
Infrastructure as code
GitOps workflows

Job description

About Tavily

We're building the infrastructure layer for agentic web interaction at scale. Our API is designed from the ground up to power Retrieval-Augmented Generation (RAG) and real-time reasoning in AI systems. By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that are not only intelligent — but also informed.

We work with some of the most innovative teams in AI — from small startups shaping the ecosystem to the largest enterprises deploying AI at scale. Whether it's powering sales assistants, research copilots, or internal knowledge tools, we're the missing link between LLMs and the real world.

The Role:
  • Managing Kubernetes clusters across multiple environments and regions
  • Owning infrastructure as code for all resources
  • Maintaining and improving CI/CD pipelines and GitOps-based deployments
  • Maintaining and optimizing real-time data pipelines that process billions of events per day across distributed queues and stream processors
  • Building out monitoring, alerting, and observability
  • Debugging production issues across services
  • Managing cloud costs and capacity planning
  • Working closely with a small engineering team — you down infra, not a slice of it
What we're looking for
  • 5-8 years in a DevOps or SRE role, working in production environments
  • Proven experience designing and operating large-scale, distributed systems, with a solid understanding of API design, reliability, and performance at scale
  • Strong Kubernetes experience in a managed cloud environment
  • Proficiency with infrastructure as code (Terraform or similar)
  • Experience with GitOps-based deployment workflows
  • Built or maintained observability stacks (logging, metrics, alerting)
  • Experience handling production incidents calmly and methodically
Nice to have
  • Multi-region deployments
  • Search infrastructure
  • Data pipeline experience (streaming, warehousing)
  • Proxy/networking infrastructure at scale
Why Tavily?
  • Full ownership — small team, you own the entire infrastructure, not a slice of it
  • Real scaling challenges — bursty scraping workloads, cache invalidation, multi-region, millions of daily requests
  • AI-native company — your infra directly powers AI agents used by leading companies in the space.
Key employee benefits in the US
  • Health insurance — 100% company-paid medical, dental, and vision coverage for employees and families.
  • 401(k) plan — Up to 4% company match with immediate vesting.
  • Parental leave — 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers.
  • Remote work reimbursement — Up to $85/month for mobile and internet.
  • Disability & life insurance — Company-paid short-term, long-term and life insurance coverage.
Equal Opportunity Statement

Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.

Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire.

If you need accommodations during the application process, please let us know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (In-Office Required)
Senior Site Reliability Engineer (In-Office Required)

Tavily • United States

Hybrid
USD 147,000 - 224,000
Health insurance
401(k) plan
Parental leave
+2
Senior Site Reliability Engineer (Agentic Search)
Senior Site Reliability Engineer (Agentic Search)

Tavily Inc. • New York (NY)

Hybrid
USD 156,000 - 262,000
100% company-paid medical, dental, and vision coverage
Up to 4% company match 401(k) plan
20 weeks paid parental leave for primary caregivers
+2
Head of Forward Deployment Engineering - Tavily
Head of Forward Deployment Engineering - Tavily

Nebius • New York (NY)

Hybrid
USD 230,000 - 310,000
100% company-paid health insurance
401(k) plan with up to 4% match
20 weeks paid parental leave
+2
Software Engineer
Software Engineer

Tavily • New York (NY)

On-site
USD 100,000 - 150,000
Front-End Engineer
Front-End Engineer

Tavily • New York (NY)

On-site
USD 100,000 - 140,000
Inclusive company culture
Building alongside a fast-moving team
Daily team lunches
+3
Head of Forward Deployment Engineering - Tavily
Head of Forward Deployment Engineering - Tavily

Nebius • City of Utica (NY)

Hybrid
USD 230,000 - 310,000
Health Insurance
401(k) Plan
Parental Leave
+2
Senior Software Engineer (Agentic Search) - Billing
Senior Software Engineer (Agentic Search) - Billing

Nebius • New York (NY)

Hybrid
USD 170,000 - 240,000
Software Engineer, Infrastructure
Software Engineer, Infrastructure

Tavus • San Francisco (CA)

On-site
USD 180,000 - 240,000
Flexible work schedule
Unlimited PTO
Competitive healthcare
+1
Web Crawling Engineer for AI-Powered Search Pipelines
Web Crawling Engineer for AI-Powered Search Pipelines

Tavily • New York (NY)

On-site
USD 140,000 - 210,000
Enterprise AI & Data Infrastructure Account Executive
Enterprise AI & Data Infrastructure Account Executive

Tavily • New York (NY)

On-site
USD 120,000 - 180,000