Software Engineer in Hardware Infrastructure Observability

Nebius B.V.

Deutschland

Remote

EUR 90.000 - 130.000

Vollzeit

Vor 2 Tagen
Sei unter den ersten Bewerbenden

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Competitive compensation
Career growth opportunities
Flexibility and ownership
Collaborative culture
Impactful AI projects
International environment

Zusammenfassung

Nebius B.V. is seeking a Senior Software Engineer to join the Hardware Infrastructure Observability team. You will design and build services and agents for deep visibility into a large server fleet and DC systems.

You will evolve metrics pipelines, build maintenance workflows, and investigate incidents hands-on. English proficiency is essential, and you should be proficient in Python and Go with strong Linux fundamentals.

Qualifikationen

  • 5+ years of professional software engineering experience.
  • Excellent knowledge of Python and Golang or you are ready to quickly switch to these languages.
  • Strong Linux fundamentals.
  • Ability to write reliable code and dig into complex problems.
  • Working proficiency in English.

Aufgaben

  • Design and develop services and agents that provide deep visibility into a large server fleet and DC engineering systems.
  • Evolve our metrics/aggregation/alerting pipelines and improve signals quality.
  • Build maintenance workflows and automation that keep fleets healthy.
  • Investigate incidents hands-on (including on-host debugging) and drive root‑cause fixes.
  • Collaborate with hardware, networking, and DC operations to improve reliability.

Kenntnisse

Python
Golang
Linux fundamentals
Reliability-focused coding
English proficiency

Jobbeschreibung

The Role Nebius is looking for a Senior Software Engineer to join the Hardware Infrastructure Observability team. You're welcome to work from our office in Amsterdam. We build and run low-level monitoring for servers and data center engineering systems to ensure reliability at scale. We also design and operate maintenance and remediation systems that enable safe, predictable fleet-wide changes and keep the infrastructure healthy.

Key Responsibilities
  • Design and develop services and agents that provide deep visibility into a large server fleet and DC engineering systems;
  • Evolve our metrics/aggregation/alerting pipelines and improve signals quality;
  • Build maintenance workflows and automation that keep fleets healthy;
  • Investigate incidents hands-on (including on-host debugging) and drive root‑cause fixes;
  • Collaborate with hardware, networking, and DC operations to improve reliability.
You should have
  • 5+ years of professional software engineering experience;
  • Excellent knowledge of Python and Golang or you are ready to quickly switch to these languages;
  • Strong Linux fundamentals;
  • Ability to write reliable code and dig into complex problems;
  • Working proficiency in English.
We expect Engineers to
  • Manage large-scale projects;
  • Break down complex tasks;
  • Be experts in specific technologies;
  • Assess task priority;
  • Have strong architectural thinking;
  • Be involved in hiring;
  • Mentor others.
Bonus
  • understanding of modern server architecture;
  • metrics/monitoring/alerting Prometheus-compatible stacks;
  • Networking knowledge;
  • Experience designing, developing, and running high-load distributed systems.
Benefits & Perks
  • Competitive compensation;
  • Career growth and learning opportunities;
  • Flexibility and ownership;
  • Collaborative and innovative culture;
  • Opportunity to work on impactful AI projects;
  • International environment and talented teams.
What it’s like to work at Nebius
  • Fast moving
  • Bold thinking
  • Constant growth
  • Meaningful impact
  • Trust and real ownership
  • Opportunity to shape the future of AI.
Equal Opportunity Statement

Nebius is an equal opportunity employer. Applicants must be authorized to work in the country of application. If you need accommodations during the application process, please let us know.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer
Site Reliability Engineer

Nebius B.V. • Deutschland

Remote
EUR 112.000 - 155.000
Health insurance (family coverage)
401(k) plan with company match
Parental leave
+2
Senior Software Engineer
Senior Software Engineer

Nebius B.V. • Deutschland

Hybrid
EUR 112.000 - 146.000
Health insurance
401(k) plan
Parental leave
+2
Software Engineer in Hardware Infrastructure
Software Engineer in Hardware Infrastructure

Nebius B.V. • Deutschland

Remote
EUR 90.000 - 150.000
Senior Backend Software Engineer (Observability)
Senior Backend Software Engineer (Observability)

Nebius B.V. • Deutschland

Remote
EUR 90.000 - 135.000
Software Engineer in Infrastructure
Software Engineer in Infrastructure

Nebius B.V. • Deutschland

Remote
EUR 155.000 - 193.000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
Senior Backend Software Engineer (Cloud Monetization Platform)
Senior Backend Software Engineer (Cloud Monetization Platform)

Meyandy LLC • Berlin

Hybrid
EUR 90.000 - 130.000
Competitive compensation
Career growth and learning opportunit�
Flexibility and ownership
+1
Technical Support Engineer
Technical Support Engineer

Nebius B.V. • Deutschland

Hybrid
EUR 94.000 - 118.000
Health Insurance
401(k) Plan
Parental Leave
+2
Senior Backend Engineer
Senior Backend Engineer

Nebius • Berlin

Hybrid
EUR 90.000 - 130.000
Competitive salary
Career growth and learning
Flexible work
+3
Infrastructure Software Engineer
Infrastructure Software Engineer

Nebius B.V. • Deutschland

Remote
EUR 129.000 - 181.000
Health insurance
Remote work reimbursement
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Meyandy LLC • Berlin

Hybrid
EUR 90.000 - 120.000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3