Technical Writer (Infrastructure L3 Support Team)

Linuxcareers

Amsterdam

On-site

EUR 50,000 - 80,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Nebius is seeking a Technical Writer for the Infrastructure L3 Support team to build and maintain a knowledge system for server hardware, GPU platforms, firmware, and Linux diagnostics across the global data center fleet.

You will translate complex engineering investigations into clear, repeatable procedures for L1/L2 technicians, ensure governance of documentation, and support hardware readiness with hands-on in-data-center work across EMEA and beyond.

Qualifications

  • Hands-on experience in data center, server infrastructure, production operations, or site reliability engineering.
  • Proficient in Linux, server hardware, firmware, and out-of-band management technologies such as IPMI, BMC, OpenBMC, or Redfish.
  • Demonstrated record of creating runbooks, SOPs, or troubleshooting guides successfully used by operations teams.
  • Experience translating complex engineering investigations into safe, repeatable operational procedures with strong organizational and writing skills.
  • Fluent written and spoken English and willingness to travel regularly to data center locations.

Responsibilities

  • Build and maintain the L3 knowledge base for GPU server platforms, server hardware, firmware, out-of-band management, and Linux-level diagnostics, including runbooks, SOPs, troubleshooting guides, and error-code documentation.
  • Work with L3 and R&D engineers during investigations to capture symptoms, root causes, and resolutions, translating complex technical findings into clear, repeatable operational procedures for L1 and L2 technicians.
  • Write layered documentation for audiences at different technical levels, explaining complex concepts with appropriate terminology, diagrams, examples, and clear prerequisites, warnings, decision points, and success criteria.
  • Validate procedures end-to-end in appropriate environments and test with representative L1 and L2 users, using their feedback and escalation patterns to continuously improve the knowledge base.
  • Define and maintain documentation governance including templates, quality standards, metadata, ownership rules, approval workflows, review cycles, and document status tracking.
  • Support new hardware platform readiness by creating complete documentation packages and traveling to data centers in EMEA and other locations to observe procedures and capture operational knowledge.

Skills

Technical writing
Knowledge base management
Runbooks & SOPs
Linux diagnostics documentation
Fluent English
Data center operations understanding

Tools

Git-based workflows
Documentation platforms
Log analysis tooling

Job description

Nebius is a Nasdaq-listed AI cloud infrastructure company building full-stack platforms for GPU orchestration and AI deployment. This role is a Technical Writer for the Infrastructure L3 Support team, responsible for building and maintaining the knowledge system for server hardware, GPU platforms, firmware, and Linux diagnostics across the global data center fleet.

What You’ll Do
  • Build and maintain the L3 knowledge base for GPU server platforms, server hardware, firmware, out-of-band management, and Linux-level diagnostics, including runbooks, SOPs, troubleshooting guides, and error-code documentation
  • Work with L3 and R&D engineers during investigations to capture symptoms, root causes, and resolutions, translating complex technical findings into clear, repeatable operational procedures for L1 and L2 technicians
  • Write layered documentation for audiences at different technical levels, explaining complex concepts with appropriate terminology, diagrams, examples, and clear prerequisites, warnings, decision points, and success criteria
  • Validate procedures end-to-end in appropriate environments and test with representative L1 and L2 users, using their feedback and escalation patterns to continuously improve the knowledge base
  • Define and maintain documentation governance including templates, quality standards, metadata, ownership rules, approval workflows, review cycles, and document status tracking
  • Support new hardware platform readiness by creating complete documentation packages and traveling to data centers in EMEA and other locations to observe procedures and capture operational knowledge
What You Need
  • Hands-on experience in data center, server infrastructure, production operations, or site reliability engineering
  • Working knowledge of Linux, server hardware, firmware, and out-of-band management technologies such as IPMI, BMC, OpenBMC, or Redfish
  • Demonstrated record of creating runbooks, SOPs, or troubleshooting guides successfully used by operations teams
  • Experience translating complex engineering investigations into safe, repeatable operational procedures with strong organizational and writing skills
  • Fluent written and spoken English and willingness to travel regularly to data center locations
Nice to Have
  • Experience with NVIDIA GPU server platforms and tools such as nvidia-smi, DCGM, dcgmi, and log-correlation tooling
  • Experience with HGX or other large-scale AI infrastructure platforms
  • Exposure to OCP-based platforms or ODM manufacturing ecosystems
  • Experience using Bash or Python for log collection, diagnostics, or operational automation
  • Experience with documentation-as-code, Git-based workflows, wikis, or large-scale knowledge-base platforms
  • Experience defining documentation metrics or using incident and escalation data to prioritize improvements

Competitive compensation, career growth and learning opportunities, flexibility and ownership, collaborative and innovative culture

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Writer — AI Infra & GPU Server Ops
Technical Writer — AI Infra & GPU Server Ops

Linuxcareers • Amsterdam

On-site
EUR 50,000 - 80,000
Technical Writer (Infrastructure L3 Support Team)
Technical Writer (Infrastructure L3 Support Team)

Nebius • Amsterdam

On-site
EUR 70,000 - 110,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
L3 Infrastructure Documentation Engineer
L3 Infrastructure Documentation Engineer

Nebius • Amsterdam

On-site
EUR 70,000 - 110,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
Technical Project Manager (Hardware)
Technical Project Manager (Hardware)

Nebius • Amsterdam

On-site
EUR 90,000 - 130,000
Competitive pay
Career growth
Flexibility
+3
Infrastructure Engineer
Infrastructure Engineer

Nebul • Leiden

On-site
EUR 55,000 - 75,000
Senior Technical Program Manager - New Data Center Launches
Senior Technical Program Manager - New Data Center Launches

Nebius • Amsterdam

On-site
EUR 75,000 - 95,000
Competitive salary
Professional growth opportunities
Flexible working arrangements
+1
AI/ML Specialist Solutions Architect
AI/ML Specialist Solutions Architect

Meyandy LLC • Amsterdam

Hybrid
EUR 120,000 - 190,000
Competitive compensation
Career growth opportunities
Flexible work arrangement
+1
Technical Product Manager – AI Compute Platform
Technical Product Manager – AI Compute Platform

ApplyMint • Netherlands

Hybrid
EUR 70,000 - 100,000
Competitive compensation
Career growth opportunities
Flexible work environment
Service Delivery Manager
Service Delivery Manager

Nebius B.V. • Amsterdam

On-site
EUR 90,000 - 120,000
Senior ML Engineer (Token Factory)
Senior ML Engineer (Token Factory)

Nebius • Amsterdam

On-site
EUR 70,000 - 90,000
Competitive compensation
Career growth opportunities
Collaborative culture
+1