Nebius is looking for a Sr. Solution Engineer, you will act as the primary technical partner for customers deploying and operating GPU clusters and AI infrastructure on Nebius. You will bridge customer requirements with internal engineering capabilities, ensuring successful deployment, stable operations, and ongoing optimization of complex, high-performance environments.
Your responsibilities will include:
- Acting as the main technical interface for customers running workloads on Nebius GPU infrastructure
- Supporting customers in deploying, configuring, and tuning GPU-based environments for performance and reliability
- Investigating and resolving complex issues spanning hardware, networking, operating systems, and cluster-level behavior
- Partnering with internal teams to coordinate and drive resolution of customer-impacting issues
- Converting customer requirements into practical architectures, configurations, and execution plans
- Identifying opportunities to improve system performance, stability, and overall customer experience
- Developing and maintaining technical documentation, including solution patterns, troubleshooting guides, and operational best practices
- Contributing to continuous improvement by surfacing recurring issues, gaps, and optimization opportunities to internal teams
We expect you to have:
- Experience in a customer-facing technical role
- Strong understanding of GPU infrastructure
- Hands-on experience with Linux systems and system-level troubleshooting
- Familiarity with large-scale compute environments such as GPU clusters, AI infrastructure, or supercomputing systems
- Ability to diagnose issues across hardware, networking, and software layers
- Strong analytical and problem-solving skills
- Excellent communication skills
- A proactive, ownership-driven approach
Key employee benefits include
- Health insurance
- 401(k)
- Parental leave
- Remote work reimbursement
- Disability & life insurance
- Competitive compensation
Compensation: $180K-$220K OTE; RSUs may be available.
Additional summary
- Experience in a customer-facing technical role
- Strong understanding of GPU infrastructure
- Hands-on experience with Linux systems
- Familiarity with large-scale compute environments
- Ability to diagnose issues across hardware, networking and software layers
- Excellent communication skills