Solution Architect (Kernel Optimization & ML Performance)

EPAM Systems

Town of Poland (NY)

On-site

USD 140,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

EPAM Systems is seeking a Solution Architect specializing in Kernel Optimization and ML Performance. This role requires extensive experience in software engineering and architecture, focusing on high-performance machine learning solutions. As a key member, you’ll define architectures for ML workloads, collaborate with global teams, and mentor development teams.

The ideal candidate will have a Bachelor’s or Master’s degree in Computer Science, along with in-depth knowledge of machine learning frameworks and performance optimization techniques.

Qualifications

  • 12+ years of software engineering experience, including 5+ years in architecture or technical leadership roles.
  • In-depth knowledge of ML frameworks and Core ML concepts.
  • Hands-on expertise in performance optimization at the kernel level targeting TPUs/GPUs.

Responsibilities

  • Define and own the architecture for performance-critical ML workloads.
  • Create a strategic roadmap for kernel optimization and performance improvement.
  • Provide technical leadership and mentorship to development teams.

Skills

ML frameworks (JAX, PyTorch, TensorFlow)
Performance optimization
C++/Python for high-performance computing
Communication skills

Education

Bachelor’s or Master’s degree in Computer Science

Tools

MLIR
OpenXLA

Job description

Join us as a Solution Architect (Kernel Optimization & ML Performance) and play a key role in shaping impactful, scalable solutions that drive real business value. You will work at the intersection of advanced hardware acceleration and machine learning infrastructure, translating complex performance and optimization requirements into robust architectural designs. This role offers the opportunity to collaborate with global, cross‑functional teams in a dynamic and fast‑paced environment, working closely with ML researchers, compiler engineers, and systems architects. You’ll have a direct influence on the technical direction of high‑performance ML solutions, solution quality, and overall efficiency across large‑scale AI workloads. At EPAM, we value innovation, ownership, and a proactive mindset, giving you the space to make a tangible technological impact. If you are ready to take on a strategic architect‑level role in AI infrastructure and grow your career in a global setting, we encourage you to apply.

Responsibilities
  • Define and own the architecture for performance‑critical ML workloads, leveraging custom kernels on TPUs and GPUs.
  • Create a strategic roadmap for kernel optimization, framework integration, and large‑scale performance improvement.
  • Collaborate with the client’s technical leadership, ML researchers, and engineers to capture requirements and design scalable solutions.
  • Evaluate and select the right technologies, toolchains, and design patterns to optimize compute‑intensive operations.
  • Guide the development of benchmarking infrastructure, autotuning frameworks, performance profiling tools, and regression suites.
  • Advocate ML performance best practices, influencing framework and compiler enhancements across teams.
  • Provide technical leadership and mentorship to development teams implementing architectural solutions.
  • Ensure adherence to security, scalability, and maintainability standards in solution design.
Requirements
  • Bachelor’s or Master’s degree in Computer Science or equivalent practical experience.
  • 12+ years of software engineering experience, including 5+ years in architecture or technical leadership roles.
  • In‑depth knowledge of ML frameworks (JAX, PyTorch, TensorFlow) and Core ML concepts.
  • Hands‑on expertise in performance optimization at the kernel level targeting TPUs/GPUs.
  • Strong experience with C++/Python for high‑performance computing.
  • Solid grasp of compiler principles, graph optimizations, and toolchains such as MLIR or OpenXLA.
  • Proven experience designing scalable, production‑grade ML systems with emphasis on performance and efficiency.
  • Excellent communication, stakeholder management, and solution delivery skills.
  • Nice to have: familiarity with emerging hardware accelerators, heterogeneous compute, and scale‑out performance patterns.
  • Experience with autotuning systems, benchmarking methodologies, and performance profiling tools.
  • Contributions to open‑source projects in ML performance, kernels, or developer infrastructure.
  • Demonstrated ability to present architectural solutions to technical and non‑technical audiences.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Solution Architect - ML Kernel & Performance
Senior Solution Architect - ML Kernel & Performance

EPAM Systems • Town of Poland (NY)

On-site
USD 140,000 - 180,000
Lead Kernel Engineer/Architect (m/f/d)
Lead Kernel Engineer/Architect (m/f/d)

EPAM Systems • Germany (OH)

Hybrid
USD 104,000 - 152,000
System Performance & AI Architect
System Performance & AI Architect

Majestic Labs ai • Los Altos (CA)

On-site
USD 140,000 - 190,000
Senior AI Accelerator Performance Leader
Senior AI Accelerator Performance Leader

EPAM Systems • New York (NY)

On-site
USD 180,000 - 220,000
Medical, Dental and Vision Insurance
Health Savings Account
Flexible Spending Accounts
+10
Senior ML Performance Engineer
Senior ML Performance Engineer

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Competitive salary
Equity and bonus opportunities
Medical, dental, and vision coverage
+2
Lead Kernel Architect for AI Accelerators
Lead Kernel Architect for AI Accelerators

EPAM Systems • Germany (OH)

Hybrid
USD 104,000 - 152,000
Computer Architect
Computer Architect

Acceler8 Talent • Mountain View (CA)

Hybrid
USD 180,000 - 300,000
Performance Architect - AI Hardware
Performance Architect - AI Hardware

TEEMA • United States

Remote
TPU Kernel Engineer for High-Performance ML Systems
TPU Kernel Engineer for High-Performance ML Systems

SignalAI • New York (NY)

Hybrid
USD 280,000 - 850,000
AI Platform Architect
AI Platform Architect

Cerebras • Austin (TX)

On-site