AI ML Compiler Developer NPU Acceleration

AMD

Hyderabad

On-site

INR 3,000,000 - 6,000,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AMD Hyderabad seeks a Lead/Staff Software Engineer to develop AI/ML kernels in C/C++ for XDNA NPU and map LLMs/Stable Diffusion on AMD Ryzen platforms. You will optimize vector processors, collaborate across teams, and ensure performance and correctness of ML operators.

The role emphasizes kernel design, testing, and integration with the software stack, with emphasis on SIMD ILP parallelism and hardware awareness.

Qualifications

  • BS/MS/PhD with 5/8/10 years of relevant experience respectively.
  • Excellent C/C++ and Python coding skills.
  • Experience with vectorized programming (SIMD) and parallel computing.
  • Familiarity with ML frameworks like TensorFlow or PyTorch.
  • Knowledge of hardware details and pre-silicon validation is a plus.

Responsibilities

  • Design and implement highly optimized C++ kernels for NPU/GPU.
  • Collaborate with research and software teams to integrate kernels into the stack.
  • Optimize for vector processors and ILP/SIMD execution.
  • Develop tests and validate ML operators in C++/Python.
  • Profile and tune kernel performance across hardware platforms.
  • Document design specs and maintain code with Git.

Skills

C/C++
Python
SIMD
Vector processors
Parallel computing
ML frameworks
MLU/Emulation platforms
Problem solving

Education

BS in Computer Science / Electrical Engineering
MS in Computer Science / Electrical Engineering
PhD in Computer Science / Electrical Engineering

Tools

TensorFlow
PyTorch

Job description

Job Description:

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.

Whether you’re designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we’ll advance your career.

MTS SOFTWARE DEVELOPMENT ENGINEER
THE ROLE:

We are looking for a dynamic, energetic Lead / Staff Software Engineer to join our growing team in AI (Artificial Intelligence) group. In this role, the individual will be responsible for developing AI/ML specific C/C++ kernels and dataflow schedules for AMD Ryzen Processors built on XDNA Neural Processor Units (NPU) to map LLMs, Stable Diffusion networks on NPU. As a C++ Kernel Developer, you will play a crucial role in designing, optimizing, and implementing machine learning kernels specifically tailored for vector processors. Your work will directly impact the efficiency, speed, and accuracy of our machine learning models.

Key Responsibilities
  • Kernel Development: Design and implement highly optimized C++ kernel library for NPU/GPU.
  • Collaborate with the research and software teams to integrate these kernels into the existing software stack.
  • Vector Processor Optimization: Work closely with hardware engineers to understand the architecture of VLIW vector core units such as MAC, GeMM, and non-linear functions.
  • Develop vectorized code that leverages SIMD (Single Instruction, Multiple Data) and ILP (instruction level parallelism) for maximum performance.
  • Performance Profiling and Tuning: Profile and analyze the performance of existing kernels.
  • Identify bottlenecks and optimize critical sections for better throughput.
  • Testing and Validation: Develop CPU models for the ML operators in C++/ Python to validate accuracy.
  • Write unit tests and integration tests to ensure correctness and reliability.
  • Validate kernel performance across different hardware platforms.
  • Documentation and Collaboration: Document design specs for new kernels and the performance improvements.
  • Follow coding guidelines, use tools like git to maintain code and create pull-requests, and documentation.
  • Collaborate with cross-functional teams, including machine learning researchers and software engineers.
PREFERRED EXPERIENCE
  • Excellent C/C++ and Python coding skills
  • Good understanding of SIMD/Tensor/VLIW processor architecture to exploit parallelism.
  • Experience with vectorized programming (SIMD) and parallel computing.
  • Familiarity with machine learning frameworks (e.g., TensorFlow, PyTorch) is a plus.
  • Experience with silicon bring-up and pre-silicon validation on Emulation platforms is a plus.
  • Knowledge of low-level hardware details (cache hierarchy, memory access patterns) is desirable.
  • Excellent problem-solving skills and a passion for performance optimization.

BS/Masters/PhD degree in Computer Science, Electrical Engineering, or a related field with around 10/8/5 year experience respectively.

Benefits offered are described: AMD benefits at a glance.

AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI/ML Compiler Developer (NPU Acceleration)
AI/ML Compiler Developer (NPU Acceleration)

Advanced Micro Devices • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Senior AI/ML - NPU Validation Engineer
Senior AI/ML - NPU Validation Engineer

AMD • Hyderabad

On-site
INR 2,000,000 - 2,800,000
Senior AI/ML - NPU Validation Engineer
Senior AI/ML - NPU Validation Engineer

Advanced Micro Devices • Hyderabad

On-site
INR 2,400,000 - 3,600,000
Lead Machine Learning Engineer
Lead Machine Learning Engineer

AMD • Telangana

On-site
INR 4,000,000 - 7,000,000
AMD benefits at a glance
Lead Machine Learning Engineer
Lead Machine Learning Engineer

Advanced Micro Devices • Hyderabad

On-site
INR 3,500,000 - 5,200,000
Lead Machine Learning Engineer
Lead Machine Learning Engineer

AMD • Hyderabad

On-site
INR 2,500,000 - 4,500,000
AI Engineer – Model Optimization & Acceleration
AI Engineer – Model Optimization & Acceleration

AMD • Bengaluru

On-site
INR 2,500,000 - 4,500,000
Lead Performance and Optimization Engineer
Lead Performance and Optimization Engineer

AMD • Bengaluru Urban

On-site
INR 900,000 - 1,500,000
Senior/Lead Linux Kernel Development Engineer
Senior/Lead Linux Kernel Development Engineer

Advanced Micro Devices • Bengaluru

On-site
INR 1,500,000 - 2,000,000
AI Engineer – Model Optimization & Acceleration
AI Engineer – Model Optimization & Acceleration

AMD • Bengaluru Urban

On-site
INR 2,000,000 - 2,800,000