Lead II - Embedded Software

UST

Bengaluru

On-site

INR 1,000,000 - 1,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

UST in Bengaluru is seeking a skilled professional for benchmarking AI models, optimizing performance, and analyzing system-level metrics. This role involves working with various AI models and hardware platforms, offering insights from performance comparisons.

The ideal candidate will have experience in performance benchmarking, AI/ML framework knowledge, and programming skills in C/C++ and Python. A strong understanding of Linux systems and profiling tools is required.

Qualifications

  • Strong experience in performance benchmarking and system analysis.
  • Good understanding of AI/ML models and inference frameworks (PyTorch, ONNX, etc.).
  • Experience with Linux systems and performance tools.
  • Knowledge of CPU/GPU/NPU architectures and memory systems.
  • Experience with profiling tools such as perf, VTune, and Nsight.
  • Programming proficiency in C/C++, Python, or Perl.

Responsibilities

  • Benchmark AI models such as OpenVLA, YOLOv8, and Whisper.
  • Measure performance metrics including latency and throughput.
  • Set up and run benchmarking pipelines across hardware platforms.
  • Perform system-level analysis to identify performance bottlenecks.
  • Optimize AI models using profiling tools.
  • Integrate AI models into application pipelines and measure performance.
  • Compare performance across platforms and generate insights.
  • Automate benchmarking execution and data collection.
  • Prepare technical reports and performance summaries.

Skills

Performance benchmarking
AI/ML models understanding
Linux systems
CPU/GPU/NPU architectures
Profiling tools
C/C++ programming
Python programming
Benchmarking
Graphics processing unit knowledge

Tools

perf
VTune
Nsight

Job description

Role Description

Benchmark AI models such as OpenVLA, YOLOv8, Whisper, Pi Droid, Groot, and MLPerf Client and similar workloads.

Measure performance metrics including latency, throughput (FPS), memory usage, and power efficiency.

Set up and run repeatable benchmarking pipelines across different hardware platforms.

Perform system‑level analysis to identify bottlenecks in CPU, GPU, memory, and data pipelines.

Optimize AI models and pipelines using profiling tools and performance tuning techniques.

Integrate AI models into application pipelines (e.g., robotics or edge AI workflows) and measure end‑to‑end performance.

Compare performance across platforms (AMD, NVIDIA, Qualcomm, etc.) and generate competitive insights.

Automate benchmark execution, data collection, and reporting.

Prepare clear technical reports and performance summaries for stakeholders.

Key Responsibilities
  • Benchmark AI models such as OpenVLA, YOLOv8, Whisper, Pi Droid, Groot, and MLPerf Client and similar workloads.
  • Measure performance metrics including latency, throughput (FPS), memory usage, and power efficiency.
  • Set up and run repeatable benchmarking pipelines across different hardware platforms.
  • Perform system‑level analysis to identify bottlenecks in CPU, GPU, memory, and data pipelines.
  • Optimize AI models and pipelines using profiling tools and performance tuning techniques.
  • Integrate AI models into application pipelines (e.g., robotics or edge AI workflows) and measure end‑to‑end performance.
  • Compare performance across platforms (AMD, NVIDIA, Qualcomm, etc.) and generate competitive insights.
  • Automate benchmark execution, data collection, and reporting.
  • Prepare clear technical reports and performance summaries for stakeholders.
Required Skills
  • Strong experience in performance benchmarking and system analysis.
  • Good understanding of AI/ML models and inference frameworks (PyTorch, ONNX, etc.).
  • Experience with Linux systems and performance tools.
  • Knowledge of CPU/GPU/NPU architectures and memory systems.
  • Experience with profiling tools such as perf, VTune, Nsight or similar.
  • Programming proficiency in C/C++, Python/Peral.
Skills
  • Embedded software development
  • Graphics processing unit
  • Robotics
  • Artificial intelligence
  • Benchmarking
  • Python
  • C++
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Benchmarking & Performance Engineer
AI Benchmarking & Performance Engineer

Sunrise Biztech Systems • Bangalore Rural

On-site
INR 1,400,000 - 2,300,000
AI Engineer Model Optimization & Acceleration
AI Engineer Model Optimization & Acceleration

Sunrise Biztech Systems • Bangalore Rural

On-site
INR 1,200,000 - 2,400,000
Senior AI Software Performance Engineer
Senior AI Software Performance Engineer

BigStep Technologies • Gurugram District

On-site
INR 1,500,000 - 2,000,000
AIML, Python with Linux
AIML, Python with Linux

Tata Consultancy Services • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Senior ML/AI Engineer
Senior ML/AI Engineer

Syren Cloud Inc. • Hyderabad

On-site
INR 1,800,000 - 3,000,000
AI Benchmark Engineer (Data Analysis)
AI Benchmark Engineer (Data Analysis)

Turing • Delhi

Remote
INR 1,200,000 - 1,800,000
Work on cutting-edge AI projects
Collaborate on high-impact work
Flexible opportunities with global teams
Artificial Intelligence Engineer
Artificial Intelligence Engineer

LeadSoc Technologies Pvt Ltd • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Edge AI deployment
Generative AI systems
Distributed inference pipelines
Senior Performance Engineer - DGX Cloud
Senior Performance Engineer - DGX Cloud

NVIDIA AI • Hyderabad

On-site
INR 2,500,000 - 4,000,000
AI Architect
AI Architect

Larsen & Toubro • Chennai District

On-site
INR 4,000,000 - 7,000,000
AI Infrastructure and Platform Architect
AI Infrastructure and Platform Architect

Ignatiuz Inc. • Indore District

On-site
INR 3,500,000 - 6,000,000