AI Systems Performance Specialist

Bright Vision Technologies

Fishers (IN, KY)

Remote

USD 130.000 - 180.000

Vollzeit

Vor 4 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die darauf ausgerichtet sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Bright Vision Technologies is seeking an experienced AI Systems Performance Specialist to optimize AI training and inference workloads across enterprise-scale platforms. You will lead GPU optimization, multi-GPU training, and production AI system enhancements, with a strong focus on CUDA, Python, and C++.

The role requires 10+ years in AI infrastructure, HPC, or performance engineering, plus hands-on expertise with distributed frameworks and cloud deployments (AWS/Azure/GCP).

Qualifikationen

  • 10+ years of professional experience in performance engineering.
  • Experience with GPU-accelerated AI workloads and HPC.
  • Proficiency in Python and C++.
  • Experience deploying AI workloads on AWS/Azure/GCP.
  • Strong leadership and communication skills.

Aufgaben

  • Optimize AI training and inference pipelines for throughput, latency, and scale.
  • Analyze and improve GPU utilization, memory, and multi-GPU performance.
  • Implement optimization techniques (quantization, pruning, mixed precision, batching).
  • Profile AI apps with industry tools and identify bottlenecks across compute, memory, networking, storage.
  • Coordinate with researchers, ML engineers, and infra teams to improve production AI performance.
  • Build benchmarking frameworks, performance dashboards, and regression tests.
  • Evaluate new AI hardware and frameworks to improve enterprise AI capabilities.
  • Drive AI infrastructure cost optimization and FinOps best practices.
  • Mentor teams on AI systems architecture and GPU optimization.

Kenntnisse

Python
C++
Performance engineering
GPU optimization

Ausbildung

Bachelor's or Master's in CS/CE/EE/AI

Tools

CUDA
NCCL
DeepSpeed
Triton Inference Server
FasterTransformer
vLLM
TensorRT-LLM
MPI

Jobbeschreibung

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title

AI Systems Performance Specialist

Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $130,000–$180,000 Annually
Experience Required: 10+ Years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

Bright Vision Technologies is seeking a highly experienced AI Systems Performance Specialist with 10+ years of experience in AI infrastructure, machine learning systems, High-Performance Computing (HPC), and performance engineering. The ideal candidate will optimize AI training and inference workloads for maximum performance, scalability, reliability, and cost efficiency. This role requires deep expertise in GPU optimization, distributed training, Large Language Model (LLM) inference, Python, C++, CUDA, and production AI systems, along with the ability to lead performance optimization initiatives across enterprise-scale AI platforms.

Key Responsibilities
  • Optimize AI training and inference pipelines for maximum throughput, low latency, scalability, and infrastructure efficiency.
  • Analyze and improve GPU utilization, memory management, kernel execution, and multi-GPU performance across production AI workloads.
  • Design and implement optimization techniques including quantization, pruning, mixed precision, batching, caching, speculative decoding, and model parallelism.
  • Profile AI applications using industry-standard performance analysis tools and identify bottlenecks across compute, memory, networking, and storage.
  • Optimize distributed training and inference using NCCL, DeepSpeed, PyTorch Distributed, Ray, MPI, or similar distributed computing framework.
  • Collaborate with AI researchers, ML engineers, platform engineers, and infrastructure teams to improve model performance and production reliability.
  • Build automated benchmarking frameworks, performance dashboards, monitoring solutions, and regression testing pipelines.
  • Evaluate emerging AI hardware, GPU architectures, inference frameworks, and optimization technologies to improve enterprise AI capabilities.
  • Drive AI infrastructure cost optimization through efficient resource utilization, cloud optimization, and FinOps best practices.
  • Mentor engineering teams and provide technical leadership on AI systems architecture, GPU optimization, and performance engineering.
Required Qualifications
  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, Artificial Intelligence, or a related technical discipline.
  • 10+ years of professional experience in performance engineering, AI infrastructure, machine learning systems, High-Performance Computing (HPC), or distributed computing.
  • Expert-level programming skills in Python and C++.
  • Extensive experience optimizing GPU-accelerated AI workloads using CUDA, distributed training frameworks, and modern deep learning libraries.
  • Strong knowledge of Large Language Models (LLMs), deep learning frameworks, model serving, and production AI inference.
  • Hands-on experience with profiling tools such as NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, TensorBoard, or similar performance analysis tools.
  • Experience deploying and optimizing AI workloads on AWS, Microsoft Azure, or Google Cloud Platform (GCP).
  • Strong understanding of distributed systems, networking, storage optimization, and AI infrastructure architecture.
  • Excellent analytical, troubleshooting, communication, and technical leadership skills.
Preferred Qualifications
  • Experience optimizing production-scale LLM inference and serving large foundation models.
  • Hands-on experience with vLLM, TensorRT-LLM, DeepSpeed, Triton Inference Server, CUTLASS, FasterTransformer, or similar AI optimization frameworks.
  • Knowledge of model compression, KV cache optimization, speculative decoding, and advanced inference optimization techniques.
  • Experience implementing FinOps strategies for AI infrastructure cost optimization and resource management.
  • Contributions to AI systems research, open-source AI infrastructure projects, patents, or technical publications.
  • Familiarity with emerging AI accelerator technologies, including AMD ROCm, Intel oneAPI, or custom AI hardware.

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI Performance Engineer
AI Performance Engineer

Bright Vision Technologies • Naperville (IL)

Remote
USD 75.000 - 100.000
AI Infrastructure Engineer
AI Infrastructure Engineer

Bright Vision Technologies • Tampa (FL)

Remote
USD 100.000 - 150.000
ML Performance Engineer
ML Performance Engineer

Bright Vision Technologies • Plymouth (MN)

Remote
USD 100.000 - 150.000
AI Optimization Engineer
AI Optimization Engineer

Bright Vision Technologies • USA

Remote
USD 85.000 - 115.000
AI Systems Engineer
AI Systems Engineer

Bright Vision Technologies • USA

Remote
USD 90.000 - 100.000
AI Systems Engineer
AI Systems Engineer

Bright Vision Technologies • Elk Grove (CA)

Remote
USD 90.000 - 100.000
High-Performance Computing Engineer
High-Performance Computing Engineer

Bright Vision Technologies • USA

Remote
USD 130.000 - 150.000
Parallel Computing Engineer
Parallel Computing Engineer

Bright Vision Technologies • Fishers (IN)

Remote
USD 130.000 - 180.000
Parallel Computing Engineer
Parallel Computing Engineer

Bright Vision Technologies • USA

Remote
USD 130.000 - 180.000
AI Infrastructure Engineer
AI Infrastructure Engineer

Bright Vision Technologies • Edison (NJ)

Remote
USD 100.000 - 160.000