Datacenter & Agentic AI Workload Performance Optimization Engineer

Tenstorrent

California

A distancia

USD 100.000 - 500.000

Jornada completa

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Consigue una respuesta de este empleador — un currículum y una carta de presentación adaptados exactamente a lo que busca para contratar.

Supera los filtros ATS

Descripción de la vacante

Tenstorrent is seeking a Workload Performance Optimization Engineer to optimize software workloads on our next‑gen RISC‑V platforms. You will identify bottlenecks across runtimes, compilers, and hardware, and implement optimizations to improve throughput, latency, and energy efficiency on data center AI workloads.

Role spans Java, Python, C/C++, and Rust, with opportunities to profile real applications, explore Vector/Matrix capabilities, and contribute to AI-assisted automation of profiling and

Formación

  • Master’s or PhD in Computer Engineering, Electrical Engineering, Computer Science, or a related field, with strong experience in performance optimization, computer architecture, compilers, or systems software.
  • Hands‑on experience with runtime or compiler optimization, such as OpenJDK/JIT, LLVM, GCC, V8, Python, or equivalent systems.
  • Strong understanding of CPU performance, memory hierarchies, concurrency, garbage collection, vector/SIMD optimization, and RISC-V architecture.
  • Expertise with performance profiling and analysis tools such as Linux perf, runtime profilers, QEMU, tracing tools, and performance modeling environments.
  • Strong programming skills in Java, Python, C/C++, and RISC-V assembly, with the ability to work effectively across multiple software layers.

Conocimientos

Performance engineering
RISC-V experience
Profiling & optimization
Cross-layer collaboration

Educación

MS/PhD in Computer Engineering/CS/EE
Advanced degree in CS/EE

Herramientas

Linux perf
QEMU
LLVM
GCC

Descripción del empleo

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.

Tenstorrent is looking for a Workload Performance Optimization Engineer to help optimize the software workloads that run on our next-generation RISC-V platforms. You’ll work across modern datacenter and agentic AI workloads—including Java, Python, PHP, Node.js, Lua, Go, and Rust—to identify performance bottlenecks and develop software, compiler, runtime, and hardware-aware optimizations that improve throughput, latency, and efficiency. This role sits at the intersection of software runtimes, compilers, CPU microarchitecture, and RISC-V silicon. You’ll bring up and tune major runtimes, profile real-world applications, investigate memory and concurrency behavior, and explore optimizations using RISC-V Vector/Matrix capabilities and custom instructions. You’ll also work with AI-assisted development and automated optimization workflows to accelerate the performance engineering process. Your work will directly influence both the RISC-V software ecosystem and the architecture of future Tenstorrent CPUs.

This role isremote, based out of North America.

We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting.

Who You Are

  • You’re a performance engineer who enjoys getting deep into runtimes, compilers, applications, and CPU microarchitecture to understand why software is fast—or slow.
  • You have hands‑on experience optimizing software on RISC-V or another modern CPU architecture, with a strong understanding of the hardware/software boundary.
  • You’re comfortable profiling complex systems, finding bottlenecks, forming hypotheses, and iterating through optimizations using data.
  • You’re excited about emerging agentic AI development workflows and using AI tools to automate profiling, coding, benchmarking, and optimization.
  • You’re a strong technical collaborator who can work across compiler, runtime, systems software, hardware, and performance modeling teams.

What We Need

  • Master’s or PhD in Computer Engineering, Electrical Engineering, Computer Science, or a related field, with strong experience in performance optimization, computer architecture, compilers, or systems software.
  • Hands‑on experience with runtime or compiler optimization, such as OpenJDK/JIT, LLVM, GCC, V8, Python, or equivalent systems.
  • Strong understanding of CPU performance, memory hierarchies, concurrency, garbage collection, vector/SIMD optimization, and RISC-V architecture.
  • Expertise with performance profiling and analysis tools such as Linux perf, runtime profilers, QEMU, tracing tools, and performance modeling environments.
  • Strong programming skills in Java, Python, C/C++, and RISC-V assembly, with the ability to work effectively across multiple software layers.

What You Will Learn

  • How to optimize modern software stacks from application and runtime all the way down to CPU microarchitecture and silicon.
  • How RISC-V Vector, Matrix, and custom ISA capabilities can be used to accelerate real-world datacenter and AI workloads.
  • How runtime, compiler, memory, and concurrency decisions impact performance at datacenter scale.
  • How to build automated and AI-assisted performance optimization workflows that continuously profile, analyze, modify, and benchmark software.
  • How to influence future CPU architecture by connecting real workload behavior and software optimization opportunities to hardware design decisions.

Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made.

Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.

This offer of employment is contingent upon the applicant being eligible to access U.S. export‑controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Datacenter & Agentic AI Workload Performance Optimization Engineer
Datacenter & Agentic AI Workload Performance Optimization Engineer

Tenstorrent • EE. UU.

A distancia
USD 100.000 - 500.000
RISC V Sr. Engineer, Performance Infrastructure (RISC-V) Austin, Texas, United States; Santa Clara, California, United States
RISC V Sr. Engineer, Performance Infrastructure (RISC-V) Austin, Texas, United States; Santa Clara, California, United States

Tenstorrent Inc. • Town of Texas (WI), Northern (KY)

Híbrido
USD 100.000 - 500.000
Principal CPU Microarchitect - RISC-V & AI Compute
Principal CPU Microarchitect - RISC-V & AI Compute

Tenstorrent • California (MO)

Presencial
USD 180.000 - 260.000
Hybrid work model
Equal opportunity employer
Sr. Engineer, CPU RTL Design
Sr. Engineer, CPU RTL Design

Tenstorrent • California (MO)

Presencial
USD 100.000 - 500.000
Highly competitive compensation package
Benefits and equal opportunity employer
Fabric SOC Architect
Fabric SOC Architect

Tenstorrent • Santa Clara (CA)

A distancia
USD 100.000 - 500.000
Staff Engineer,Post-Silicon Validation
Staff Engineer,Post-Silicon Validation

Tenstorrent University Jobs • Santa Clara (CA)

Híbrido
USD 100.000 - 500.000
Sr. Performance Modeling Architect
Sr. Performance Modeling Architect

Tenstorrent • Austin (CA)

Presencial
USD 100.000 - 500.000
Competitive compensation
Benefits package
Hybrid work model
Sr. Engineer, RTL Implementation
Sr. Engineer, RTL Implementation

Tenstorrent Inc. • Austin (TX), Santa Clara (CA)

Presencial
USD 100.000 - 500.000
Staff Engineer,Post-Silicon Validation
Staff Engineer,Post-Silicon Validation

Tenstorrent • Town of Texas (WI)

Híbrido
USD 100.000 - 500.000
Fabric SOC Architect
Fabric SOC Architect

Tenstorrent • EE. UU.

A distancia
USD 100.000 - 500.000