Compiler Engineer

TetraMem - Accelerate The World

San Jose (CA)

On-site

USD 160,000 - 300,000

Full time

12 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

TetraMem in San Jose, California, is seeking a Senior Compiler Engineer to design and maintain compiler toolchains that translate ML models from common frameworks into optimized workloads for our analog in-memory computing hardware.

You will develop runtimes, libraries, and SDKs to enable efficient deployment of AI applications on TetraMem accelerators, implement advanced optimizations, and collaborate with ML engineers and hardware teams to improve performance and usability.

Qualifications

  • MS or PhD in Computer Engineering/CS/EE.
  • 5+ years industry experience as a compiler engineer or developer.
  • Experience developing compilers for GPU, dataflow compilers, or ML compilers.
  • Startup mindset/experience.

Responsibilities

  • Design, develop, and maintain compiler toolchains that translate ML models into optimized workloads for TetraMem’s analog in-memory computing hardware.
  • Develop runtime systems, software libraries, and SDK components for AI apps on TetraMem accelerators.
  • Implement compiler optimizations including graph transformations, operator fusion, memory optimization, scheduling, and code generation.
  • Research and develop techniques to improve inference speed, latency, throughput, and power across AI workloads.
  • Collaborate with ML engineers to support model conversion, validation, optimization, benchmarking, and deployment.
  • Partner with hardware architects to co-design software/hardware features for performance and usability.
  • Develop performance analysis, profiling, debugging, and benchmarking tools for AI workloads on TetraMem platforms.
  • Integrate and support frameworks like PyTorch, TensorFlow, ONNX, etc.
  • Lead technical design reviews and set best practices for scalable software development.
  • Mentor junior engineers and help define the long-term roadmap for compiler, runtime, and SDK technologies.

Skills

Compiler engineering
GPU/ML compilers
Startup mindset
Leadership
Team collaboration

Education

MS or PhD in Computer Engineering/CS/EE

Tools

GCC
Clang
MLIR
LLVM

Job description

Responsibilities
  • Design, develop, and maintain compiler toolchains that translate machine learning models from industry-standard frameworks into optimized workloads for TetraMem’s analog in-memory computing hardware.
  • Develop runtime systems, software libraries, and SDK components that enable efficient deployment, execution, and management of AI applications on TetraMem accelerators.
  • Implement compiler optimizations, including graph transformations, operator fusion, memory optimization, scheduling, and code generation to maximize performance and energy efficiency.
  • Research and develop innovative techniques to improve machine learning inference speed, latency, throughput, and power consumption across a wide range of AI workloads.
  • Collaborate closely with machine learning engineers to support model conversion, validation, optimization, benchmarking, and deployment.
  • Partner with hardware architects and silicon engineering teams to co-design software and hardware features that improve system performance, programmability, and usability.
  • Develop performance analysis, profiling, debugging, and benchmarking tools to evaluate and optimize AI workloads on current and future TetraMem platforms.
  • Integrate and support industry-standard machine learning frameworks and model formats, including PyTorch, TensorFlow, ONNX, and other emerging AI ecosystems.
  • Lead technical design reviews, contribute to software architecture decisions, and establish best practices for scalable, maintainable, and high-quality software development.
  • Mentor junior engineers, contribute to technical documentation, and help define the long-term roadmap for TetraMem’s compiler, runtime, and SDK technologies.
Responsibilities
  • Design, develop, and maintain compiler toolchains that translate machine learning models from industry-standard frameworks into optimized workloads for TetraMem’s analog in-memory computing hardware.
  • Develop runtime systems, software libraries, and SDK components that enable efficient deployment, execution, and management of AI applications on TetraMem accelerators.
  • Implement compiler optimizations, including graph transformations, operator fusion, memory optimization, scheduling, and code generation to maximize performance and energy efficiency.
  • Research and develop innovative techniques to improve machine learning inference speed, latency, throughput, and power consumption across a wide range of AI workloads.
  • Collaborate closely with machine learning engineers to support model conversion, validation, optimization, benchmarking, and deployment.
  • Partner with hardware architects and silicon engineering teams to co-design software and hardware features that improve system performance, programmability, and usability.
  • Develop performance analysis, profiling, debugging, and benchmarking tools to evaluate and optimize AI workloads on current and future TetraMem platforms.
  • Integrate and support industry-standard machine learning frameworks and model formats, including PyTorch, TensorFlow, ONNX, and other emerging AI ecosystems.
  • Lead technical design reviews, contribute to software architecture decisions, and establish best practices for scalable, maintainable, and high-quality software development.
  • Mentor junior engineers, contribute to technical documentation, and help define the long-term roadmap for TetraMem’s compiler, runtime, and SDK technologies.
Requirements
  • MS or PhD in Computer Engineering/CS/EE
  • 5+ years industry experience as a compiler engineer or developer
  • Experience developing compilers for GPU, dataflow compilers, or ML compilers
  • Startup mindset/experience
Experience in one or more of the following areas considered a strong plus:
  • Experience in RISC-V CPU/VPU kernel development and optimization
  • Experience providing technical leadership and/or guidance to other engineers
  • Knowledge of popular CPU/GPU compilers such as GCC, Clang
  • Knowledge of ML compilers such as MLIR
  • Experience with LLVM and other open-source compiler libraries and tools
  • Publications on compilation of ML or dataflow programs for HW acceleration
Salary Range:

$160,000 - $300,000 / year

TetraMem celebrates diversity and is committed to creating an inclusive environment for all employees. We are proud to be an Equal Opportunity Employer and welcome applicants from all backgrounds. Qualified candidates will receive consideration for employment without regard to race, color, religion, creed, sex, gender identity or expression, sexual orientation, national origin, ancestry, age, marital status, medical condition, disability, genetic information, military or veteran status, or any other characteristic protected by applicable federal, state, or local law.

TetraMem is committed to providing reasonable accommodations to qualified applicants with disabilities throughout the recruitment process. Applicants requiring accommodation may contact Human Resources for assistance.

To ensure a fair, consistent, and efficient hiring process, all candidates must apply through TetraMem’s official ClearCompany Applicant Tracking System (ATS). Applications submitted through the ATS allow our hiring team to evaluate candidates using a standardized process and ensure timely communication throughout the recruitment process. To promote equal consideration for all applicants, applications submitted outside of the ClearCompany ATS, including direct emails, LinkedIn messages, or unsolicited submissions to employees, may not be reviewed or considered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer I – Compiler & Runtime
Software Engineer I – Compiler & Runtime

TetraMem Inc • San Jose (CA)

On-site
USD 135,000 - 165,000
Full-time employee benefits
Equity eligibility
US 2026 Software - Compiler Engineer Intern
US 2026 Software - Compiler Engineer Intern

TetraMem Inc. • San Jose (CA)

On-site
USD 48,000 - 62,000
Compiler Engineer, MTIA Software (Technical Leadership)
Compiler Engineer, MTIA Software (Technical Leadership)

Meta • Menlo Park (CA)

On-site
USD 219,000 - 301,000
Compiler Engineer, MTIA Software (Technical Leadership)
Compiler Engineer, MTIA Software (Technical Leadership)

Meta • New York (NY)

On-site
USD 219,000 - 301,000
System Engineer
System Engineer

TETRAMEM INC • San Jose (CA)

On-site
USD 110,000 - 250,000
Compiler Engineer, Graph Compiler Performance Optimization
Compiler Engineer, Graph Compiler Performance Optimization

Meta • Menlo Park (CA)

On-site
USD 184,000 - 257,000
Principal AI Compiler & Runtime Engineer
Principal AI Compiler & Runtime Engineer

Coley S Home Remodeling • Palo Alto (CA)

Hybrid
USD 200,000 - 350,000
Medical, dental, vision
401(k) with company match
Unlimited vacation
+2
Engineering Manager, ML Compilers (Triton / Kernel DSLs)
Engineering Manager, ML Compilers (Triton / Kernel DSLs)

Socket.dev • Bellevue (WA)

On-site
USD 219,000 - 301,000
Bonus
Equity
Benefits
Compiler Engineer, Hardware
Compiler Engineer, Hardware

River AI • Palo Alto (CA)

On-site
USD 200,000 - 420,000
Health, dental, and vision benefits
Unlimited PTO
Relocation support
Compiler Code Gen Engineer
Compiler Code Gen Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 170,000 - 230,000
Equity
Medical/dental/vision
Retirement savings plan
+1