AI Chip Toolchain Architect

Forestown

California City (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

A leading tech company based in California is seeking an experienced AI Chip Toolchain Architect to design and plan innovative AI toolchains. The ideal candidate should have a master’s degree in computer science and more than 5 years of experience in model deployment and compression, with strong C++ programming skills. Responsibilities include overseeing architecture design, conducting feasibility assessments, and ensuring technological competitiveness in AI solutions. This role requires excellent communication skills and the ability to work collaboratively across teams.

Qualifications

  • Over 5 years of experience in model deployment and model compression, or over 10 years in AI algorithm development.
  • Deep understanding of AI technologies and trends.
  • Strong programming skills with experience in C++ system projects.

Responsibilities

  • Responsible for overall architecture design and planning of AI toolchain.
  • Conduct feasibility assessments of core technologies.
  • Focus on long-term competitiveness in AI toolchain.

Skills

Model deployment
Model compression
C++ programming
AI architecture development
Communication and collaboration

Education

Master's degree in computer science or related major

Tools

TensorRT

Job description

About the job AI Chip Toolchain Architect

Job Responsibilities

1. Be responsible for the overall architecture design and planning of Horizon's AI toolchain. Design a system architecture and overall solution that meet the requirements, and track and support the implementation of requirements during the product R & D process. Conduct feasibility assessments of key core technologies of the project and assist in improving the product definition.

2. Focus on the long-term technical competitiveness of the AI toolchain. Think from the perspectives of model deployment and model compression

3. Be responsible for the R & D of the model quantization and compression tool. Conduct mid - and long - term planning for AI model deployment, model compression, and model quantization technologies to ensure the technical competitiveness of the AI chip toolchain in the fields of model quantization and model compression.

4. Undertake the system and architecture design of the model quantization tool. Analyze and decompose the system problems during the deployment of AI models for autonomous driving.

Job Requirements

1. A master's degree or above in computer science or a related major. More than 5 years of work experience in model deployment and model compression, or more than 10 years of experience in AI algorithm development, architecture design, or technical management. Have an in - depth understanding of the latest AI technologies and trends.

2. Be familiar with the end - to - end details of AI model deployment, including but not limited to model quantization, compilation, and edge - side deployment optimization. Have a deep understanding of key technologies such as model compression (especially post - quantization), model deployment, etc., and be able to conduct mid - and long - term technical planning proficiently. Have an accurate prediction of the development of the model deployment field and have a relatively in - depth understanding and recognition of at least one mainstream deployment optimization tool, such as TensorRT.

3. Understand the business problems and pain points in the development process of algorithms for intelligent driving and human - machine interaction, as well as the development models. Be able to transform domain technologies and models (such as model conversion and optimization technologies, compiler technologies) into engineering architectures. Have an in - depth understanding of the future evolution of algorithms and application development models for autonomous driving and human - machine interaction, and have an in - depth understanding of the development models for algorithms and applications.

4. Be able to evaluate multiple alternative solutions, make architecture decisions, determine priorities, and guide the project and the organization in the right direction. Have strong abstraction ability to simplify complex problems and transform high - level architecture technical planning into detailed design.

5. Have strong programming skills. Be proficient in the development, upgrading, and maintenance of complex C++ system projects and have in - depth thinking at the system architecture level.

6. Have strong communication and collaboration abilities and documentation skills. Collaborate with other architects and stakeholders, align goals, document the architecture design and decisions, and communicate them to the team to unify cognition. Be able to clearly express and convey your design to the team and guide developers to implement it correctly. Preferably with experience in complex software system development.

7. Preferably with experience in AI compilers, PTQ/QAT, GPT large models, algorithms for autonomous driving and human - machine interaction, and AI architecture development.

8. Preferably with published papers on model compression and deployment in core conference journals or experience in the development of mainstream AI chip toolchains.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Chip Toolchain Architect—Model Quantization & Deployment
AI Chip Toolchain Architect—Model Quantization & Deployment

Forestown • California City (CA)

On-site
USD 120,000 - 150,000
Sr. Staff AI Engineer, Silicon Design
Sr. Staff AI Engineer, Silicon Design

Cognichip • Redwood City (CA)

On-site
USD 150,000 - 200,000
Staff AI/ML Software Engineer, Model Distillation, Fine-Tuning
Staff AI/ML Software Engineer, Model Distillation, Fine-Tuning

Jobtailor • California (MO)

Hybrid
USD 180,000 - 260,000
AI Compiler R&D Expert
AI Compiler R&D Expert

Forestown • California City (CA)

On-site
USD 100,000 - 130,000
Chip Design Engineer
Chip Design Engineer

TenX Semi • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
AI Engineer
AI Engineer

TenX Semi • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Gen AI architect
Gen AI architect

Prodapt • San Francisco (CA)

On-site
USD 150,000 - 190,000
AI Architect
AI Architect

Lorven Technologies Inc. • Foster City (CA)

On-site
USD 180,000 - 260,000
Senior AI Performance Architect
Senior AI Performance Architect

Qualcomm • Raleigh (NC)

On-site
USD 126,000 - 218,000
Competitive annual discretionary bonus
Annual RSU grants
Comprehensive benefits package
Software Engineer, AI Framework
Software Engineer, AI Framework

Black Sesame Technologies Inc • San Jose (CA)

On-site
USD 140,000 - 210,000