Full Stack AI Solution Engineer

Lenovo

Kampung Malaysia Tambahan

On-site

MYR 180,000 - 300,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Lenovo is seeking an experienced AI/ML architecture professional to lead enterprise-scale LLM, RAG, and Agent initiatives. You will design hybrid retrieval systems, orchestrate multi-tool AI agents, and implement standardized prompt engineering with NL2SQL capabilities.

The role emphasizes observability, cost optimization, and strong cross-team collaboration. You will contribute to scalable AI platforms, ensuring high availability and secure, compliant deployment across enterprise environments.

Qualifications

  • Solid practical experience in end-to-end enterprise-level LLM, RAG and Agent projects.
  • In-depth understanding of systematic prompt engineering and NL2SQL.
  • Familiar with vector database engineering and hybrid retrieval optimization.
  • Proficient in Databricks platform application and MLOps system construction.
  • Experience in AI cost optimization, security compliance architecture, and technical debt governance.
  • Strong business thinking and cross-team communication capabilities.

Responsibilities

  • Research and implement enterprise-level RAG systems, build hybrid retrieval architecture, and fine-tune reranking models.
  • Design and build complex AI agent orchestration platforms with multi-tool chaining and dynamic prompts routing.
  • Develop standardized prompt engineering systems, NL2SQL solutions, and multi-turn dialogue intent clarification mechanisms.
  • Build end-to-end observability for LLM apps, including prompt tracking and hallucination detection.
  • Create reusable AI components and AB testing systems for RAG strategies, prompts, and agent logic.
  • Optimize AI application costs via caching, vector storage tiering, and idle-powers scheduling.
  • Govern AI technical debt with unified gateway standards and scalable deployment practices.
  • Advance from POC to enterprise-scale architectures with high availability and scalable inference.

Skills

LLM systems
RAG integration
Agent orchestration
Prompt engineering
Observability
Cost optimization
Security compliance
Cross-team comms

Tools

Databricks
Delta Lake
Vector DB
MLOps
NL2SQL

Job description

We are Lenovo. We do what we say. We own what we do. We WOW our customers.

Lenovo is a US$83 billion revenue global technology powerhouse, ranked #153 in the Fortune Global 500, and serving millions of customers every day in 180 markets. Focused on a bold vision to deliver Smarter Technology for All, Lenovo has built on its success as the world’s largest PC company with a full-stack portfolio of AI-enabled, AI-ready, and AI-optimized devices (PCs, workstations, smartphones, tablets), infrastructure (server, storage, edge, high performance computing and software defined infrastructure), software, solutions, and services. Lenovo’s continued investment in world-changing innovation is building a more equitable, trustworthy, and smarter future for everyone, everywhere. Lenovo is listed on the Hong Kong stock exchange under Lenovo Group Limited (HKSE: 992) (ADR: LNVGY).

This transformation together with Lenovo’s world-changing innovation is building a more inclusive, trustworthy, and smarter future for everyone, everywhere. To find out more visit www.lenovo.com, and read about the latest news via our StoryHub.

Key Responsibilities
  • Enterprise RAG System Engineering Tuning Undertake the research and engineering implementation of enterprise-level RAG systems, build hybrid retrieval architecture integrating keyword, vector and graph retrieval, realize adaptive document chunking and reranking model fine-tuning, land full-link hallucination suppression strategies, and continuously improve knowledge retrieval accuracy and LLM answer robustness.
  • Complex Business Agent Orchestration Design and build complex business Agent orchestration platforms, implement multi-tool chained calling, hierarchical memory management, dynamic prompt routing and automatic Agent failure retry mechanisms, and standardize human-machine collaborative workflow for complex industrial and business scenarios.
  • Systematic Prompt Engineering Dialogue Optimization Build standardized prompt engineering systems, complete NL2SQL solution design and tuning, optimize conversational data analysis interaction logic and user intent recognition capabilities, establish prompt version management, quantitative effect evaluation and multi-turn dialogue intent clarification mechanisms, and improve the practicability and stability of LLM dialogue services.
  • LLM Full-link Observability Construction Build end-to-end observability systems for LLM applications, realize prompt parameter tracking, retrieval log collection and monitoring, model output hallucination detection, and automatic user QA effect scoring, support closed-loop quantitative iteration and continuous optimization of AI applications.
  • Reusable AI Component A/B Testing System Construction Encapsulate reusable AI application basic components including general knowledge base access modules, Agent tool plugin market and unified multi-model invocation gateway. Build A/B testing systems to support gray-scale comparison and iterative optimization of RAG retrieval strategies, prompt templates and Agent business logic.
  • AI Cost Optimization Security Compliance Architecture Optimize operating costs of large-scale AI applications via LLM caching, quantitative inference, vector storage tiered management and idle computing power recycling scheduling. Build AI security and compliance architecture covering prompt injection prevention, knowledge base data desensitization, model access authentication and user dialogue data standardized compliance management.
  • Unified AI Gateway Technical Debt Governance Design cross-terminal unified AI service gateway, unify logging, authentication and rate limiting standards for RAG, Agent and LLM fine-tuning businesses. Carry out AI technical debt governance, complete standardized transformation of scattered stock RAG/Agent applications, and unify AI project development, deployment and operation specifications.
  • AI Architecture Upgrade from POC to Enterprise Scale Promote the upgrade of POC-level AI solutions to enterprise-scale architectures, implement vector database sharding and incremental knowledge synchronization, optimize massive document retrieval performance. Build high-availability LLM inference architecture with full-link fault tolerance capabilities including circuit breaking, rate limiting, caching and dynamic scaling.
  • Business Value Transformation Standardized Delivery Sort out structured workflows for complex industrial Agents, convert technical architectures and indicators into quantifiable business value and high-level reporting metrics. Conduct cross-team AI architecture interpretation for product and business teams, and precipitate standardized AI POC delivery templates including cost estimation, performance baselines, risk lists and large-scale transformation plans.
Job Requirements
  • Solid practical experience in end-to-end enterprise-level LLM, RAG and Agent project landing, capable of independent architecture design, performance tuning and online iteration of large-scale AI applications.
  • In-depth understanding of systematic prompt engineering and NL2SQL technology, proficient in multi-turn dialogue context management and intent clarification, with practical experience in prompt version management and effect quantitative evaluation.
  • Familiar with vector database engineering, hybrid retrieval optimization and reranking tuning, able to solve core problems such as insufficient retrieval accuracy, model hallucination and massive data retrieval performance bottlenecks.
  • Proficient in Databricks platform application and MLOps system construction, familiar with data lake + vector database fusion architecture and Delta Lake data governance tuning.
  • Have in-depth practice in AI application cost optimization, security compliance architecture construction and technical debt governance, with enterprise-level high availability, standardization and compliance awareness.
  • Possess excellent business thinking and cross-team communication capabilities, able to translate technical advantages into business value and promote standardized and large-scale landing of AI projects.

We are an Equal Opportunity Employer and do not discriminate against any employee or applicant for employment because of race, color, sex, age, religion, sexual orientation, gender identity, national origin, status as a veteran, and basis of disability or any federal, state, or local protected class.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Full Stack Engineer
AI Full Stack Engineer

Lenovo • Kuala Lumpur

On-site
MYR 180,000 - 300,000
AI Tech Lead / Solution Architect- Finance Domain
AI Tech Lead / Solution Architect- Finance Domain

Lenovo • Kuala Lumpur

On-site
MYR 180,000 - 300,000
Senior AI Solution Architect for Enterprise LLM & RAG
Senior AI Solution Architect for Enterprise LLM & RAG

Lenovo • Kampung Malaysia Tambahan

On-site
MYR 180,000 - 300,000
Full Stack AI Engineer
Full Stack AI Engineer

RSGx • Kuala Lumpur

Hybrid
MYR 180,000 - 240,000
Hybrid work
KL office nearby
AIOps Assistant Engineer
AIOps Assistant Engineer

Lenovo • Petaling Jaya

On-site
MYR 70,000 - 110,000
Full Stack AI Engineer
Full Stack AI Engineer

Resource Services Group X Pty Ltd • Kuala Lumpur

Hybrid
MYR 180,000 - 320,000
Hybrid working arrangements in Kuala L
Senior AI Engineer
Senior AI Engineer

FinTop Consulting • Malaysia

On-site
MYR 180,000 - 260,000
AI Engineer
AI Engineer

Fourtitude Asia • Petaling Jaya

On-site
MYR 42,000 - 68,000
AI Solutions Engineer
AI Solutions Engineer

techstreet • Petaling Jaya

On-site
MYR 60,000 - 110,000
AI Engineer (GenAI / RAG / LangGraph / Python)
AI Engineer (GenAI / RAG / LangGraph / Python)

Unison Group • Kuala Lumpur

On-site
MYR 180,000 - 300,000