Production GenAI Engineer — LLM Systems & RAG Pipelines

onebyzero

Makati

On-site

PHP 1,200,000 - 1,800,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive compensation
AWS partnership access
Global team
Diversity and inclusion

Job summary

OneByZero is seeking an Applied AI Engineer with 3–4 years of ML engineering experience to design and build production‑grade GenAI systems across Asia Pacific. You will help shape end‑to‑end LLM architectures, RAG pipelines, and multi‑agent setups, with a strong emphasis on production readiness and reliable software fundamentals.

You will work on real‑world enterprise use cases, pre‑train and fine‑tune models, evaluate performance and cost, and collaborate with cross‑functional teams to deliver

Qualifications

  • 3–6 years of experience in ML engineering, LLMs, or model development roles.
  • Hands‑on experience with CPT, SFT, LoRA, or QLoRA on LLMs.
  • Strong Python programming skills.
  • Ability to write clean, testable, production‑ready code.
  • Experience running model evaluation and benchmarking pipelines.
  • Solid understanding of transformer architectures and fine‑tuning effects.
  • Experience deploying fine‑tuned and pre‑trained models to cloud environments with cost and latency awareness.
  • Strong problem‑solving and independent client‑facing work capability.
  • Solid software fundamentals: APIs, data structures, testing, debugging, and performance optimization.
  • Ability to review designs, communicate trade‑offs, and collaborate in a fast‑paced environment.
  • Experience with AWS‑native GenAI building blocks and secure enterprise deployments.
  • Experience with vector databases/search engines and retrieval optimization.
  • Experience with containerization and orchestration (Docker, Kubernetes).
  • Experience building evaluation/observability pipelines for LLM systems and implementing safety/guardrail patterns.
  • Consulting or client‑facing delivery experience.

Responsibilities

  • Design and contribute to end‑to‑end LLM system architecture for real‑world enterprise use cases.
  • Pre‑train, fine‑tune LLMs and domain models for client needs.
  • Design and run model evaluation pipelines to benchmark performance and cost.
  • Optimize models for latency, throughput, and token efficiency in production.
  • Integrate fine‑tuned models into multi‑agent pipelines with orchestration teams.
  • Implement prompt versioning, rollback strategies, and monitoring for reliability.
  • Translate client requirements into model adaptation strategies with clear success criteria.
  • Contribute to internal knowledge sharing on fine‑tuning practices and tooling.
  • Define enterprise integration patterns for GenAI systems including governance and data boundaries.
  • Improve production reliability and cost control; collaborate across cross‑functional teams.
  • Contribute hands‑on code, reviews, and raise engineering standards.

Skills

Python
ML engineering
LLMs
Production code

Tools

Docker
Kubernetes
Bedrock
OpenSearch
Lambda
ECS/EKS
Pinecone
Weaviate
Milvus
FAISS

Job description

OneByZero is seeking an Applied AI Engineer with 3–4 years of ML engineering experience to design and build production‑grade GenAI systems across Asia Pacific. You will help shape end‑to‑end LLM architectures, RAG pipelines, and multi‑agent setups, with a strong emphasis on production readiness and reliable software fundamentals.

You will work on real‑world enterprise use cases, pre‑train and fine‑tune models, evaluate performance and cost, and collaborate with cross‑functional teams to deliver

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Production GenAI Engineer for Enterprise LLMs
Production GenAI Engineer for Enterprise LLMs

Onebyzero • Philippines

Hybrid
PHP 1,200,000 - 1,800,000
GenAI Engineer: Scalable LLM Solutions & Pipelines
GenAI Engineer: Scalable LLM Solutions & Pipelines

TymblHub • Hinoba-an

On-site
PHP 900,000 - 1,300,000
GenAI/AI Software Engineer (Junior) — RAG & LLM Pipelines
GenAI/AI Software Engineer (Junior) — RAG & LLM Pipelines

Marktine • Hinoba-an

On-site
PHP 400,000 - 700,000
Docker experience
Linux environment
Cloud deployment (AWS/GCP)
Junior GenAI Engineer — LLMs, ML Apps & Deployment
Junior GenAI Engineer — LLMs, ML Apps & Deployment

TymblHub • Hinoba-an

On-site
PHP 600,000 - 1,200,000
Senior GenAI Engineer: Build Scalable LLM Solutions
Senior GenAI Engineer: Build Scalable LLM Solutions

TymblHub • Hinoba-an

On-site
PHP 1,674,000 - 3,571,000
AI Engineer: Build Production ML & LLM Solutions
AI Engineer: Build Production ML & LLM Solutions

Cctech • Hinoba-an

On-site
PHP 900,000 - 1,500,000
Cutting-edge AI projects
Growth opportunities
Senior AI/ML Engineer - GenAI & Production Pipelines
Senior AI/ML Engineer - GenAI & Production Pipelines

TymblHub • Hinoba-an

On-site
PHP 1,200,000 - 2,000,000
Production AI/ML Engineer - GenAI & MLOps
Production AI/ML Engineer - GenAI & MLOps

GECO Asia Pte. Ltd • Santo Niño 1st

On-site
PHP 1,000,000 - 2,000,000
Senior AI Engineer — Remote, LLM/RAG Production Systems
Senior AI Engineer — Remote, LLM/RAG Production Systems

D2B • Metro Manila

Remote
PHP 1,339,000 - 2,009,000
100% remote role
Paid local holidays aligned with AU/NZ
Training and professional growth
+1
AI Engineer - Generative AI (Remote)
AI Engineer - Generative AI (Remote)

Hire Feed • Mexico

On-site
PHP 7,524,000 - 11,285,000