AI Engineer

Activeloop

Mountain View (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Activeloop in Mountain View is seeking an AI Engineer who will design, develop, and deploy advanced AI search and retrieval systems that leverage RAG techniques to solve complex information access challenges.

You will collaborate with software engineers, customers, and business stakeholders to build scalable, low-latency search solutions, optimize vector databases, and evaluate performance with rigorous metrics.

Qualifications

  • Master's or PhD in CS/ML/Statistics or related field.
  • Proven production experience with ML models and cloud platforms.
  • Strong knowledge of information retrieval, RAG, and search systems.
  • Experience with distributed training, vector databases, and scaling.
  • Proficiency in Python and at least one of C++ and modern ML libraries.

Responsibilities

  • Design, develop, and deploy AI search and retrieval systems using RAG.
  • Collaborate with engineers, customers, and stakeholders to deliver AI search solutions.
  • Lead design and implementation of retrieval systems like Deep Memory by Activeloop.
  • Develop and optimize search algorithms, including semantic and hybrid search.
  • Integrate and optimize vector storage and indexing for high-dim embeddings.
  • Design query processing pipelines and evaluate system performance with metrics.
  • Ensure scalability, low latency, and real-time capabilities for large datasets.

Skills

Python
Machine Learning
Information Retrieval
Distributed Training
Cloud Platforms

Education

Master's or PhD in Computer Science / ML / Statistics

Tools

TensorFlow
PyTorch
Llama Index
LangChain
Deep Lake

Job description

- We're looking for an AI Engineer who possesses a deep understanding of large-scale information retrieval systems, deep learning, databases, and RAG architectures. The ideal candidate will have expertise in developing and optimizing search algorithms, implementing efficient indexing techniques, and leveraging RAG to enhance AI-powered search and question-answering systems.

What You Will Be Doing

As an AI Engineer, you will play a pivotal role in designing, developing, and deploying advanced search and retrieval systems that leverage RAG techniques to solve complex information access challenges. You will collaborate with software engineers, customers, and business stakeholders to develop AI search solutions that deliver significant value to the organization and our clients.

Key Responsibilities

RAG System Research and Implementation: Lead the design and implementation of advanced retrieval systems like Deep Memory by Activeloop, delivering optimized RAG systems across the entire value chain - from embedding or model fine-tuning to retrieval optimization with custom algorithms, to enhance knowledge retrieval accuracy.

Search Algorithm Optimization: Develop and refine search algorithms, including semantic search, hybrid search, and multi-modal search techniques, to improve retrieval performance and relevance ranking.

Vector Database Integration : Implement and optimize vector storage and indexing solutions within Deep Lake, ensuring efficient similarity search capabilities for high-dimensional embeddings used in RAG systems.

Query Understanding and Processing: Design and implement advanced query processing pipelines, including query expansion, intent recognition, and contextual interpretation to enhance search precision.

Information Retrieval Model Development: Create and fine-tune machine learning models specifically for information retrieval tasks, such as document ranking, query-document relevance scoring, and zero-shot retrieval.

Performance Evaluation and Metrics: Establish comprehensive evaluation frameworks for search and RAG systems, including relevance assessments, A/B testing, and user satisfaction metrics to continually improve system performance.

Scalability and Efficiency: Optimize RAG and search systems for high throughput and low latency, ensuring they can handle large-scale datasets and real-time query processing demands.

Data Ingestion and Indexing: Develop efficient data ingestion pipelines and indexing strategies to support rapid updates and real-time search capabilities across diverse data types and sources.

What We Need to See
  • Master's or PhD degree in Computer Science, Machine Learning, Statistics, or a related field.
  • Strong programming skills in one or more programming languages, such as Python, or C++, and extensive experience with machine learning libraries, such as TensorFlow, PyTorch, Llama Index, LangChain etc.
  • Proven experience in developing and deploying complex machine learning models in production environments, including experience with cloud-based platforms, edge devices, or embedded systems.
  • Strong understanding of advanced machine learning algorithms, such as deep learning, reinforcement learning, RAGs, and ensemble methods, and experience with model optimization techniques, such as hyper-parameter tuning, model compression, and quantization.
  • Solid understanding of data pre-processing, feature engineering, and data quality assurance techniques, and experience with large-scale, complex data sets.
  • Excellent problem-solving skills and ability to analyze and interpret complex data sets to extract meaningful insights and drive decision-making.
Ways To Stand Out From The Crowd
  • You have trained deep learning models in a distributed manner.
  • You are a highly motivated, curious, hardworking explorer in the field of AI
  • You have publications in top-tier machine learning and AI conferences such as ICML, NeurIPS, and CVPR (highly desirable).
  • You have a builder attitude and you love building cool things that matter
  • You're excited to work closely with the founding team in developing hyper-scalable software for ML
  • You can proactively identify and anticipate problems and provide tangible solutions
  • You enjoy the startup journey of building an endurable, scalable business
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Manager of Machine Learning Engineering
Senior Manager of Machine Learning Engineering

ServiceNow • Santa Clara (CA)

On-site
USD 230,000 - 320,000
Generous family leave
Matched donations
Annual learning stipends
+3
AI Search Engineer
AI Search Engineer

Compunnel, Inc. • Plano (TX)

On-site
USD 100,000 - 135,000
AI/ML Engineer
AI/ML Engineer

RiskForce • Northern (KY)

Hybrid
USD 120,000 - 155,000
AI Engineer
AI Engineer

LeoTechnologies • Boca Raton (FL), Northern (KY)

On-site
USD 140,000 - 190,000
RAG-Powered AI Search Engineer
RAG-Powered AI Search Engineer

Activeloop • Mountain View (CA)

On-site
USD 150,000 - 210,000
Senior AI Engineer
Senior AI Engineer

Harnham • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior Machine Learning Engineer/ RAG
Senior Machine Learning Engineer/ RAG

Motion Recruitment • Dallas (TX)

On-site
USD 150,000 - 190,000
AI/ML Engineer
AI/ML Engineer

Winaxis LLC • Dallas (TX)

On-site
USD 120,000 - 160,000
Lead Al Engineer
Lead Al Engineer

Namely • Mountain View (CA)

On-site
USD 19,000 - 35,000
Senior AI Software Engineer
Senior AI Software Engineer

Harnham • San Francisco (CA)

On-site
USD 180,000 - 260,000