Edge AI Architect: LLM & Inference Optimization

NextGen Federal Systems

Aberdeen (WA)

On-site

USD 130,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NextGen Federal Systems seeks an Edge AI/Model Optimization Engineer to deploy, optimize, and sustain AI capabilities in edge and tactical computing environments. You will evaluate LLMs and embedding models for latency, memory usage, and reliability on constrained hardware such as the X9 Spider Mission Computer.

You will tune runtime configurations, collaborate with stakeholders to assess mission requirements, and build robust benchmarks and validation procedures to ensure dependable AI

Qualifications

  • Bachelor's degree in a technical field (CS/EE/CE/AI) or related discipline.
  • 5+ years in AI/ML deployment, model optimization, edge computing, or AI inference operations.
  • Experience deploying and optimizing LLMs or embeddings on constrained hardware.
  • Proficiency with GPU-enabled systems and inference optimization tools (CUDA, TensorRT, ONNX Runtime).
  • Experience tuning runtime configurations (quantization, batching, memory).
  • Familiarity with Linux, containerized deployments (Docker/Kubernetes).
  • Strong analytical, troubleshooting, and communication skills.
  • Active Security Clearance required.

Responsibilities

  • Evaluate LLMs and embedding models for quality, latency, memory, and reliability on edge-enabled platforms (X9 Spider).
  • Tune runtime configurations for edge deployment including quantization, batching, and memory optimization.
  • Collaborate with stakeholders to assess mission requirements and platform tradeoffs.
  • Benchmark AI workflows and model-serving architectures against hardware constraints.
  • Develop repeatable performance and stress-testing frameworks for edge environments.
  • Package, deploy, and sustain local model-serving components for edge operations.
  • Support integration teams to ensure mission effectiveness after optimizations.

Skills

AI model optimization
Edge computing
GPU acceleration
Linux
Python
Communication
Security clearance

Education

Bachelor's degree in Computer Science / Electrical Engineering or related

Tools

CUDA
TensorRT
ONNX Runtime
vLLM
Ollama

Job description

NextGen Federal Systems seeks an Edge AI/Model Optimization Engineer to deploy, optimize, and sustain AI capabilities in edge and tactical computing environments. You will evaluate LLMs and embedding models for latency, memory usage, and reliability on constrained hardware such as the X9 Spider Mission Computer.

You will tune runtime configurations, collaborate with stakeholders to assess mission requirements, and build robust benchmarks and validation procedures to ensure dependable AI

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Edge AI Optimization Engineer for Tactical Systems
Edge AI Optimization Engineer for Tactical Systems

Nextgenfed • Aberdeen (MD)

On-site
USD 140,000 - 210,000
Edge AI Engineer: LLMs & Model Optimization
Edge AI Engineer: LLMs & Model Optimization

NextGen Federal Systems • Aberdeen (MD)

On-site
USD 90,000 - 120,000
Equal Opportunity Employer
Edge AI/Model Optimization Engineer
Edge AI/Model Optimization Engineer

Nextgenfed • Aberdeen (MD)

On-site
USD 140,000 - 210,000
Edge AI/Model Optimization Engineer
Edge AI/Model Optimization Engineer

NextGen Federal Systems • Aberdeen (MD)

On-site
USD 90,000 - 120,000
Equal Opportunity Employer
Edge AI/Model Optimization Engineer
Edge AI/Model Optimization Engineer

NextGen Federal Systems • Aberdeen (WA)

On-site
USD 130,000 - 190,000
Edge AI/ML Engineer for Production LLMs & Multimodal AI
Edge AI/ML Engineer for Production LLMs & Multimodal AI

Expression Networks • Washington

Hybrid
USD 140,000 - 190,000
401k matching
Medical/dental/vision insurance
Education reimbursement up to $10,000/
+5
Remote Edge AI Engineer - On-Device ML & Optimization
Remote Edge AI Engineer - On-Device ML & Optimization

Visa Hunt • United States

On-site
USD 100,000 - 155,000
Edge AI Systems Engineer — Real-Time ML Ops
Edge AI Systems Engineer — Real-Time ML Ops

Evolved • City of Elmira (NY)

On-site
USD 120,000 - 160,000
Senior Edge AI Inference Optimization Engineer
Senior Edge AI Inference Optimization Engineer

PVH (Tommy Hilfiger/Calvin Klein) • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Intel Benefits
Lead Edge AI/ML Engineer for Real-Time Embedded Systems
Lead Edge AI/ML Engineer for Real-Time Embedded Systems

Arcfield • Home Creek (VA)

On-site
USD 101,000 - 201,000