PRX- AI Engineer

Apexon Technology

New York (NY)

On-site

USD 120,000 - 150,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health Insurance with Dental & Vision
401K Plan
Life Insurance, STD & LTD
Paid Vacations & Holidays

Job summary

Apexon Technology is looking for a skilled engineer in New York to build agentic AI systems and productionize Large Language Models (LLMs). You'll design and implement advanced AI solutions, ensuring safety and compliance, while optimizing performance and costs.

Your role will involve collaborating with diverse teams to address real-world challenges through innovative AI applications. Benefits include health insurance, 401K, life insurance, and paid time off.

Qualifications

  • 5+ years of software development in languages like Python, C/C++, Go, Java.
  • Practical experience with LLMs and cloud infrastructure, preferably AWS.
  • Strong analytical problem-solving and effective communication skills.

Responsibilities

  • Design and implement AI systems using retrieval and reasoning.
  • Productionize LLMs by building evaluation frameworks.
  • Collaborate with teams to translate production pain points into actionable plans.

Skills

Software development
Large Language Models (LLMs)
Cloud infrastructure
Analytical problem-solving

Job description

Responsibilities
  • Build agentic AI systems: design and implement tool‑calling agents that combine retrieval, structured reasoning, and secure action execution (function calling, change orchestration, policy enforcement) following MCP protocol. Engineer robust guardrails for safety, compliance, and least‑privilege access.
  • Productionize LLMs: build an evaluation framework for open‑source and foundational LLMs; implement retrieval pipelines, prompt synthesis, response validation, and self‑correction loops tailored to production operations.
  • Integrate with runtime ecosystems: connect agents to observability, incident management, and deployment systems to enable automated diagnostics, runbook execution, remediation, and post‑incident summarization with full traceability.
  • Collaborate directly with users: partner with production engineers and application teams to translate production pain points into agentic AI roadmaps; define objective functions linked to reliability, risk reduction, and cost; and deliver auditable, business‑aligned outcomes.
  • Safety, reliability, and governance: build validator models, adversarial prompts, and policy checks into the stack; enforce deterministic fallbacks, circuit breakers, and rollback strategies; instrument continuous evaluations for usefulness, correctness, and risk.
  • Scale and performance: optimize cost and latency via prompt engineering, context management, caching, model routing, and distillation; leverage batching, streaming, and parallel tool‑calls to meet stringent SLOs under real‑world load.
  • Build a RAG pipeline: curate domain knowledge; build data‑quality validation framework; establish feedback loops and milestone framework to maintain knowledge freshness.
  • Raise the bar: drive design reviews, experiment rigor, and high‑quality engineering practices; mentor peers on agent architectures, evaluation methodologies, and safe deployment patterns.
Qualifications
  • 5+ years of software development in one or more languages (Python, C/C++, Go, Java); strong hands‑on experience building and maintaining large‑scale Python applications preferred.
  • 3+ years designing, architecting, testing, and launching production ML systems, including model deployment/serving, evaluation and monitoring, data processing pipelines, and model fine‑tuning workflows.
  • Practical experience with Large Language Models (LLMs): API integration, prompt engineering, fine‑tuning/adaptation, and building applications using RAG and tool‑using agents (vector retrieval, function calling, secure tool execution).
  • Understanding of different LLMs, both commercial and open source, and their capabilities (OpenAI, Gemini, Llama, Qwen, Claude).
  • Solid grasp of applied statistics, core ML concepts, algorithms, and data structures to deliver efficient and reliable solutions.
  • Strong analytical problem‑solving, ownership, and urgency; ability to communicate complex ideas simply and collaborate effectively across global teams with a focus on measurable business impact.
  • Preferred: Proficiency building and operating on cloud infrastructure (ideally AWS), including containerized services (ECS/EKS), serverless (Lambda), data services (S3, DynamoDB, Redshift), orchestration (Step Functions), model serving (SageMaker), and infra‑as‑code (Terraform/CloudFormation).
Commitment to Diversity & Inclusion

Apexon is committed to being an equal opportunity employer and promoting diversity in the workplace. We take affirmative action to ensure equal employment opportunity for all qualified individuals. Apexon strictly prohibits discrimination and harassment of any kind and provides equal employment opportunities to employees and applicants without regard to gender, race, color, ethnicity or national origin, age, disability, religion, sexual orientation, gender identity or expression, veteran status, or any other applicable characteristics protected by law.

Commitment to Environment

Actively contribute to Apexon's commitment to environmental responsibility by following sustainable practices and supporting ESG initiatives.

Benefits
  • Health Insurance with Dental & Vision
  • 401K Plan
  • Life Insurance, STD & LTD
  • Paid Vacations & Holidays
  • FSA Dependent & Limited Purpose Care
Job Location

New York, United States

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Engineer
Senior AI Engineer

Beacon Roofing Supply, Inc • Seattle (WA)

On-site
USD 150,000 - 220,000
Annual performance bonus
401(k) with employer match
Medical, dental, and vision insurance
+1
Senior AI Engineer
Senior AI Engineer

Apperture Solutions • Charlotte (NC)

On-site
USD 140,000 - 190,000
Employee Stock Ownership Plan (ESOP)
Retirement plan
Medical/Dental/Vision Insurance
+6
AI Engineer, Product Software
AI Engineer, Product Software

Equinix, Inc. • Redwood City (CA)

On-site
USD 142,000 - 212,000
Agentic AI Engineer
Agentic AI Engineer

Trilagen • Bethesda (MD)

On-site
USD 120,000 - 160,000
401K
Health Insurance
Dental Insurance
+2
Senior AI Engineer
Senior AI Engineer

CCC Intelligent Solutions Inc. • Chicago (IL)

On-site
USD 120,000 - 160,000
401K Match
Paid time off
Performance Bonus
+2
AI Infrastructure Lead
AI Infrastructure Lead

Apex Group Ltd • Boston (MA)

On-site
USD 225,000 - 275,000
Dynamic and fast-paced team
Exposure to all business aspects
Training and development opportunities
+1
AI Engineer, Product Software
AI Engineer, Product Software

Equinix • Redwood City (CA)

On-site
USD 142,000 - 212,000
Employee Assistance Program
US Benefits: health, life, disability,
retirement
Staff Engineer, AI/LLM Platform
Staff Engineer, AI/LLM Platform

Simulations Plus • Northern (KY)

On-site
USD 100,000 - 120,000
Fully remote work
Flexible schedules
Generous vacation policy
+2
Principal AI Engineer
Principal AI Engineer

NAM Info Inc • New York (NY)

Hybrid
USD 210,000 - 320,000
Hybrid work model
Enterprise - AI Engineer - Python, LLM, AWS
Enterprise - AI Engineer - Python, LLM, AWS

Erias Ventures • Maryland

Hybrid
USD 150,000 - 265,000
Above Market Hourly Pay
401k with immediate vesting
Spot bonuses for business development
+5