Principal AI Engineer

IMR Soft Llc

New York (NY)

On-site

USD 180,000 - 280,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

IMR Soft Llc in the New York area is seeking a Principal AI Engineer to lead enterprise-scale agentic AI implementations for Fortune 500 clients. This hands-on role focuses on designing and shipping autonomous, tool-calling AI systems that reason over enterprise context and operate at scale with strict latency, cost, and governance constraints.

You will own the full stack—from data pipelines and backend services to orchestration and cloud infrastructure—while mentoring teams and shaping

Qualifications

  • 8–14 years of software engineering with Python
  • Experience with systems languages (Go, Rust, Java, C/C++)
  • Strong data structures and algorithms
  • APIs, microservices, and distributed systems
  • Data pipelines and production-grade systems

Responsibilities

  • Design and build agentic systems with tool-calling agents
  • Productionize LLM applications with retrieval pipelines and evaluation
  • Own the full stack from data pipelines to orchestration
  • Engineer for reliability, governance, and deterministic fallbacks
  • Optimize for cost and latency with caching and routing

Skills

Python
System design
APIs
Data pipelines
LLM engineering
Distributed systems
Token optimization

Tools

LangGraph
Google ADK
CrewAI
Claude Agent SDK
Milvus
Pinecone
Weaviate
FAISS
Terraform
CloudFormation

Job description

Principal AI Engineer

Location: US(NYC)- 3(WFO)

Employment Type: Full Time (Overlapping EST)

About the Role

Hiring a Staff/Principal AI Engineer to lead enterprise-scale agentic AI implementations for Fortune 500 clients.

This is a hands-on engineering role focused on designing and shipping autonomous, tool-calling AI systems - agents that reason over enterprise context, invoke real systems through secure interfaces, and operate reliably at scale under strict latency, cost, and governance constraints.

You will own these systems end to end: the data pipelines feeding them, the backend services around them, the agent orchestration layer, the evaluation harness that keeps them honest, and the cloud infrastructure they run on. We are looking for engineers with genuine software engineering and data science depth who have taken agentic systems all the way to production.

What We're Looking For
Engineering foundation
  • 8 14 years of software engineering experience, with strong hands-on large-scale Python
  • Working depth in at least one systems or backend language - Go, Rust, Java, or C/C++ - and the judgment to know when to reach for it
  • Strong data structures and algorithms.
  • Strong understanding of APIs, microservices, and system design
  • Hands-on experience building and operating data pipelines and production-grade distributed systems.
Agentic AI and LLMs
  • 2+ years of hands-on LLM engineering, with at least couple agentic system you designed and took to production
  • Production experience with agent frameworks - LangGraph, Google ADK, CrewAI, Claude Agent SDK, or equivalent - and the fluency to move between them as the ecosystem evolves
  • Experience building MCP (Model Context Protocol) servers and tool-calling interfaces
  • RAG from first principles: chunking strategy, embeddings, vector and hybrid retrieval, reranking, and response validation
  • Strong experience with vector databases (Milvus, Pinecone, Weaviate, FAISS, etc. or cloud equivalents)
  • Design of guardrails and reliability patterns - validators, policy checks, self-correction loops, deterministic fallbacks, circuit breakers, and rollback paths
Optimization
  • Deep familiarity with token optimization and context-window management - context shaping, pruning, and compaction
  • Latency and cost optimization through caching, model routing, batching, streaming, and parallel tool calls
  • Performance testing and tuning systems against defined SLOs
Evaluation
  • Experience building evaluation frameworks for LLM systems - offline eval sets, continuous online evaluation, and regression detection
  • Instrumentation and traceability suitable for regulated enterprise environments using tools like LangSmith, Langfuse, etc.
Cloud
  • Hands-on AWS: containerized services (ECS/EKS), serverless (Lambda), data services (S3, DynamoDB, Redshift) and orchestration (Step Functions); Azure or GCP equivalents also valued
  • Familiarity with CI/CD pipelines and DevOps practices
  • Infrastructure as code with Terraform or CloudFormation, and mature CI/CD practice
Working traits
  • Strong analytical problem-solving with a bias to ownership and urgency
  • Clear cross-team communication, working directly with client stakeholders to translate business problems into technical roadmaps
  • Able to work productively in ambiguity from system-level documentation and ramp quickly in unfamiliar codebases
Good to Have
  • Experience with managed AI platforms - Amazon Bedrock, Vertex AI, Azure AI - paired with fluency in the underlying fundamentals
Roles & Responsibilities
  • Design and build agentic systems: Lead the architecture and implementation of tool-calling agents that combine retrieval, structured reasoning, and secure action execution with least-privilege access.
  • Productionize LLM applications: Build retrieval pipelines, prompt synthesis, response validation, and self-correction loops, backed by rigorous evaluation.
  • Own the full stack: Deliver the data pipelines, backend services, distributed compute, and orchestration layer that agentic systems depend on - not only the model invocation.
  • Engineer for reliability and governance: Build validator models, adversarial test suites, and policy checks; enforce deterministic fallbacks and rollback strategies; instrument continuous evaluation.
  • Optimize for cost and latency: Drive measurable improvements in token efficiency, response time, and unit economics against defined SLOs.
  • Codebase ownership: Build, maintain, and review high-quality Python and SQL, with an emphasis on reusable components, scalability, and performance.
  • Cloud integration: Deploy AI applications on AWS, Azure, or GCP with optimized resource usage and robust CI/CD.
  • Cross-functional collaboration: Partner with product owners, data scientists, and business SMEs to define requirements and deliver impactful AI products.
  • Mentoring and technical leadership: Set engineering standards and share knowledge across the team, raising the bar on AI and software engineering practice.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Agentic AI Engineer
Agentic AI Engineer

Tactical Edge • Washington

On-site
USD 150,000 - 210,000
Principal AI System Architect
Principal AI System Architect

Fruition Group US • San Jose (CA)

On-site
USD 190,000 - 260,000
Agentic AI Engineer
Agentic AI Engineer

Compunnel, Inc. • Dallas (TX)

On-site
USD 120,000 - 150,000
Senior Director, AI Engineering Lead
Senior Director, AI Engineering Lead

The Coca-Cola Company • Atlanta (GA)

On-site
USD 250,000 - 320,000
Title: Principal AI Engineer – Agentic AI | Contract to Hire | Irvine, CA (Hybrid) | AS
Title: Principal AI Engineer – Agentic AI | Contract to Hire | Irvine, CA (Hybrid) | AS

Central Business Solutions, Inc • Irvine (CA)

Hybrid
USD 180,000 - 240,000
Senior AI Architect
Senior AI Architect

Programmers.io • Dallas (TX)

On-site
USD 190,000 - 230,000
Lead Agentic AI engineer
Lead Agentic AI engineer

Envision Technology Solutions • Dallas (TX)

Hybrid
USD 140,000 - 200,000
Principle AI engineer
Principle AI engineer

Urbane Systems • New York (NY)

On-site
USD 180,000 - 260,000
AI Engineer
AI Engineer

Kaleidoscope Innovation • Fort Worth (TX)

On-site
USD 180,000 - 260,000
Principal Software Engineer
Principal Software Engineer

Compunnel, Inc. • Dallas (TX)

On-site
USD 130,000 - 180,000