Staff AI Engineer/AI Solution Architect, Conversation Intelligence Systems

Xenoss

New York (NY)

On-site

USD 180,000 - 240,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Xenoss is seeking a Staff AI Engineer / AI Solution Architect to lead the applied AI architecture for a long-term In-Call Assistant initiative in New York, NY (onsite). This role shapes a real-time conversational AI system with compliance guardrails and confidence management, guiding PoC-first development toward production-grade conversation intelligence.

You will define the end-to-end AI approach, set taxonomies for live conversations, and collaborate with AI engineering, data engineering,

Qualifications

  • Hands-on experience with applied AI/ML in production environments.
  • Experience with NLP or transcript-based systems.
  • Ability to design evaluation frameworks, not only run experiments.
  • Experience building or validating structured datasets from unstructured text.
  • Strong knowledge of LLM-based extraction, RAG, and fine-tuning trade-offs.

Responsibilities

  • Lead the applied AI architecture across the In-Call Assistant lifecycle, from data design to production readiness.
  • Design end-to-end AI architecture for the In-Call Assistant.
  • Define signal/taxonomies for live conversations.
  • Design training strategies for signal detection and specialist recommendation models.
  • Shape data preparation, annotation, and SME validation workflows.
  • Evaluate fine-tuning, post-training, RAG, and hybrid approaches.
  • Design low-latency signal detection, routing, and confidence management.
  • Define grounding, guardrails, policy-compliance behavior, and governance.

Skills

Applied AI/ML
NLP/Conversational AI
LLM Extraction
RAG & Fine-tuning
MLOps
Data Wrangling
Evaluation Frameworks

Tools

PyTorch
Hugging Face
RAG Pipelines

Job description

Xenoss is seeking a Staff AI Engineer / AI Solution Architect to lead the applied AI architecture for a long-term In-Call Assistant initiative in New York, NY (onsite). This role shapes a real-time conversational AI system that helps front-office employees during live customer interactions, with compliance guardrails, confidence management, and a clear path from PoC-first development to production-grade conversation intelligence.

What you’ll build

You’ll define the end-to-end AI approach for conversation intelligence, including the pipeline for low-latency signal detection, context preparation, specialist recommendation generation, and RAG over approved product and policy knowledge. The system is designed to identify customer needs, objections, and buying signals, then recommend required process steps using grounded, policy-compliant behavior.

Responsibilities
  • Lead the applied AI architecture across the In-Call Assistant lifecycle, from data and taxonomy design to model training, evaluation, and production readiness.
  • Design the end-to-end AI architecture for the In-Call Assistant.
  • Define signal and trigger taxonomies for live conversations.
  • Design training strategies for signal detection and specialist recommendation models.
  • Shape data preparation, annotation, and SME validation workflows.
  • Evaluate fine-tuning, post-training, RAG, and hybrid approaches.
  • Design low-latency signal detection, routing, context preparation, and confidence management.
  • Design evaluation frameworks, golden datasets, and model improvement cycles.
  • Define grounding, guardrails, abstention, and policy-compliance behavior.
  • Make trade-offs across model quality, latency, cost, explainability, and governance.
  • Partner with AI engineering, data engineering, MLOps, and client SMEs.
  • Act as an escalation point for AI architecture, evaluation, and data strategy decisions.
  • Translate ambiguous business use cases into testable AI hypotheses and validation plans.
PoC-first scope and delivery context
  • Define the AI approach for the conversation intelligence PoC.
  • Establish the event / intent / insight taxonomy.
  • Define the golden dataset strategy and annotation workflow.
  • Establish evaluation frameworks and acceptance criteria.
  • Drive trade-offs between accuracy, explainability, latency, cost, and governance.
  • Decide which modeling approaches fit each use case.
  • Work within a cross-functional team spanning AI engineering, data engineering, MLOps, solution architecture, and client stakeholders.
What you bring
  • Hands-on experience with applied AI / ML systems in production-oriented environments.
  • Experience with NLP, conversational AI, or transcript-based intelligence systems.
  • Ability to design evaluation frameworks, not only run experiments.
  • Experience building or validating structured datasets from unstructured text.
  • Strong understanding of LLM-based extraction, classification, RAG, and fine-tuning trade-offs.
  • Practical knowledge of classical ML or predictive modeling.
  • Understanding of probability-based prediction, calibration, and outcome evaluation.
  • Comfort working with messy enterprise data and incomplete labels.
  • Ability to communicate with both technical teams and business stakeholders.
  • Strong ownership of ambiguity, scope control, and PoC validation strategy.
  • Financial services domain exposure.
  • Experience with sales, call center, or customer conversation analytics.
  • Speech / ASR pipeline familiarity.
  • Model governance and auditability experience.
  • Experience with real-time AI systems or low-latency inference.
  • Experience combining unstructured conversation signals with structured CRM, transaction, or customer profile data.
  • Experience designing golden datasets and SME review workflows.
Technologies you’ll use
  • LLM, SFT, DPO, preference optimization, LoRA, QLoRA, PEFT, PyTorch, Hugging Face
  • RAG, embeddings, retrieval
  • MLOps, monitoring, feedback loops, and model governance

Infrastructure and data residency: work is executed within the client perimeter using the client environment only (no external training or data processing environments). Delivery is PoC-first, with evolution toward live production conversation intelligence and prediction systems. Engagement is FTE-equivalent via long-term B2B contract.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff AI Engineer/AI Architect, Conversation Intelligence Systems (On-Site, New York)
Staff AI Engineer/AI Architect, Conversation Intelligence Systems (On-Site, New York)

Xenoss • Northern (KY), New York (NY)

Hybrid
USD 150,000 - 190,000
Staff AI Architect: Real-Time In-Call Assistant
Staff AI Architect: Real-Time In-Call Assistant

Xenoss • New York (NY)

On-site
USD 180,000 - 240,000
Staff AI Architect – Real-Time Conversation Intelligence (NY)
Staff AI Architect – Real-Time Conversation Intelligence (NY)

Xenoss • Northern (KY), New York (NY)

Hybrid
USD 150,000 - 190,000
Sr AI Engineer - Software
Sr AI Engineer - Software

Snowrelic Inc • Jersey City (NJ)

On-site
USD 150,000 - 230,000
Applied AI Engineer
Applied AI Engineer

SherlockTalent • Miami (FL)

On-site
USD 120,000 - 140,000
Solid Benefits
Referral bonus of $2,500
Senior AI Engineer, Voice & Agentic Systems
Senior AI Engineer, Voice & Agentic Systems

FIS Management Services LLC • Jacksonville (FL)

On-site
USD 140,000 - 190,000
Senior Conversational AI Delivery Engineer
Senior Conversational AI Delivery Engineer

Omilia • Myrtle Point (OR)

On-site
USD 120,000 - 180,000
Vacation days
Professional development
Apple gear
AI Agents Applied Research/Engineering Lead - Vice President
AI Agents Applied Research/Engineering Lead - Vice President

JPMorgan Chase & Co. • New York (NY)

On-site
USD 180,000 - 240,000
Staff Software Engineer, AI Systems
Staff Software Engineer, AI Systems

Via Licensing Corporation • Atlanta (GA)

On-site
USD 160,000 - 260,000
Senior AI Engineer
Senior AI Engineer

Propio • Overland Park (KS)

Hybrid
USD 120,000 - 180,000