Production AI Architect | Real-Time Voice Copilot Lead

Neurons Lab

Roma

In loco

EUR 90.000 - 130.000

Tempo pieno

5 giorni fa
Candidati tra i primi
Generatore di candidature

Una candidatura apposita per questa offerta — un curriculum e una lettera di presentazione personalizzati, perfettamente in linea con l'annuncio.

Supera i filtri ATS

Descrizione del lavoro

Neurons Lab in Rome, Italy, is seeking a senior AI engineer to own the architecture and delivery of the voice copilot product for the veterinary care client. You will manage real-time streaming STT, LLM field extraction, Chrome-extension delivery, and the AWS Bedrock infrastructure to production readiness.

The role focuses on reducing latency, validating accuracy, and controlling per-call costs, with ongoing knowledge transfer to the client team and Neurons Lab engineers.

Competenze

  • Real-time voice AI in production — shipped at least one real-time voice or speech product to real users.
  • 6+ years hands-on AI/ML engineering with strong recent LLM production practice.
  • Latency improvements and concurrency fixes demonstrated on a live system.
  • Consulting / client-facing seniority; calm and precise under detailed UAT scrutiny.

Mansioni

  • Own the full pipeline: streaming STT, LLM field extraction, Chrome-extension delivery, and AWS infrastructure.
  • Drive latency work: reduce P95 from ~6s toward ~2s and remove post-processing lag.
  • Run model A/B tests with golden-set evaluation for name and email accuracy.
  • Own evaluation and cost: Langfuse traces, dashboards, and per-call cost optimization.
  • Harden for production: 5–10+ concurrent calls, data isolation, monitoring, and safe rollback.
  • Ship epics end to end (SES email briefing); maintain a demo fallback for reliability.

Conoscenze

Real-time voice pipelines
LLM engineering
Observability
Low-latency inference
Python
English communication
Cost optimization

Strumenti

Langfuse
AWS Bedrock
SES
Chrome Extension

Descrizione del lavoro

Neurons Lab in Rome, Italy, is seeking a senior AI engineer to own the architecture and delivery of the voice copilot product for the veterinary care client. You will manage real-time streaming STT, LLM field extraction, Chrome-extension delivery, and the AWS Bedrock infrastructure to production readiness.

The role focuses on reducing latency, validating accuracy, and controlling per-call costs, with ongoing knowledge transfer to the client team and Neurons Lab engineers.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Agent Engineer, Italy (Rome)
Agent Engineer, Italy (Rome)

Wonderful Ltd. • Roma

In loco
EUR 45.000 - 65.000
Agent Engineer, Italy (Rome)
Agent Engineer, Italy (Rome)

Wonderful Ltd. • Roma

In loco
EUR 40.000 - 60.000
Agent Engineer, Italy (Milan)
Agent Engineer, Italy (Milan)

Wonderful Ltd. • Milano

In loco
EUR 40.000 - 70.000
AI Engineer - Site: Rome
AI Engineer - Site: Rome

Aieng • Roma

Ibrido
EUR 60.000 - 90.000
Restaurant-sized perks? No explicit
Senior AI Engineer - Client-Facing, Europe Travel
Senior AI Engineer - Client-Facing, Europe Travel

Nearform • Italia

In loco
EUR 90.000 - 150.000
Annual Bonus
Remote Working
Paid Time Off 28 days
+3
AI Research Engineer (Agentic Post-training)
AI Research Engineer (Agentic Post-training)

Jobgether • Milano

Remoto
EUR 70.000 - 110.000
Remote-first environment
Global team
Competitive compensation
Senior AI Engineer – LLMs & Agentic Systems - Freelance
Senior AI Engineer – LLMs & Agentic Systems - Freelance

Shakers • Roma

Ibrido
EUR 60.000 - 80.000
AI Engineer - Site: Rome
AI Engineer - Site: Rome

Aieng • Roma

Ibrido
EUR 55.000 - 75.000
Enterprise Solutions Engineer - Italy Italy
Enterprise Solutions Engineer - Italy Italy

ElevenLabs • Italia

In loco
EUR 55.000 - 75.000
AI Engineer: LLMs, RAG & Local Inference
AI Engineer: LLMs, RAG & Local Inference

Aieng • Lazio

Ibrido
EUR 50.000 - 75.000