Head of AI Engineering (f/m/x)

Neoshare

Berlin

Vor Ort

EUR 120.000 - 180.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Hebe dich für diese Rolle von der Masse ab — erstelle in etwa einer Minute einen maßgeschneiderten Lebenslauf und ein Anschreiben.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

30 vacation days
Jobticket
Urban Sports/EGYM subsidy
Workation
Flexible working hours

Zusammenfassung

Neoshare is a Munich-based fintech scale-up with offices across Germany and Bulgaria, seeking an experienced AI/ML engineering leader. You will transform a multi-year ML team into a high-throughput, production-grade function and partner with the CTO to define strategy and build a unified AI platform.

You will guide architecture for LLM access, RAG pipelines, and backend services, while driving reliability, cost efficiency, and scalable delivery across programs.

Qualifikationen

  • 5+ years in backend engineering with AI/ML in production.
  • 4+ years leading AI/ML teams.
  • Strong architecture and scalability experience.

Aufgaben

  • Lead and evolve AI/ML engineering function.
  • Partner with CTO on strategy and platform roadmap.
  • Own LLM gateway and RAG pipelines with robust observability.
  • Drive SLOs, cost tracking, and low-latency inference.
  • Mentor engineers, establish coding and research review practices.

Kenntnisse

Backend architecture
Java (JVM)
Node.js
Distributed systems
LLM/AI tooling
MLOps

Tools

NestJS
Kubernetes
Terraform
AWS

Jobbeschreibung

About neoshare

We’re a Munich-based AI-first fintech scale-up (founded 2019) with offices in Munich, Frankfurt, Berlin and Sofia. Our SaaS platform brings banks, investors, and advisors together to collaborate on complex financial deals making due diligence faster, smarter, and more transparent. Our AI features are already live with leading banks. Now we’re scaling.


The Role

Own and evolve our AI engineering function — transforming a 15–20 person ML team from research-heavy to a high-throughput, production-grade organization. You’ll partner with the CTO on strategy, build the platform that unifies LLM access, RAG, and backend services, and ship reliable, scalable AI features that change how banks work.


Key responsibilities


  • Team leadership and org build

    • Hire, mentor, and develop a high-performing team; set the technical bar, operating rhythms, and code/research review practices

    • Organize sub-teams (e.g., Core Modeling, AI Platform/Infra, Integrations) with clear ownership, SLOs, and on-call

    • Manage roadmap, capacity planning, and delivery across parallel initiatives



  • Architecture and platform

    • Own the LLM gateway: unified APIs and proxy layers for multi-provider routing (OpenAI, Gemini, Bedrock), with rate limits, fallbacks, and cost tracking

    • Build high-performance RAG pipelines (ingestion, embeddings, vector stores, caching) with robust observability and safety guardrails

    • Partner with Java/ NestJS teams to define clean async contracts, schemas, and eventing patterns; drive low-latency, scalable inference



  • Model lifecycle and operations

    • Lead end-to-end model and prompt lifecycle: data curation, training/fine-tuning, evaluation, deployment, rollback

    • Establish LLMOps / MLOps : model/prompt registries, CI/CD, canary/A/B tests, offline/online evals, drift and cost monitoring

    • Optimize inference throughput and cost (autoscaling, batching, quantization/distillation, caching)



  • Strategy and collaboration

    • Translate company goals into an AI/ML roadmap with measurable outcomes; balance exploration with reliability and cost

    • Own build-vs-buy/vendor strategy for models, infrastructure, and data services; manage budgets and SLAs



  • Governance and security

    • Implement data privacy, security, and compliance practices (RBAC, secrets, auditability); track prompt/model lineage and reproducibility

    • Define incident response, runbooks, and postmortems for AI features




Your profile


  • 5+ years as a backend engineer and 4+ years leading AI/ML engineering in production (10+ years total experience ideal)

  • Deep architecture expertise in Java (JVM) and/or Node.js ( NestJS ), distributed systems, APIs, microservices, and messaging/streaming

  • Hands-on with LLM stacks: orchestration (e.g., LangChain / LlamaIndex or custom), vector DBs (Pinecone, Qdrant , FAISS), cloud AI (e.g., AWS Bedrock)

  • Proven operation of systems at scale (millions of daily API calls) with strong SLOs, observability, and incident management

  • MLOps foundations: model registries, experiment tracking, CI/CD, Kubernetes, IaC (e.g., Terraform), security best practices

  • Excellent communication and stakeholder management; strong product sense focused on shipping user-fac­ing feature

  • Fluent German and English for daily team collaboration, stakeholder management, and technical documentation


Nice to have


  • Experience with GPU/accelerator serving and optimization ( vLLM , TGI, Triton, ONNX Runtime)

  • Cost optimization for LLM workloads (token budgets, dynamic routing, caching)

  • Evaluation and safety/red-teaming for generative systems; startup/high-growth experience


Impact metrics


  • Platform: adoption of a unified LLM gateway; standardized observability and cost reporting

  • Delivery: 2–3 user-fac­ing AI features shipped with clear SLOs and measurable impact

  • Reliability/cost: reduced average latency and cost per request; autoscaling and caching in place

  • Org: sub-team structure established ; improved code quality and on-time delivery; targeted hiring completed


Our stack


  • Backend: Java (JVM), Node.js ( NestJS ); event-driven microservices; API gateways/proxies

  • AI platform: Python, PyTorch , LLM orchestration, prompt pipelines/registry; vector DBs (Pinecone, Qdrant ); RAG services

  • Infra/DevOps: AWS (incl. Bedrock), Kubernetes, Terraform, CI/CD, Observability ( OpenTelemetry , Prometheus/Grafana)


Why us?

International & Inclusive Team: Collaboration with diverse teams at our locations in Munich, Frankfurt, Berlin, and Sofia.


Modern & Dog-friendly Offices: Ergonomic, green, and inspiring for collaboration and productivity.


Flexibility: 30 vacation days, flexible working hours.


Special Time Off: Additional half-day off on Christmas Eve and New Year’s Eve.


Workation: Work remotely for a limited period each year from selected destinations.


Wellbeing & Mobility Benefits: Support for well-being and sustainable lifestyle:



  • Urban Sports/EGYM Club subsidy:Monthly support for your membership.

  • Jobticket:50% monthly subsidy for the Deutschlandticket.

  • JobRad:Leasing of bicycles or e-bikes at attractive conditions.


Candidates must have the right to work in the EU; visa sponsorship is not provided for this role.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Head of AI Engineering (f/m/x)
Head of AI Engineering (f/m/x)

Neoshare • München

Vor Ort
EUR 120.000 - 180.000
30 vacation days
Flexible working hours
Time off on Christmas Eve/New Year’s E
+4
Head of AI Engineering (f/m/x)
Head of AI Engineering (f/m/x)

Neoshare • Frankfurt

Vor Ort
EUR 120.000 - 170.000
Head of AI Engineering (f/m/x)
Head of AI Engineering (f/m/x)

neoshare AG • München

Hybrid
EUR 180.000 - 240.000
International & inclusive team across?
30 vacation days
Hybrid work
+2
AI Platform Engineer (m/f/x)
AI Platform Engineer (m/f/x)

Scalable GmbH • Bayern

Vor Ort
EUR 70.000 - 90.000
Flexible and discounted sports activities
Monthly contribution for ‘Deutschland Jobticket’
Internal knowledge sharing sessions
+4
Junior AI Developer
Junior AI Developer

reeeliance IM GmbH • Hamburg

Vor Ort
EUR 45.000 - 60.000
Mentorship and onboarding
Cutting-edge workspace
Long-term stability
+2
Senior AI Software Engineer (m/f/d)
Senior AI Software Engineer (m/f/d)

Peter Park System GmbH • München

Hybrid
EUR 80.000 - 100.000
Company pension plan
Corporate benefits
JobRad bike leasing
+2
AI Platform Engineer (m/f/x)
AI Platform Engineer (m/f/x)

Scalable Capital • München

Vor Ort
EUR 90.000 - 140.000
Relocation support
Company pension
German language classes
+2
AI Platform Engineer (m/f/x)
AI Platform Engineer (m/f/x)

Scalable Capital • Berlin

Vor Ort
EUR 90.000 - 130.000
Deutschland Jobticket 50% monthly
Education Budget
German language classes
+3
Principal Engineer (m/f/x) - AI & Digital Systems
Principal Engineer (m/f/x) - AI & Digital Systems

appliedAI Initiative GmbH • München

Hybrid
EUR 110.000 - 150.000
Hybrid work model
Sports membership (Wellpass)
JobRad bike leasing
+2
AI Engineer (m/w/d)
AI Engineer (m/w/d)

STATWORX GmbH • Frankfurt

Vor Ort
EUR 90.000 - 120.000
Mentoring program
Flat hierarchies
Wellbeing offers
+2