Schick keinen Standard-Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.
Lyceum in Berlin is offering an internship as a Forward Deployed Engineer Intern. You will own the technical side of AI inference deals, help customers select models, GPUs and configurations, and run technical sessions to drive deal closure.
This role blends engineering with client-facing work, turning manual processes into scalable product features, and requires strong English communication while collaborating with the commercial team.
As a Forward Deployed Engineer Intern, you own the technical side of our AI inference deals. You help customers figure out which models, GPUs and configurations fit their needs, run technical sessions with them, and work hand in hand with our commercial team to get deals closed.
This is not a pure engineering role: you'll spend a lot of time with customers, and we're looking for someone who enjoys exactly that. Much of this work is still manual today – you'll help us turn it into product.
Match customer requirements to the right models, GPUs and configurations for dedicated inference
Run technical sessions with customers and help them make confident decisions
Work in tandem with our commercial team to move deals forward
Support serverless and API customizations, and help turn recurring ones into product
Translate customer needs into clear technical specs for our engineering team
Collect benchmarks and learnings that help us automate matching and customizations
Studies in computer science, data science or a closely related field
Interest in or first exposure to AI inference: LLMs, inference engines, GPUs
Real excitement about working with customers and the commercial side – not just the technical one
Strong communication skills: you talk confidently to customers and engineers alike
An entrepreneurial mindset: give you an outcome, and you find a way without getting blocked
You stay calm and constructive when your ideas are challenged
Fluent English
Coursework or projects on inference engines or ML systems (e.g. vLLM, SGLang, TensorRT-LLM)
Experience with GPU sizing, model serving, benchmarking or performance optimization
A previous internship at an AI infrastructure or inference company
Startup experience
German
Outstanding team: Work with some of the best engineers in the world, coming from hedge funds, big tech, AI startups and top universities
Once in a lifetime opportunity: Early-stage company in the fastest-growing market in the world
Ownership: Shape how European AI companies access GPU compute
European mission: Build sovereign, GDPR-compliant AI infrastructure for the next generation of deep-tech