🇩🇪 Berlin, Germany · 6h ago
Forward Deployed Engineer AI Inference (Intern)
Lyceum
LinkedInjuniorEnglish-friendly
About LyceumLyceum is a sovereign European AI inference provider. We run open-source models on our own GPU infrastructure, powered by 100% renewable energy, so teams can build with AI on their own terms – without giving up their data or getting locked into a single vendor. Backed by tier-1 investors, we're growing fast and our inference business is about to scale strongly.The RoleAs a Forward Deployed Engineer Intern, you own the technical side of our AI inference deals. You help customers figure out which models, GPUs and configurations fit their needs, run technical sessions with them, and work hand in hand with our commercial team to get deals closed.This is not a pure engineering role: you'll spend a lot of time with customers, and we're looking for someone who enjoys exactly that. Much of this work is still manual today – you'll help us turn it into product.What You'll DoMatch customer requirements to the right models, GPUs and configurations for dedicated inferenceRun technical sessions with customers and help them make confident decisionsWork in tandem with our commercial team to move deals forwardSupport serverless and API customizations, and help turn recurring ones into productTranslate customer needs into clear technical specs for our engineering teamCollect benchmarks and learnings that help us automate matching and customizationsWhat We're Looking ForStudies in computer science, data science or a closely related fieldInterest in or first exposure to AI inference: LLMs, inference engines, GPUsReal excitement about working with customers and the commercial side – not just the technical oneStrong communication skills: you talk confidently to customers and engineers alikeAn entrepreneurial mindset: give you an outcome, and you find a way without getting blockedYou stay calm and constructive when your ideas are challengedFluent EnglishBonus PointsCoursework or projects on inference engines or ML systems (e.g. vLLM, SGLang, TensorRT-LLM)Experience with GPU sizing, model serving, benchmarking or performance optimizationA previous internship at an AI infrastructure or inference companyStartup experienceGermanWhy Join UsCutting-edge work: Solve real AI inference problems with real customers from day oneRare mix: Combine technical and commercial work – unusual for an internshipReal ownership: Help build what becomes our product, from GPU matching to customizationsFounder access: Work directly with our product team and the foundersMission-driven team: Build sustainable, 100% renewable compute for the AI era, backed by top-tier investorsSourced from LinkedIn. Relocantly aggregates public job postings; apply on the original site.