Lyceum
Visit websiteForward Deployed Engineer – AI Inference
Salary not disclosedOnsite
- Engineering
- Berlin
- Internship
- Today
About the role
As a Forward Deployed Engineer Intern, you will manage the technical aspects of AI inference deals by assisting customers with model and hardware configurations. You will collaborate with the commercial team to close deals and translate customer requirements into technical specifications for the engineering team. This role bridges the gap between technical engineering and customer-facing commercial work.
Responsibilities
- Match customer requirements to the right models, GPUs and configurations for dedicated inference
- Run technical sessions with customers and help them make confident decisions
- Work in tandem with our commercial team to move deals forward
- Support serverless and API customizations, and help turn recurring ones into product
- Translate customer needs into clear technical specs for our engineering team
- Collect benchmarks and learnings that help us automate matching and customizations
Required skills
- AI inference
- LLMs
- Inference engines
- GPUs
- Communication skills
Nice to have
- vLLM
- SGLang
- TensorRT-LLM
- GPU sizing
- Model serving
- Benchmarking
- Performance optimization
- German
Qualifications
- Studies in computer science, data science or a closely related field
About the Company
Lyceum is an early-stage company building sovereign, GDPR-compliant AI infrastructure for the next generation of deep-tech. The team consists of engineers from hedge funds, big tech, and AI startups.