Alina Dallmann
Alina Dallmann is a computer scientist currently working as a Data Scientist at scieneers GmbH. Her enthusiasm for classical software development and data-driven projects has recently come together in various projects focused on building retrieval-augmented generation (RAG) systems.
Session
Using LiteLLM in a Real-World RAG System: What Worked and What Didn’t
LiteLLM provides a unified interface to work with multiple LLM providers—but how well does it hold up in practice? In this talk, I’ll share how we used LiteLLM in a production system to simplify model access and handle token budgets. I’ll outline the benefits, the hidden trade-offs, and the situations where the abstraction helped—or got in the way. This is a practical, developer-focused session on integrating LiteLLM into real workflows, including lessons learned around deployment, limitations, and decision points. If you’re considering LiteLLM, this talk offers a grounded look at using it beyond simple prototypes.