👤 Speaker : Simone Balloccu
💬 Titre : How expert is your "expert AI" in the real world?
Abstract : Large-scale data annotation, training, and alignment bring the promise of "Expert AI". Depending on who you ask, this ranges from models capable of assisting domain experts in their work to superhuman AGI, which may replace you in the near future. Evidence of this "expertise" often comes from benchmarks, leaderboards, and sometimes even ad-hoc cases. In this talk, we take a closer (and perhaps, critical) look at this evaluation paradigm: how "expert" are these models when we apply them to really challenging real-world scenarios? We will cover three hard cases: irrelevant information, subjective preferences, and enterprise QA. Our "expert AI" might gracefully collapse, teaching us a lesson on what we need to do before we deploy these systems in sensitive domains.
Short bio : Simone Balloccu is a computer scientist with 9 years of research experience in NLP & AI. He worked within several EU-funded projects, including Horizon 2020, and ERC, focusing on AI for mental health and behaviour change, human evaluation, and more generally on AI applied to expert domains. He leads the “NLP for expert domains” lab at TU Darmstadt, researching Expert-AI interaction and cooperation. His current research involves efficient RAG systems over corporate knowledge, Multimodal NLP applied to mental health, and modelling expert preferences in LLMs.
📅 Date : 02/10/2026
📍 Local : Maison Des Langues A118 (FIAL) ( Google Map link )