Description
Title: Language models explainability
Summary: We will examine the specificity of language in explainability for both classification and generation. Study how explainability methods are adapted to text and what the state of the art currently is. Finally, through a tutorial, we will apply the methods seen in theory, notably with a bias detection use case.
Bio: Antonin is a 2nd-year PhD Student in the Explainability of Language Models, between the IRT Saint Exupéry and the IRIT. He has been working on explainability for the last 5 years and is part of the developing teams of the Xplique and Interpreto explainability libraries.