From Models to Systems: A Survey of Explainability for Tool-Augmented Language Models and AI Agents
Large language models (LLMs) are increasingly being used as part of complex agentic systems that orchestrate the use of external tools, such as retrieval mechanisms or code interpreters. In this survey, we argue that this development necessitates a rethinking of the goals of explainable artificial intelligence (XAI): Rather than focusing on providing users with explanations for monolithic machine learning models, we need system-level explanations that also provide information about which and how tools are used, as well as how external execution traces causally influence system behavior. We provide an overview of the existing methods in explainable AI and discuss the limitations of monolithic XAI methods in agentic contexts. Finally, we highlight open challenges in providing faithful explanations for LLM-based systems.
Top
- Roth, Benjamin
- Edwards, Nicholas
- Hong, Pingjun
- Schoenegger, Loris
- Schuster, Sebastian
Top
Category |
Technical Report (Discussion Paper) |
Divisions |
Data Mining and Machine Learning |
Subjects |
Kuenstliche Intelligenz |
Date |
13 January 2026 |
Export |
Top
