TY - GEN
T1 - Evaluating Argumentation Graphs as Global Explainable Surrogate Models for Dense Neural Networks and Their Comparison with Decision Trees
AU - Vilone, Giulia
AU - Longo, Luca
N1 - DBLP License: DBLP's bibliographic metadata records provided through http://dblp.org/ are distributed under a Creative Commons CC0 1.0 Universal Public Domain Dedication. Although the bibliographic metadata records are provided consistent with CC0 1.0 Dedication, the content described by the metadata records is not. Content may be subject to copyright, rights of privacy, rights of publicity and other restrictions.
PY - 2025/10/19
Y1 - 2025/10/19
N2 - Rule-based methods are often used to learn surrogates of black-box models within Explainable Artificial Intelligence. Decision trees, among others, are routinely used for such purposes and inherently possess more explainability. Unfortunately, they might be convoluted in large-scale scenarios, with large sizes and many branches, thus hampering such inherent property. They also fail at modelling contrastive information and conflictuality among rules. This research proposes a novel method based on computational argumentation that aims to solve such shortcomings of decision trees. In particular, it proposes a mechanism for automatically extracting rules from trained dense neural networks, the arguments. It then describes a procedure for automatically extracting their conflicts using the notion of attacks. Arguments and attacks are integrated into argumentation frameworks, which are directed graphs that can be used as surrogate models for explaining black boxes. The dialectical status of the arguments in such graphs can be evaluated with formal semantics and then aggregated toward a rational outcome corresponding to the target classes of the black-box models. Such graphs are empirically evaluated against eight objective metrics, including completeness, correctness, fidelity, robustness, number of rules, average rule length, fraction of classes and fraction overlap. They are also compared with the corresponding surrogate decision trees. Findings show how argumentation graphs are highly comparable to decision trees regarding explainability across selected objective metrics. However, it is potentially more appealing given that argumentation graphs offer richer justification and explanations by modelling rules’ conflictuality.
AB - Rule-based methods are often used to learn surrogates of black-box models within Explainable Artificial Intelligence. Decision trees, among others, are routinely used for such purposes and inherently possess more explainability. Unfortunately, they might be convoluted in large-scale scenarios, with large sizes and many branches, thus hampering such inherent property. They also fail at modelling contrastive information and conflictuality among rules. This research proposes a novel method based on computational argumentation that aims to solve such shortcomings of decision trees. In particular, it proposes a mechanism for automatically extracting rules from trained dense neural networks, the arguments. It then describes a procedure for automatically extracting their conflicts using the notion of attacks. Arguments and attacks are integrated into argumentation frameworks, which are directed graphs that can be used as surrogate models for explaining black boxes. The dialectical status of the arguments in such graphs can be evaluated with formal semantics and then aggregated toward a rational outcome corresponding to the target classes of the black-box models. Such graphs are empirically evaluated against eight objective metrics, including completeness, correctness, fidelity, robustness, number of rules, average rule length, fraction of classes and fraction overlap. They are also compared with the corresponding surrogate decision trees. Findings show how argumentation graphs are highly comparable to decision trees regarding explainability across selected objective metrics. However, it is potentially more appealing given that argumentation graphs offer richer justification and explanations by modelling rules’ conflictuality.
UR - https://doi.org/10.1007/978-3-032-08333-3_5
U2 - 10.1007/978-3-032-08333-3_5
DO - 10.1007/978-3-032-08333-3_5
M3 - Conference proceeding
T3 - Communications in Computer and Information Science
SP - 89
EP - 112
BT - Explainable Artificial Intelligence (xAI 2025)
PB - Springer
T2 - 3rd World Conference on Explainable Artificial Intelligence, xAI 2025
Y2 - 9 July 2025 through 11 July 2025
ER -