Towards a fuller understanding of neurons with Clustered Compositional Explanations

Biagio La Rosa*

Leilani H. Gilpin*

Roberto Capobianco

* External authors

NeurIPS 2023

2023

Abstract

Compositional Explanations is a method for identifying logical formulas of concepts that approximate the neurons' behavior. However, these explanations are linked to the small spectrum of neuron activations used to check the alignment (i.e., the highest ones), thus lacking completeness. In this paper, we propose a generalization, called Clustered Compositional Explanations, that combines Compositional Explanations with clustering and a novel search heuristic to approximate a broader spectrum of the neuron behavior. We define, and address the problems connected to the application of these methods to multiple ranges of activations, analyze the insights retrievable by using our algorithm, and propose some desiderata qualities that can be used to study the explanations returned by different algorithms.

Related Publications

XAI-Guided Continual Learning: Rationale, Methods, and Future Directions

WIREs, 2025
Michela Proietti*, Alessio Ragno*, Roberto Capobianco

Providing neural networks with the ability to learn new tasks sequentially represents one of the main challenges in artificial intelligence. Unlike humans, neural networks are prone to losing previously acquired knowledge upon learning new information, a phenomenon known as …

Interpretable Memory-based Prototypical Pooling

WSDM, 2025
Alessio Ragno*, Roberto Capobianco

Graph Neural Networks (GNNs) have proven their effectiveness in various graph-structured data applications. However, one of the significant challenges in the realm of GNNs is representation learning, a critical concept that bridges graph pooling, aimed at creating compressed…

Intermediate Layers of LLMs Align Best With the Brain by Balancing Short- and Long-Range Information

CCN, 2025
Michela Proietti*, Roberto Capobianco, Mariya Toneva

Contextual integration is fundamental to human language comprehension. Language models are a powerful tool for studying how contextual information influences brain activity. In this work, we analyze the brain alignment of three types of language models, which vary in how the…

SEE ALL

HOME
Publications
Towards a fuller understanding of neurons with Clustered Compositional Explanations

JOIN US

Shape the Future of AI with Sony AI

We want to hear from those of you who have a strong desire
to shape the future of AI.

LEARN MORE