[Re] On the Reproducibility of Post-Hoc Concept Bottleneck Models
| Authors |
|
|---|---|
| Publication date | 05-2024 |
| Journal | Transactions on Machine Learning Research |
| Article number | 2212 |
| Volume | Issue number | 2024 | 5 |
| Number of pages | 28 |
| Organisations |
|
| Abstract |
To obtain state-of-the-art performance, many deeper artificial intelligence models sacrifice human explainability in their decision-making. One solution proposed for achieving top performance and retaining explainability is the Post-Hoc Concept Bottleneck Model (PCBM) (Yuksekgonul et al., 2023), which can convert the embeddings of any deep neural network into a set of human-interpretable concept weights. In this work, we reproduce and expand upon the findings of Yuksekgonul et al. (2023), showing that while their claims and results do generally hold, some of them could not be sufficiently replicated. Specifically, the claims relating to PCBM performance preservation and its non-requirement of labeled concept datasets were generally reproduced, whereas the one claiming its model editing capabilities was not. Beyond these results, our contributions to their work include evidence that PCBMs may work for audio classification problems, verification of the interpretability of their methods, and updates to their code for missing implementations. The code for our implementations can be found in https://github.com/dgcnz/FACT.
|
| Document type | Article |
| Language | English |
| Published at |
https://openreview.net/forum?id=8UfhCZjOV7
(Final published version)
|
| Other links | |
| Downloads |
[Re] On the Reproducibility of Post-Hoc Concept Bottleneck Models
(Final published version)
|
| Permalink to this page | |