Abstract: Neural networks increasingly inform consequential decisions, making their reliability increasingly important. Yet their internal mechanisms provide little evidence of whether decisions remain grounded in the training cases and which cases ultimately support or oppose their outcomes. Without this connection between decisions and training cases, users cannot determine whether a model has learned reliable decision patterns from data. This motivates a fundamental question: do neural networks preserve case structure? We establish a connection between neural networks and Case-Based Decision Theory (CBDT), showing that trained neural networks can preserve a recoverable case structure through their learned representations. Such a structure allows fitted decision margins to be decomposed into individual case contributions. We identify the conditions under which this recovered case structure admits a CBDT interpretation. We further establish decision consistency between this interpretation and the decision selected by the original neural network. Experiments on a controlled CBDT setting and three decision tasks based on real-world data validate our approach. These results connect neural network decisions with the cases that shape them. This connection allows model choices to be traced back to supporting and opposing cases, providing a basis for assessing the reliability of neural network decisions.
Read the original article:
