Paper ID: 2303.02655

On Modifying a Neural Network's Perception

Manuel de Sousa Ribeiro, João Leite

Artificial neural networks have proven to be extremely useful models that have allowed for multiple recent breakthroughs in the field of Artificial Intelligence and many others. However, they are typically regarded as black boxes, given how difficult it is for humans to interpret how these models reach their results. In this work, we propose a method which allows one to modify what an artificial neural network is perceiving regarding specific human-defined concepts, enabling the generation of hypothetical scenarios that could help understand and even debug the neural network model. Through empirical evaluation, in a synthetic dataset and in the ImageNet dataset, we test the proposed method on different models, assessing whether the performed manipulations are well interpreted by the models, and analyzing how they react to them.

Submitted: Mar 5, 2023

Topics

Neural Network
Neural Network Model
Perception Aware
ImageNet Dataset
Human Concept

Links

arXiv PDF