Peering inside the 'Black Box': understanding and refining deep neural networks with representational similarity analysis
Deep neural networks, and Transformer models in particular, have achieved unprecedented success in natural language processing tasks. Despite this success, they are infamous for their status as black boxes. Specific details on how they encode and process high-level linguistic task-relevant information remain difficult...