Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Towards Explaining STEM Document Classification using Mathematical Entity Linking

Sep 02, 2021

Philipp Scharpf, Moritz Schubotz, Bela Gipp

Figure 1 for Towards Explaining STEM Document Classification using Mathematical Entity Linking

Figure 2 for Towards Explaining STEM Document Classification using Mathematical Entity Linking

Figure 3 for Towards Explaining STEM Document Classification using Mathematical Entity Linking

Figure 4 for Towards Explaining STEM Document Classification using Mathematical Entity Linking

Share this with someone who'll enjoy it:

Abstract:Document subject classification is essential for structuring (digital) libraries and allowing readers to search within a specific field. Currently, the classification is typically made by human domain experts. Semi-supervised Machine Learning algorithms can support them by exploiting the labeled data to predict subject classes for unclassified new documents. However, while humans partly do, machines mostly do not explain the reasons for their decisions. Recently, explainable AI research to address the problem of Machine Learning decisions being a black box has increasingly gained interest. Explainer models have already been applied to the classification of natural language texts, such as legal or medical documents. Documents from Science, Technology, Engineering, and Mathematics (STEM) disciplines are more difficult to analyze, since they contain both textual and mathematical formula content. In this paper, we present first advances towards STEM document classification explainability using classical and mathematical Entity Linking. We examine relationships between textual and mathematical subject classes and entities, mining a collection of documents from the arXiv preprint repository (NTCIR and zbMATH dataset). The results indicate that mathematical entities have the potential to provide high explainability as they are a crucial part of a STEM document.

View paper on

Share this with someone who'll enjoy it:

Title:Towards Explaining STEM Document Classification using Mathematical Entity Linking

Paper and Code