Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Adrian Paschke

Retrieval-Augmented Generation-based Relation Extraction

Apr 20, 2024

Sefika Efeoglu, Adrian Paschke

Abstract:Information Extraction (IE) is a transformative process that converts unstructured text data into a structured format by employing entity and relation extraction (RE) methodologies. The identification of the relation between a pair of entities plays a crucial role within this framework. Despite the existence of various techniques for relation extraction, their efficacy heavily relies on access to labeled data and substantial computational resources. In addressing these challenges, Large Language Models (LLMs) emerge as promising solutions; however, they might return hallucinating responses due to their own training data. To overcome these limitations, Retrieved-Augmented Generation-based Relation Extraction (RAG4RE) in this work is proposed, offering a pathway to enhance the performance of relation extraction tasks. This work evaluated the effectiveness of our RAG4RE approach utilizing different LLMs. Through the utilization of established benchmarks, such as TACRED, TACREV, Re-TACRED, and SemEval RE datasets, our aim is to comprehensively evaluate the efficacy of our RAG4RE approach. In particularly, we leverage prominent LLMs including Flan T5, Llama2, and Mistral in our investigation. The results of our study demonstrate that our RAG4RE approach surpasses performance of traditional RE approaches based solely on LLMs, particularly evident in the TACRED dataset and its variations. Furthermore, our approach exhibits remarkable performance compared to previous RE methodologies across both TACRED and TACREV datasets, underscoring its efficacy and potential for advancing RE tasks in natural language processing.

* Submitted to Semantic Web Journal. Under Review

Via

Access Paper or Ask Questions

Hybrid Quantum Machine Learning Assisted Classification of COVID-19 from Computed Tomography Scans

Oct 04, 2023

Leo Sünkel, Darya Martyniuk, Julia J. Reichwald, Andrei Morariu, Raja Havish Seggoju, Philipp Altmann, Christoph Roch, Adrian Paschke

Figure 1 for Hybrid Quantum Machine Learning Assisted Classification of COVID-19 from Computed Tomography Scans

Figure 2 for Hybrid Quantum Machine Learning Assisted Classification of COVID-19 from Computed Tomography Scans

Figure 3 for Hybrid Quantum Machine Learning Assisted Classification of COVID-19 from Computed Tomography Scans

Figure 4 for Hybrid Quantum Machine Learning Assisted Classification of COVID-19 from Computed Tomography Scans

Abstract:Practical quantum computing (QC) is still in its infancy and problems considered are usually fairly small, especially in quantum machine learning when compared to its classical counterpart. Image processing applications in particular require models that are able to handle a large amount of features, and while classical approaches can easily tackle this, it is a major challenge and a cause for harsh restrictions in contemporary QC. In this paper, we apply a hybrid quantum machine learning approach to a practically relevant problem with real world-data. That is, we apply hybrid quantum transfer learning to an image processing task in the field of medical image processing. More specifically, we classify large CT-scans of the lung into COVID-19, CAP, or Normal. We discuss quantum image embedding as well as hybrid quantum machine learning and evaluate several approaches to quantum transfer learning with various quantum circuits and embedding techniques.

Via

Access Paper or Ask Questions

Hateful Messages: A Conversational Data Set of Hate Speech produced by Adolescents on Discord

Sep 04, 2023

Jan Fillies, Silvio Peikert, Adrian Paschke

Abstract:With the rise of social media, a rise of hateful content can be observed. Even though the understanding and definitions of hate speech varies, platforms, communities, and legislature all acknowledge the problem. Therefore, adolescents are a new and active group of social media users. The majority of adolescents experience or witness online hate speech. Research in the field of automated hate speech classification has been on the rise and focuses on aspects such as bias, generalizability, and performance. To increase generalizability and performance, it is important to understand biases within the data. This research addresses the bias of youth language within hate speech classification and contributes by providing a modern and anonymized hate speech youth language data set consisting of 88.395 annotated chat messages. The data set consists of publicly available online messages from the chat platform Discord. ~6,42% of the messages were classified by a self-developed annotation schema as hate speech. For 35.553 messages, the user profiles provided age annotations setting the average author age to under 20 years old.

Via

Access Paper or Ask Questions

ContCommRTD: A Distributed Content-based Misinformation-aware Community Detection System for Real-Time Disaster Reporting

Jan 30, 2023

Elena-Simona Apostol, Ciprian-Octavian Truică, Adrian Paschke

Abstract:Real-time social media data can provide useful information on evolving hazards. Alongside traditional methods of disaster detection, the integration of social media data can considerably enhance disaster management. In this paper, we investigate the problem of detecting geolocation-content communities on Twitter and propose a novel distributed system that provides in near real-time information on hazard-related events and their evolution. We show that content-based community analysis leads to better and faster dissemination of reports on hazards. Our distributed disaster reporting system analyzes the social relationship among worldwide geolocated tweets, and applies topic modeling to group tweets by topics. Considering for each tweet the following information: user, timestamp, geolocation, retweets, and replies, we create a publisher-subscriber distribution model for topics. We use content similarity and the proximity of nodes to create a new model for geolocation-content based communities. Users can subscribe to different topics in specific geographical areas or worldwide and receive real-time reports regarding these topics. As misinformation can lead to increase damage if propagated in hazards related tweets, we propose a new deep learning model to detect fake news. The misinformed tweets are then removed from display. We also show empirically the scalability capabilities of the proposed system.

Via

Access Paper or Ask Questions

EDSA-Ensemble: an Event Detection Sentiment Analysis Ensemble Architecture

Jan 30, 2023

Alexandru Petrescu, Ciprian-Octavian Truică, Elena-Simona Apostol, Adrian Paschke

Abstract:As global digitization continues to grow, technology becomes more affordable and easier to use, and social media platforms thrive, becoming the new means of spreading information and news. Communities are built around sharing and discussing current events. Within these communities, users are enabled to share their opinions about each event. Using Sentiment Analysis to understand the polarity of each message belonging to an event, as well as the entire event, can help to better understand the general and individual feelings of significant trends and the dynamics on online social networks. In this context, we propose a new ensemble architecture, EDSA-Ensemble (Event Detection Sentiment Analysis Ensemble), that uses Event Detection and Sentiment Analysis to improve the detection of the polarity for current events from Social Media. For Event Detection, we use techniques based on Information Diffusion taking into account both the time span and the topics. To detect the polarity of each event, we preprocess the text and employ several Machine and Deep Learning models to create an ensemble model. The preprocessing step includes several word representation models, i.e., raw frequency, TFIDF, Word2Vec, and Transformers. The proposed EDSA-Ensemble architecture improves the event sentiment classification over the individual Machine and Deep Learning models.

Via

Access Paper or Ask Questions

Knowledge Augmented Machine Learning with Applications in Autonomous Driving: A Survey

May 10, 2022

Julian Wörmann, Daniel Bogdoll, Etienne Bührle, Han Chen, Evaristus Fuh Chuo, Kostadin Cvejoski, Ludger van Elst, Tobias Gleißner, Philip Gottschall, Stefan Griesche(+36 more)

Figure 1 for Knowledge Augmented Machine Learning with Applications in Autonomous Driving: A Survey

Figure 2 for Knowledge Augmented Machine Learning with Applications in Autonomous Driving: A Survey

Figure 3 for Knowledge Augmented Machine Learning with Applications in Autonomous Driving: A Survey

Figure 4 for Knowledge Augmented Machine Learning with Applications in Autonomous Driving: A Survey

Abstract:The existence of representative datasets is a prerequisite of many successful artificial intelligence and machine learning models. However, the subsequent application of these models often involves scenarios that are inadequately represented in the data used for training. The reasons for this are manifold and range from time and cost constraints to ethical considerations. As a consequence, the reliable use of these models, especially in safety-critical applications, is a huge challenge. Leveraging additional, already existing sources of knowledge is key to overcome the limitations of purely data-driven approaches, and eventually to increase the generalization capability of these models. Furthermore, predictions that conform with knowledge are crucial for making trustworthy and safe decisions even in underrepresented scenarios. This work provides an overview of existing techniques and methods in the literature that combine data-based models with existing knowledge. The identified approaches are structured according to the categories integration, extraction and conformity. Special attention is given to applications in the field of autonomous driving.

* 93 pages

Via

Access Paper or Ask Questions

Variational Quanvolutional Neural Networks with enhanced image encoding

Jun 23, 2021

Denny Mattern, Darya Martyniuk, Henri Willems, Fabian Bergmann, Adrian Paschke

Figure 1 for Variational Quanvolutional Neural Networks with enhanced image encoding

Figure 2 for Variational Quanvolutional Neural Networks with enhanced image encoding

Figure 3 for Variational Quanvolutional Neural Networks with enhanced image encoding

Figure 4 for Variational Quanvolutional Neural Networks with enhanced image encoding

Abstract:Image classification is an important task in various machine learning applications. In recent years, a number of classification methods based on quantum machine learning and different quantum image encoding techniques have been proposed. In this paper, we study the effect of three different quantum image encoding approaches on the performance of a convolution-inspired hybrid quantum-classical image classification algorithm called quanvolutional neural network (QNN). We furthermore examine the effect of variational - i.e. trainable - quantum circuits on the classification results. Our experiments indicate that some image encodings are better suited for variational circuits. However, our experiments show as well that there is not one best image encoding, but that the choice of the encoding depends on the specific constraints of the application.

Via

Access Paper or Ask Questions

TopicsRanksDC: Distance-based Topic Ranking applied on Two-Class Data

May 17, 2021

Malik Yousef, Jamal Al Qundus, Silvio Peikert, Adrian Paschke

Figure 1 for TopicsRanksDC: Distance-based Topic Ranking applied on Two-Class Data

Figure 2 for TopicsRanksDC: Distance-based Topic Ranking applied on Two-Class Data

Figure 3 for TopicsRanksDC: Distance-based Topic Ranking applied on Two-Class Data

Figure 4 for TopicsRanksDC: Distance-based Topic Ranking applied on Two-Class Data

Abstract:In this paper, we introduce a novel approach named TopicsRanksDC for topics ranking based on the distance between two clusters that are generated by each topic. We assume that our data consists of text documents that are associated with two-classes. Our approach ranks each topic contained in these text documents by its significance for separating the two-classes. Firstly, the algorithm detects topics using Latent Dirichlet Allocation (LDA). The words defining each topic are represented as two clusters, where each one is associated with one of the classes. We compute four distance metrics, Single Linkage, Complete Linkage, Average Linkage and distance between the centroid. We compare the results of LDA topics and random topics. The results show that the rank for LDA topics is much higher than random topics. The results of TopicsRanksDC tool are promising for future work to enable search engines to suggest related topics.

* International Conference on Database and Expert Systems Applications DEXA 2020: Database and Expert Systems Applications pp 11-21
* 10 pages, 5 figures

Via

Access Paper or Ask Questions

AI supported Topic Modeling using KNIME-Workflows

Apr 15, 2021

Jamal Al Qundus, Silvio Peikert, Adrian Paschke

Figure 1 for AI supported Topic Modeling using KNIME-Workflows

Figure 2 for AI supported Topic Modeling using KNIME-Workflows

Figure 3 for AI supported Topic Modeling using KNIME-Workflows

Figure 4 for AI supported Topic Modeling using KNIME-Workflows

Abstract:Topic modeling algorithms traditionally model topics as list of weighted terms. These topic models can be used effectively to classify texts or to support text mining tasks such as text summarization or fact extraction. The general procedure relies on statistical analysis of term frequencies. The focus of this work is on the implementation of the knowledge-based topic modelling services in a KNIME workflow. A brief description and evaluation of the DBPedia-based enrichment approach and the comparative evaluation of enriched topic models will be outlined based on our previous work. DBpedia-Spotlight is used to identify entities in the input text and information from DBpedia is used to extend these entities. We provide a workflow developed in KNIME implementing this approach and perform a result comparison of topic modeling supported by knowledge base information to traditional LDA. This topic modeling approach allows semantic interpretation both by algorithms and by humans.

* 7 pages, 7 figures. Qurator2020 - Conference on Digital Curation Technologies

Via

Access Paper or Ask Questions

ROC: An Ontology for Country Responses towards COVID-19

Apr 15, 2021

Jamal Al Qundus, Ralph Schäfermeier, Naouel Karam, Silvio Peikert, Adrian Paschke

Figure 1 for ROC: An Ontology for Country Responses towards COVID-19

Figure 2 for ROC: An Ontology for Country Responses towards COVID-19

Figure 3 for ROC: An Ontology for Country Responses towards COVID-19

Abstract:The ROC ontology for country responses to COVID-19 provides a model for collecting, linking and sharing data on the COVID-19 pandemic. It follows semantic standardization (W3C standards RDF, OWL, SPARQL) for the representation of concepts and creation of vocabularies. ROC focuses on country measures and enables the integration of data from heterogeneous data sources. The proposed ontology is intended to facilitate statistical analysis to study and evaluate the effectiveness and side effects of government responses to COVID-19 in different countries. The ontology contains data collected by OxCGRT from publicly available information. This data has been compiled from information provided by ECDC for most countries, as well as from various repositories used to collect data on COVID-19.

* Qurator2021 - Conference on Digital Curation Technologies
* 10 pages, 3 figures

Via

Access Paper or Ask Questions