Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Kaung Myat Kyaw

Graph Fusion Across Languages using Large Language Models

Mar 22, 2026

Kaung Myat Kyaw, Khush Agarwal, Jonathan Chan

Abstract:Combining multiple knowledge graphs (KGs) across linguistic boundaries is a persistent challenge due to semantic heterogeneity and the complexity of graph environments. We propose a framework for cross-lingual graph fusion, leveraging the in-context reasoning and multilingual semantic priors of Large Language Models (LLMs). The framework implements structural linearization by mapping triplets directly into natural language sequences (e.g., [head] [relation] [tail]), enabling the LLM to map relations and reconcile entities between an evolving fused graph ($G_{c}^{(t-1)}$) and a new candidate graph ($G_{t}$). Evaluated on the DBP15K dataset, this exploratory study demonstrates that LLMs can serve as a universal semantic bridge to resolve cross-lingual discrepancies. Results show the successful sequential agglomeration of multiple heterogeneous graphs, offering a scalable, modular solution for continuous knowledge synthesis in multi-source, multilingual environments.

Via

Access Paper or Ask Questions

A Framework for Synthetic Audio Conversations Generation using Large Language Models

Sep 02, 2024

Kaung Myat Kyaw, Jonathan Hoyin Chan

Figure 1 for A Framework for Synthetic Audio Conversations Generation using Large Language Models

Figure 2 for A Framework for Synthetic Audio Conversations Generation using Large Language Models

Figure 3 for A Framework for Synthetic Audio Conversations Generation using Large Language Models

Figure 4 for A Framework for Synthetic Audio Conversations Generation using Large Language Models

Abstract:In this paper, we introduce ConversaSynth, a framework designed to generate synthetic conversation audio using large language models (LLMs) with multiple persona settings. The framework first creates diverse and coherent text-based dialogues across various topics, which are then converted into audio using text-to-speech (TTS) systems. Our experiments demonstrate that ConversaSynth effectively generates highquality synthetic audio datasets, which can significantly enhance the training and evaluation of models for audio tagging, audio classification, and multi-speaker speech recognition. The results indicate that the synthetic datasets generated by ConversaSynth exhibit substantial diversity and realism, making them suitable for developing robust, adaptable audio-based AI systems.

* This work has been submitted for consideration at the WI-IAT'24 to be held in December 2024

Via

Access Paper or Ask Questions