Picture for Peter Bell

Peter Bell

An Evaluation Framework for Text-to-Speech Voice Reconstruction

Add code
Jun 19, 2026
Viaarxiv icon

LISE : Listenable Interpretable Speaker Embeddings

Add code
Jun 19, 2026
Viaarxiv icon

UNet-Based Fusion and Exponential Moving Average Adaptation for Noise-Robust Speaker Recognition

Add code
Apr 28, 2026
Viaarxiv icon

The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs

Add code
Mar 15, 2026
Viaarxiv icon

Rethinking Discrete Speech Representation Tokens for Accent Generation

Add code
Jan 27, 2026
Viaarxiv icon

Learning Speech Representations with Variational Predictive Coding

Add code
Dec 31, 2025
Viaarxiv icon

TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systems

Add code
Jun 24, 2025
Figure 1 for TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systems
Figure 2 for TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systems
Figure 3 for TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systems
Figure 4 for TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systems
Viaarxiv icon

A Practitioner's Guide to Building ASR Models for Low-Resource Languages: A Case Study on Scottish Gaelic

Add code
Jun 05, 2025
Viaarxiv icon

Language Bias in Self-Supervised Learning For Automatic Speech Recognition

Add code
Jan 31, 2025
Figure 1 for Language Bias in Self-Supervised Learning For Automatic Speech Recognition
Figure 2 for Language Bias in Self-Supervised Learning For Automatic Speech Recognition
Figure 3 for Language Bias in Self-Supervised Learning For Automatic Speech Recognition
Figure 4 for Language Bias in Self-Supervised Learning For Automatic Speech Recognition
Viaarxiv icon

Beyond Oversmoothing: Evaluating DDPM and MSE for Scalable Speech Synthesis in ASR

Add code
Oct 16, 2024
Figure 1 for Beyond Oversmoothing: Evaluating DDPM and MSE for Scalable Speech Synthesis in ASR
Figure 2 for Beyond Oversmoothing: Evaluating DDPM and MSE for Scalable Speech Synthesis in ASR
Figure 3 for Beyond Oversmoothing: Evaluating DDPM and MSE for Scalable Speech Synthesis in ASR
Viaarxiv icon