Picture for Yanmin Qian

Yanmin Qian

Luna-TTS Family Technical Report

Add code
Aug 12, 2026
Viaarxiv icon

Towards Array-Invariant Speech Enhancement via Geometry-Aware Dynamic Convolution

Add code
Jul 21, 2026
Viaarxiv icon

TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

Add code
Jun 30, 2026
Viaarxiv icon

JASTIN: Aligning LLMs for Zero-Shot Audio and Speech Evaluation via Natural Language Instructions

Add code
May 06, 2026
Viaarxiv icon

Representation-Regularized Convolutional Audio Transformer for Audio Understanding

Add code
Jan 29, 2026
Viaarxiv icon

SLM-SS: Speech Language Model for Generative Speech Separation

Add code
Jan 27, 2026
Viaarxiv icon

DeepASMR: LLM-Based Zero-Shot ASMR Speech Generation for Anyone of Any Voice

Add code
Jan 22, 2026
Viaarxiv icon

ICASSP 2026 URGENT Speech Enhancement Challenge

Add code
Jan 20, 2026
Viaarxiv icon

USE: A Unified Model for Universal Sound Separation and Extraction

Add code
Dec 24, 2025
Viaarxiv icon

A Data-Centric Approach to Generalizable Speech Deepfake Detection

Add code
Dec 24, 2025
Figure 1 for A Data-Centric Approach to Generalizable Speech Deepfake Detection
Figure 2 for A Data-Centric Approach to Generalizable Speech Deepfake Detection
Figure 3 for A Data-Centric Approach to Generalizable Speech Deepfake Detection
Figure 4 for A Data-Centric Approach to Generalizable Speech Deepfake Detection
Viaarxiv icon