Picture for Neil Zeghidour

Neil Zeghidour

PSL, FAIR, LSCP

MAD Speech: Measures of Acoustic Diversity of Speech

Add code
Apr 16, 2024
Figure 1 for MAD Speech: Measures of Acoustic Diversity of Speech
Figure 2 for MAD Speech: Measures of Acoustic Diversity of Speech
Figure 3 for MAD Speech: Measures of Acoustic Diversity of Speech
Figure 4 for MAD Speech: Measures of Acoustic Diversity of Speech
Viaarxiv icon

MusicRL: Aligning Music Generation to Human Preferences

Add code
Feb 06, 2024
Viaarxiv icon

TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition

Add code
Aug 21, 2023
Figure 1 for TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition
Figure 2 for TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition
Figure 3 for TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition
Viaarxiv icon

AudioPaLM: A Large Language Model That Can Speak and Listen

Add code
Jun 22, 2023
Figure 1 for AudioPaLM: A Large Language Model That Can Speak and Listen
Figure 2 for AudioPaLM: A Large Language Model That Can Speak and Listen
Figure 3 for AudioPaLM: A Large Language Model That Can Speak and Listen
Figure 4 for AudioPaLM: A Large Language Model That Can Speak and Listen
Viaarxiv icon

SoundStorm: Efficient Parallel Audio Generation

Add code
May 16, 2023
Figure 1 for SoundStorm: Efficient Parallel Audio Generation
Figure 2 for SoundStorm: Efficient Parallel Audio Generation
Figure 3 for SoundStorm: Efficient Parallel Audio Generation
Figure 4 for SoundStorm: Efficient Parallel Audio Generation
Viaarxiv icon

LMCodec: A Low Bitrate Speech Codec With Causal Transformer Models

Add code
Mar 23, 2023
Figure 1 for LMCodec: A Low Bitrate Speech Codec With Causal Transformer Models
Figure 2 for LMCodec: A Low Bitrate Speech Codec With Causal Transformer Models
Figure 3 for LMCodec: A Low Bitrate Speech Codec With Causal Transformer Models
Figure 4 for LMCodec: A Low Bitrate Speech Codec With Causal Transformer Models
Viaarxiv icon

Speech Intelligibility Classifiers from 550k Disordered Speech Samples

Add code
Mar 15, 2023
Figure 1 for Speech Intelligibility Classifiers from 550k Disordered Speech Samples
Figure 2 for Speech Intelligibility Classifiers from 550k Disordered Speech Samples
Figure 3 for Speech Intelligibility Classifiers from 550k Disordered Speech Samples
Figure 4 for Speech Intelligibility Classifiers from 550k Disordered Speech Samples
Viaarxiv icon

DNArch: Learning Convolutional Neural Architectures by Backpropagation

Add code
Feb 10, 2023
Figure 1 for DNArch: Learning Convolutional Neural Architectures by Backpropagation
Figure 2 for DNArch: Learning Convolutional Neural Architectures by Backpropagation
Figure 3 for DNArch: Learning Convolutional Neural Architectures by Backpropagation
Figure 4 for DNArch: Learning Convolutional Neural Architectures by Backpropagation
Viaarxiv icon

Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision

Add code
Feb 07, 2023
Figure 1 for Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
Figure 2 for Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
Figure 3 for Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
Figure 4 for Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
Viaarxiv icon

SingSong: Generating musical accompaniments from singing

Add code
Jan 30, 2023
Figure 1 for SingSong: Generating musical accompaniments from singing
Figure 2 for SingSong: Generating musical accompaniments from singing
Figure 3 for SingSong: Generating musical accompaniments from singing
Figure 4 for SingSong: Generating musical accompaniments from singing
Viaarxiv icon