Picture for Paarth Neekhara

Paarth Neekhara

VoiceChat-TTS: A Low-Latency Continuous Speech Synthesis Model for Interactive Agents

Add code
Aug 13, 2026
Viaarxiv icon

MagpieTTS-LF: Inference-Time Long-Form Speech Generation Without Training on Long-Form data

Add code
Jun 16, 2026
Viaarxiv icon

HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset

Add code
Jun 04, 2025
Viaarxiv icon

Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance

Add code
Feb 07, 2025
Figure 1 for Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
Figure 2 for Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
Figure 3 for Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
Figure 4 for Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
Viaarxiv icon

Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference

Add code
Sep 18, 2024
Figure 1 for Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference
Viaarxiv icon

Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment

Add code
Jun 25, 2024
Figure 1 for Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
Figure 2 for Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
Figure 3 for Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
Figure 4 for Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
Viaarxiv icon

REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models

Add code
Oct 18, 2023
Figure 1 for REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models
Figure 2 for REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models
Figure 3 for REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models
Figure 4 for REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models
Viaarxiv icon

SelfVC: Voice Conversion With Iterative Refinement using Self Transformations

Add code
Oct 14, 2023
Figure 1 for SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
Figure 2 for SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
Figure 3 for SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
Figure 4 for SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
Viaarxiv icon

ACE-VC: Adaptive and Controllable Voice Conversion using Explicitly Disentangled Self-supervised Speech Representations

Add code
Feb 16, 2023
Viaarxiv icon

FastStamp: Accelerating Neural Steganography and Digital Watermarking of Images on FPGAs

Add code
Sep 26, 2022
Figure 1 for FastStamp: Accelerating Neural Steganography and Digital Watermarking of Images on FPGAs
Figure 2 for FastStamp: Accelerating Neural Steganography and Digital Watermarking of Images on FPGAs
Figure 3 for FastStamp: Accelerating Neural Steganography and Digital Watermarking of Images on FPGAs
Figure 4 for FastStamp: Accelerating Neural Steganography and Digital Watermarking of Images on FPGAs
Viaarxiv icon