Picture for Yossi Adi

Yossi Adi

Sid

Beyond Words: Towards Effective Modeling of Non-Verbal Vocalizations in ASR

Add code
Jul 02, 2026
Viaarxiv icon

Interleaved Speech Language Models Latently Work In Text

Add code
Jun 21, 2026
Viaarxiv icon

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

Add code
Apr 12, 2026
Viaarxiv icon

LLMs versus the Halting Problem: Revisiting Program Termination Prediction

Add code
Jan 26, 2026
Viaarxiv icon

What Does It Take to Be a Good AI Research Agent? Studying the Role of Ideation Diversity

Add code
Nov 19, 2025
Viaarxiv icon

DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion

Add code
Oct 23, 2025
Viaarxiv icon

GmSLM : Generative Marmoset Spoken Language Modeling

Add code
Sep 11, 2025
Figure 1 for GmSLM : Generative Marmoset Spoken Language Modeling
Figure 2 for GmSLM : Generative Marmoset Spoken Language Modeling
Figure 3 for GmSLM : Generative Marmoset Spoken Language Modeling
Figure 4 for GmSLM : Generative Marmoset Spoken Language Modeling
Viaarxiv icon

Discrete Audio Tokens: More Than a Survey!

Add code
Jun 12, 2025
Viaarxiv icon

Auto-Regressive vs Flow-Matching: a Comparative Study of Modeling Paradigms for Text-to-Music Generation

Add code
Jun 11, 2025
Viaarxiv icon

StressTest: Can YOUR Speech LM Handle the Stress?

Add code
May 28, 2025
Viaarxiv icon