Picture for Yilin Yang

Yilin Yang

Allocation Before Ranking: Decoupled Token Compression for OmniLLMs

Add code
Aug 03, 2026
Viaarxiv icon

Holographic MIMO-assisted Multiuser Transmission with Electromagnetic Exposure Constraints

Add code
Jul 13, 2026
Viaarxiv icon

From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users

Add code
May 30, 2026
Viaarxiv icon

HII-DPO: Eliminate Hallucination via Accurate Hallucination-Inducing Counterfactual Images

Add code
Feb 11, 2026
Viaarxiv icon

The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes

Add code
Jan 15, 2026
Viaarxiv icon

Prototype-Driven Multi-Feature Generation for Visible-Infrared Person Re-identification

Add code
Sep 09, 2024
Figure 1 for Prototype-Driven Multi-Feature Generation for Visible-Infrared Person Re-identification
Figure 2 for Prototype-Driven Multi-Feature Generation for Visible-Infrared Person Re-identification
Figure 3 for Prototype-Driven Multi-Feature Generation for Visible-Infrared Person Re-identification
Figure 4 for Prototype-Driven Multi-Feature Generation for Visible-Infrared Person Re-identification
Viaarxiv icon

Language-Informed Beam Search Decoding for Multilingual Machine Translation

Add code
Aug 11, 2024
Viaarxiv icon

MSLM-S2ST: A Multitask Speech Language Model for Textless Speech-to-Speech Translation with Speaker Style Preservation

Add code
Mar 19, 2024
Figure 1 for MSLM-S2ST: A Multitask Speech Language Model for Textless Speech-to-Speech Translation with Speaker Style Preservation
Figure 2 for MSLM-S2ST: A Multitask Speech Language Model for Textless Speech-to-Speech Translation with Speaker Style Preservation
Figure 3 for MSLM-S2ST: A Multitask Speech Language Model for Textless Speech-to-Speech Translation with Speaker Style Preservation
Figure 4 for MSLM-S2ST: A Multitask Speech Language Model for Textless Speech-to-Speech Translation with Speaker Style Preservation
Viaarxiv icon

An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis

Add code
Mar 19, 2024
Figure 1 for An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis
Figure 2 for An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis
Figure 3 for An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis
Figure 4 for An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis
Viaarxiv icon

Seamless: Multilingual Expressive and Streaming Speech Translation

Add code
Dec 08, 2023
Figure 1 for Seamless: Multilingual Expressive and Streaming Speech Translation
Figure 2 for Seamless: Multilingual Expressive and Streaming Speech Translation
Figure 3 for Seamless: Multilingual Expressive and Streaming Speech Translation
Figure 4 for Seamless: Multilingual Expressive and Streaming Speech Translation
Viaarxiv icon