Picture for Ke Cheng

Ke Cheng

Noise-Aware Shrinkage for Differentially Private Zeroth-Order Fine-Tuning of Large Language Models

Add code
Aug 04, 2026
Viaarxiv icon

M$^\text{4}$World: A Multi-view Multimodal Driving World Model for Interactive Object Manipulation and Minute-long Streaming

Add code
Jul 15, 2026
Viaarxiv icon

VistaGEN: Consistent Driving Video Generation with Fine-Grained Control Using Multiview Visual-Language Reasoning

Add code
Mar 30, 2026
Viaarxiv icon

When Convenience Becomes Risk: A Semantic View of Under-Specification in Host-Acting Agents

Add code
Mar 22, 2026
Viaarxiv icon

Recurrent Preference Memory for Efficient Long-Sequence Generative Recommendation

Add code
Feb 13, 2026
Viaarxiv icon

PRISM: Parallel Residual Iterative Sequence Model

Add code
Feb 12, 2026
Viaarxiv icon

Differentially Private Subspace Fine-Tuning for Large Language Models

Add code
Jan 16, 2026
Viaarxiv icon

RelayGR: Scaling Long-Sequence Generative Recommendation via Cross-Stage Relay-Race Inference

Add code
Jan 05, 2026
Viaarxiv icon

P/D-Device: Disaggregated Large Language Model between Cloud and Devices

Add code
Aug 12, 2025
Viaarxiv icon

CAKE: Cascading and Adaptive KV Cache Eviction with Layer Preferences

Add code
Mar 16, 2025
Figure 1 for CAKE: Cascading and Adaptive KV Cache Eviction with Layer Preferences
Figure 2 for CAKE: Cascading and Adaptive KV Cache Eviction with Layer Preferences
Figure 3 for CAKE: Cascading and Adaptive KV Cache Eviction with Layer Preferences
Figure 4 for CAKE: Cascading and Adaptive KV Cache Eviction with Layer Preferences
Viaarxiv icon