Picture for Yihong Zhuang

Yihong Zhuang

RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States

Add code
Aug 03, 2026
Viaarxiv icon

BridgeAlign: Bridging Preference Alignment for Humanities and Social Sciences

Add code
Jul 29, 2026
Viaarxiv icon

HSS-Synth: Humanities and Social Sciences Data Synthesis for LLMs

Add code
Jul 29, 2026
Viaarxiv icon

LLaDA2.1: Speeding Up Text Diffusion via Token Editing

Add code
Feb 09, 2026
Viaarxiv icon

Training-Trajectory-Aware Token Selection

Add code
Jan 15, 2026
Viaarxiv icon

HeartBench: Probing Core Dimensions of Anthropomorphic Intelligence in LLMs

Add code
Dec 26, 2025
Figure 1 for HeartBench: Probing Core Dimensions of Anthropomorphic Intelligence in LLMs
Figure 2 for HeartBench: Probing Core Dimensions of Anthropomorphic Intelligence in LLMs
Figure 3 for HeartBench: Probing Core Dimensions of Anthropomorphic Intelligence in LLMs
Figure 4 for HeartBench: Probing Core Dimensions of Anthropomorphic Intelligence in LLMs
Viaarxiv icon

LLaDA2.0: Scaling Up Diffusion Language Models to 100B

Add code
Dec 24, 2025
Viaarxiv icon

Merge-of-Thought Distillation

Add code
Sep 10, 2025
Figure 1 for Merge-of-Thought Distillation
Figure 2 for Merge-of-Thought Distillation
Figure 3 for Merge-of-Thought Distillation
Figure 4 for Merge-of-Thought Distillation
Viaarxiv icon

Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts

Add code
Aug 11, 2025
Viaarxiv icon

Knowledge Condensation Distillation

Add code
Jul 12, 2022
Figure 1 for Knowledge Condensation Distillation
Figure 2 for Knowledge Condensation Distillation
Figure 3 for Knowledge Condensation Distillation
Figure 4 for Knowledge Condensation Distillation
Viaarxiv icon