Picture for Zihe Huang

Zihe Huang

When Does Muon Help Agentic Reinforcement Learning?

Add code
Jul 20, 2026
Viaarxiv icon

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade

Add code
Jul 07, 2026
Viaarxiv icon

Chain-of-Memory: Lightweight Memory Construction with Dynamic Evolution for LLM Agents

Add code
Jan 14, 2026
Viaarxiv icon

Projecting Out the Malice: A Global Subspace Approach to LLM Detoxification

Add code
Jan 09, 2026
Viaarxiv icon