Picture for Zijian Zhao

Zijian Zhao

Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperative Multi-Agent Reinforcement Learning

Add code
Aug 12, 2026
Viaarxiv icon

Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads

Add code
Aug 12, 2026
Viaarxiv icon

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization

Add code
Jul 20, 2026
Viaarxiv icon

Stage Light is Sequence$^2$: Multi-Light Control via Imitation Learning

Add code
May 05, 2026
Viaarxiv icon

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

Add code
Apr 15, 2026
Viaarxiv icon

Pushing the Boundaries of Natural Reasoning: Interleaved Bonus from Formal-Logic Verification

Add code
Jan 30, 2026
Viaarxiv icon

AutoFed: Manual-Free Federated Traffic Prediction via Personalized Prompt

Add code
Dec 31, 2025
Viaarxiv icon

Zero-Effort Image-to-Music Generation: An Interpretable RAG-based VLM Approach

Add code
Sep 26, 2025
Viaarxiv icon

Each to Their Own: Exploring the Optimal Embedding in RAG

Add code
Jul 23, 2025
Figure 1 for Each to Their Own: Exploring the Optimal Embedding in RAG
Figure 2 for Each to Their Own: Exploring the Optimal Embedding in RAG
Figure 3 for Each to Their Own: Exploring the Optimal Embedding in RAG
Figure 4 for Each to Their Own: Exploring the Optimal Embedding in RAG
Viaarxiv icon

A Short Overview of Multi-Modal Wi-Fi Sensing

Add code
May 10, 2025
Viaarxiv icon