Picture for Zhijing Wu

Zhijing Wu

The Weakest Link Tells It All: Outcome-Supervised Process Reward Modeling via Learnable Credit Assignment

Add code
Jun 26, 2026
Viaarxiv icon

EvoRubrics: Dynamic Rubrics as Rewards via Adversarial Co-Evolution for LLM Reinforcement Learning

Add code
Jun 22, 2026
Viaarxiv icon

PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning

Add code
May 11, 2026
Viaarxiv icon

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations?

Add code
Apr 26, 2026
Viaarxiv icon

Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model

Add code
Apr 23, 2026
Viaarxiv icon

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Survey

Add code
Jan 13, 2026
Viaarxiv icon

Efficient and Robust Video Defense Framework against 3D-field Personalized Talking Face

Add code
Dec 24, 2025
Viaarxiv icon

Is It Truly Necessary to Process and Fit Minutes-Long Reference Videos for Personalized Talking Face Generation?

Add code
Nov 11, 2025
Viaarxiv icon

PRO-V: An Efficient Program Generation Multi-Agent System for Automatic RTL Verification

Add code
Jun 13, 2025
Viaarxiv icon

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations

Add code
Jun 06, 2025
Viaarxiv icon