Picture for Shaokun Zhang

Shaokun Zhang

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning

Add code
Jul 22, 2026
Viaarxiv icon

An Empirical Study for Android-to-OpenHarmony GUI Test Migration

Add code
Jul 14, 2026
Viaarxiv icon

Who&When Pro: Can LLMs Really Attribute Failures in AI Agents?

Add code
Jul 10, 2026
Viaarxiv icon

ProCUA-SFT Technical Report

Add code
Jun 15, 2026
Viaarxiv icon

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

Add code
Apr 27, 2026
Viaarxiv icon

ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents

Add code
Mar 19, 2026
Viaarxiv icon

Golden Goose: A Simple Trick to Synthesize Unlimited RLVR Tasks from Unverifiable Internet Text

Add code
Jan 30, 2026
Viaarxiv icon

NVIDIA Nemotron Nano V2 VL

Add code
Nov 07, 2025
Viaarxiv icon

A Survey of Self-Evolving Agents: On Path to Artificial Super Intelligence

Add code
Jul 28, 2025
Figure 1 for A Survey of Self-Evolving Agents: On Path to Artificial Super Intelligence
Figure 2 for A Survey of Self-Evolving Agents: On Path to Artificial Super Intelligence
Figure 3 for A Survey of Self-Evolving Agents: On Path to Artificial Super Intelligence
Figure 4 for A Survey of Self-Evolving Agents: On Path to Artificial Super Intelligence
Viaarxiv icon

Divide, Optimize, Merge: Fine-Grained LLM Agent Optimization at Scale

Add code
May 06, 2025
Viaarxiv icon