Picture for Li Sheng

Li Sheng

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

Add code
Jul 30, 2026
Viaarxiv icon

CMSL: Constructive Multi-Sequence Learning for Recommendation Systems

Add code
Jun 26, 2026
Viaarxiv icon

Qwen-AgentWorld: Language World Models for General Agents

Add code
Jun 23, 2026
Viaarxiv icon

Post-Trained MoE Can Skip Half Experts via Self-Distillation

Add code
May 18, 2026
Viaarxiv icon

How Far Can Unsupervised RLVR Scale LLM Training?

Add code
Mar 09, 2026
Viaarxiv icon

P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads

Add code
Feb 10, 2026
Viaarxiv icon

P1: Mastering Physics Olympiads with Reinforcement Learning

Add code
Nov 17, 2025
Viaarxiv icon

TTRL: Test-Time Reinforcement Learning

Add code
Apr 22, 2025
Figure 1 for TTRL: Test-Time Reinforcement Learning
Figure 2 for TTRL: Test-Time Reinforcement Learning
Figure 3 for TTRL: Test-Time Reinforcement Learning
Figure 4 for TTRL: Test-Time Reinforcement Learning
Viaarxiv icon

UltraIF: Advancing Instruction Following from the Wild

Add code
Feb 06, 2025
Figure 1 for UltraIF: Advancing Instruction Following from the Wild
Figure 2 for UltraIF: Advancing Instruction Following from the Wild
Figure 3 for UltraIF: Advancing Instruction Following from the Wild
Figure 4 for UltraIF: Advancing Instruction Following from the Wild
Viaarxiv icon

Depression Detection on Social Media with Large Language Models

Add code
Mar 16, 2024
Figure 1 for Depression Detection on Social Media with Large Language Models
Figure 2 for Depression Detection on Social Media with Large Language Models
Figure 3 for Depression Detection on Social Media with Large Language Models
Figure 4 for Depression Detection on Social Media with Large Language Models
Viaarxiv icon