Picture for Sixian Li

Sixian Li

ETA: A New Agentic Paradigm for Embodied Tasks

Add code
Aug 04, 2026
Viaarxiv icon

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias

Add code
Jul 13, 2026
Viaarxiv icon

CoRE-VLA: Towards Scalable and Robust Vision-Language-Action Modeling via Conditional Routing of Experts

Add code
Jul 04, 2026
Viaarxiv icon

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models

Add code
Jun 05, 2026
Viaarxiv icon

DFPO: Scaling Value Modeling via Distributional Flow towards Robust and Generalizable LLM Post-Training

Add code
Feb 05, 2026
Viaarxiv icon

FRoM-W1: Towards General Humanoid Whole-Body Control with Language Instructions

Add code
Jan 19, 2026
Viaarxiv icon

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models

Add code
May 29, 2025
Figure 1 for SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models
Figure 2 for SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models
Figure 3 for SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models
Figure 4 for SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models
Viaarxiv icon

TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use

Add code
Dec 20, 2024
Figure 1 for TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use
Figure 2 for TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use
Figure 3 for TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use
Figure 4 for TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use
Viaarxiv icon

SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance

Add code
Jun 26, 2024
Figure 1 for SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance
Figure 2 for SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance
Figure 3 for SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance
Figure 4 for SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance
Viaarxiv icon

Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution

Add code
Feb 27, 2024
Figure 1 for Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution
Figure 2 for Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution
Figure 3 for Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution
Figure 4 for Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution
Viaarxiv icon