Picture for Jiyuan Zhang

Jiyuan Zhang

The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

Add code
Jul 23, 2026
Viaarxiv icon

SlimPer: Make Personalization Model Slim and Smart

Add code
Jul 14, 2026
Viaarxiv icon

Differentiable Efficient Operator Search

Add code
Jun 03, 2026
Viaarxiv icon

NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results

Add code
Apr 12, 2026
Viaarxiv icon

ERNIE 5.0 Technical Report

Add code
Feb 04, 2026
Viaarxiv icon

OmniSparse: Training-Aware Fine-Grained Sparse Attention for Long-Video MLLMs

Add code
Nov 18, 2025
Viaarxiv icon

X-Driver: Explainable Autonomous Driving with Vision-Language Models

Add code
May 08, 2025
Figure 1 for X-Driver: Explainable Autonomous Driving with Vision-Language Models
Figure 2 for X-Driver: Explainable Autonomous Driving with Vision-Language Models
Figure 3 for X-Driver: Explainable Autonomous Driving with Vision-Language Models
Figure 4 for X-Driver: Explainable Autonomous Driving with Vision-Language Models
Viaarxiv icon

Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding

Add code
Jan 28, 2025
Figure 1 for Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
Figure 2 for Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
Figure 3 for Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
Figure 4 for Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
Viaarxiv icon

Automatic Item Generation for Personality Situational Judgment Tests with Large Language Models

Add code
Dec 10, 2024
Viaarxiv icon

Evaluating and Advancing Multimodal Large Language Models in Ability Lens

Add code
Nov 22, 2024
Viaarxiv icon