Picture for Hewei Guo

Hewei Guo

Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging

Add code
Jul 11, 2026
Viaarxiv icon

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy

Add code
Jun 26, 2026
Viaarxiv icon

SenseNova-MARS: Empowering Multimodal Agentic Reasoning and Search via Reinforcement Learning

Add code
Dec 30, 2025
Viaarxiv icon

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Add code
Apr 29, 2024
Figure 1 for How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Figure 2 for How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Figure 3 for How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Figure 4 for How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Viaarxiv icon

Unsupervised Domain Adaptation GAN Inversion for Image Editing

Add code
Nov 22, 2022
Figure 1 for Unsupervised Domain Adaptation GAN Inversion for Image Editing
Figure 2 for Unsupervised Domain Adaptation GAN Inversion for Image Editing
Figure 3 for Unsupervised Domain Adaptation GAN Inversion for Image Editing
Figure 4 for Unsupervised Domain Adaptation GAN Inversion for Image Editing
Viaarxiv icon