Picture for Huyang Sun

Huyang Sun

MemeMind: Reference-Guided Trace Construction for Offline Context Optimization

Add code
Aug 10, 2026
Viaarxiv icon

MemeBench: What LVLMs Miss When Interpreting Culture-Dependent Memes

Add code
Jul 30, 2026
Viaarxiv icon

Community-Aware Assessment of Social Textual Engagement and Resonance: A Human-Centric Perspective on User-Generated Content Evaluation

Add code
Jun 01, 2026
Viaarxiv icon

TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis

Add code
May 23, 2025
Figure 1 for TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis
Figure 2 for TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis
Figure 3 for TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis
Figure 4 for TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis
Viaarxiv icon

Aligning Anime Video Generation with Human Feedback

Add code
Apr 14, 2025
Figure 1 for Aligning Anime Video Generation with Human Feedback
Figure 2 for Aligning Anime Video Generation with Human Feedback
Figure 3 for Aligning Anime Video Generation with Human Feedback
Figure 4 for Aligning Anime Video Generation with Human Feedback
Viaarxiv icon

AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era

Add code
Dec 19, 2024
Figure 1 for AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era
Figure 2 for AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era
Figure 3 for AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era
Figure 4 for AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era
Viaarxiv icon

Exploring the Frontiers of Animation Video Generation in the Sora Era: Method, Dataset and Benchmark

Add code
Dec 13, 2024
Figure 1 for Exploring the Frontiers of Animation Video Generation in the Sora Era: Method, Dataset and Benchmark
Figure 2 for Exploring the Frontiers of Animation Video Generation in the Sora Era: Method, Dataset and Benchmark
Figure 3 for Exploring the Frontiers of Animation Video Generation in the Sora Era: Method, Dataset and Benchmark
Figure 4 for Exploring the Frontiers of Animation Video Generation in the Sora Era: Method, Dataset and Benchmark
Viaarxiv icon

DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training

Add code
Aug 01, 2024
Figure 1 for DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training
Figure 2 for DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training
Figure 3 for DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training
Figure 4 for DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training
Viaarxiv icon

Video Moment Retrieval from Text Queries via Single Frame Annotation

Add code
Apr 26, 2022
Figure 1 for Video Moment Retrieval from Text Queries via Single Frame Annotation
Figure 2 for Video Moment Retrieval from Text Queries via Single Frame Annotation
Figure 3 for Video Moment Retrieval from Text Queries via Single Frame Annotation
Figure 4 for Video Moment Retrieval from Text Queries via Single Frame Annotation
Viaarxiv icon

Boosting the Performance of Video Compression Artifact Reduction with Reference Frame Proposals and Frequency Domain Information

Add code
May 31, 2021
Figure 1 for Boosting the Performance of Video Compression Artifact Reduction with Reference Frame Proposals and Frequency Domain Information
Figure 2 for Boosting the Performance of Video Compression Artifact Reduction with Reference Frame Proposals and Frequency Domain Information
Figure 3 for Boosting the Performance of Video Compression Artifact Reduction with Reference Frame Proposals and Frequency Domain Information
Figure 4 for Boosting the Performance of Video Compression Artifact Reduction with Reference Frame Proposals and Frequency Domain Information
Viaarxiv icon