Picture for Yuqing Du

Yuqing Du

University of British Columbia

Intrinsically-Motivated Humans and Agents in Open-World Exploration

Add code
Mar 31, 2025
Viaarxiv icon

Imagen 3

Add code
Aug 13, 2024
Viaarxiv icon

Semi-Supervised One-Shot Imitation Learning

Add code
Aug 09, 2024
Figure 1 for Semi-Supervised One-Shot Imitation Learning
Figure 2 for Semi-Supervised One-Shot Imitation Learning
Figure 3 for Semi-Supervised One-Shot Imitation Learning
Figure 4 for Semi-Supervised One-Shot Imitation Learning
Viaarxiv icon

Teaching Large Language Models to Reason with Reinforcement Learning

Add code
Mar 07, 2024
Figure 1 for Teaching Large Language Models to Reason with Reinforcement Learning
Figure 2 for Teaching Large Language Models to Reason with Reinforcement Learning
Figure 3 for Teaching Large Language Models to Reason with Reinforcement Learning
Figure 4 for Teaching Large Language Models to Reason with Reinforcement Learning
Viaarxiv icon

Learning to Model the World with Language

Add code
Jul 31, 2023
Figure 1 for Learning to Model the World with Language
Figure 2 for Learning to Model the World with Language
Figure 3 for Learning to Model the World with Language
Figure 4 for Learning to Model the World with Language
Viaarxiv icon

DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models

Add code
May 25, 2023
Figure 1 for DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models
Figure 2 for DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models
Figure 3 for DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models
Figure 4 for DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models
Viaarxiv icon

Vision-Language Models as Success Detectors

Add code
Mar 13, 2023
Viaarxiv icon

Aligning Text-to-Image Models using Human Feedback

Add code
Feb 23, 2023
Viaarxiv icon

Guiding Pretraining in Reinforcement Learning with Large Language Models

Add code
Feb 13, 2023
Figure 1 for Guiding Pretraining in Reinforcement Learning with Large Language Models
Figure 2 for Guiding Pretraining in Reinforcement Learning with Large Language Models
Figure 3 for Guiding Pretraining in Reinforcement Learning with Large Language Models
Figure 4 for Guiding Pretraining in Reinforcement Learning with Large Language Models
Viaarxiv icon

It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation

Add code
Feb 22, 2022
Figure 1 for It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation
Figure 2 for It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation
Figure 3 for It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation
Figure 4 for It Takes Four to Tango: Multiagent Selfplay for Automatic Curriculum Generation
Viaarxiv icon