Picture for Geon-Hyeong Kim

Geon-Hyeong Kim

A Regret Minimization Framework on Preference Learning in Large Language Models

Add code
Jun 08, 2026
Viaarxiv icon

SafeDPO: A Simple Approach to Direct Preference Optimization with Enhanced Safety

Add code
May 26, 2025
Viaarxiv icon

LobsDICE: Offline Imitation Learning from Observation via Stationary Distribution Correction Estimation

Add code
Feb 28, 2022
Figure 1 for LobsDICE: Offline Imitation Learning from Observation via Stationary Distribution Correction Estimation
Figure 2 for LobsDICE: Offline Imitation Learning from Observation via Stationary Distribution Correction Estimation
Figure 3 for LobsDICE: Offline Imitation Learning from Observation via Stationary Distribution Correction Estimation
Viaarxiv icon

Variational Interaction Information Maximization for Cross-domain Disentanglement

Add code
Dec 08, 2020
Figure 1 for Variational Interaction Information Maximization for Cross-domain Disentanglement
Figure 2 for Variational Interaction Information Maximization for Cross-domain Disentanglement
Figure 3 for Variational Interaction Information Maximization for Cross-domain Disentanglement
Figure 4 for Variational Interaction Information Maximization for Cross-domain Disentanglement
Viaarxiv icon