Picture for Han Qi

Han Qi

Harvard University

Fair ASR: Re-Evaluating Black-Box Jailbreaks under Shared Target-Call Budgets

Add code
Aug 18, 2026
Viaarxiv icon

JailbreakSkill: Scaling Automated Red-Teaming with Reusable and Ever-Evolving Skills

Add code
Aug 17, 2026
Viaarxiv icon

X4Val: Learning Neural Surrogates for Variance-Reduced Policy Evaluation

Add code
Jun 03, 2026
Viaarxiv icon

Agentic Trading: When LLM Agents Meet Financial Markets

Add code
May 19, 2026
Viaarxiv icon

MAGIC: A Co-Evolving Attacker-Defender Adversarial Game for Robust LLM Safety

Add code
Feb 02, 2026
Viaarxiv icon

Compose by Focus: Scene Graph-based Atomic Skills

Add code
Sep 19, 2025
Figure 1 for Compose by Focus: Scene Graph-based Atomic Skills
Figure 2 for Compose by Focus: Scene Graph-based Atomic Skills
Figure 3 for Compose by Focus: Scene Graph-based Atomic Skills
Figure 4 for Compose by Focus: Scene Graph-based Atomic Skills
Viaarxiv icon

MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook

Add code
Sep 17, 2025
Figure 1 for MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
Figure 2 for MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
Figure 3 for MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
Figure 4 for MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
Viaarxiv icon

SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law

Add code
Jul 24, 2025
Figure 1 for SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law
Figure 2 for SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law
Figure 3 for SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law
Figure 4 for SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law
Viaarxiv icon

Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute

Add code
Apr 02, 2025
Figure 1 for Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
Figure 2 for Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
Figure 3 for Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
Figure 4 for Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
Viaarxiv icon

Graph Feedback Bandits on Similar Arms: With and Without Graph Structures

Add code
Jan 24, 2025
Figure 1 for Graph Feedback Bandits on Similar Arms: With and Without Graph Structures
Figure 2 for Graph Feedback Bandits on Similar Arms: With and Without Graph Structures
Figure 3 for Graph Feedback Bandits on Similar Arms: With and Without Graph Structures
Figure 4 for Graph Feedback Bandits on Similar Arms: With and Without Graph Structures
Viaarxiv icon