Picture for Han Qiu

Han Qiu

The Model's Tell: Measuring Context-Leakage Attack Signals with Behavior Gauges

Add code
Aug 18, 2026
Viaarxiv icon

Can Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora?

Add code
Aug 11, 2026
Viaarxiv icon

Auditing Chinese Web-scale Corpora via Sampled BPE Token Statistics

Add code
Aug 11, 2026
Viaarxiv icon

Leak-Resistant Unlearning: A New Benchmark for Evaluating Multi-Hop Reasoning Consistency and Recovery Robustness

Add code
Aug 05, 2026
Viaarxiv icon

Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

Add code
Aug 01, 2026
Viaarxiv icon

EvoVid: Temporal-Centric Self-Evolution for Video Large Language Models

Add code
May 21, 2026
Viaarxiv icon

LeakDojo: Decoding the Leakage Threats of RAG Systems

Add code
May 07, 2026
Viaarxiv icon

LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety

Add code
Apr 13, 2026
Viaarxiv icon

State-Dependent Safety Failures in Multi-Turn Language Model Interaction

Add code
Mar 15, 2026
Viaarxiv icon

Survive at All Costs: Exploring LLM's Risky Behaviors under Survival Pressure

Add code
Mar 05, 2026
Viaarxiv icon