Picture for Dilek Hakkani-Tur

Dilek Hakkani-Tur

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

Add code
Jul 28, 2026
Viaarxiv icon

IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations

Add code
Jun 26, 2026
Viaarxiv icon

ReasoningFlow: Discourse Structures for Understanding LLM Reasoning Traces

Add code
Jun 03, 2026
Viaarxiv icon

MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models

Add code
May 27, 2026
Viaarxiv icon

Too Polite to Disagree: Understanding Sycophancy Propagation in Multi-Agent Systems

Add code
Apr 03, 2026
Viaarxiv icon

Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration

Add code
Mar 12, 2026
Viaarxiv icon

SIMU: Selective Influence Machine Unlearning

Add code
Oct 09, 2025
Viaarxiv icon

Question Generation for Assessing Early Literacy Reading Comprehension

Add code
Jul 30, 2025
Figure 1 for Question Generation for Assessing Early Literacy Reading Comprehension
Figure 2 for Question Generation for Assessing Early Literacy Reading Comprehension
Viaarxiv icon

Reinforcement Learning Finetunes Small Subnetworks in Large Language Models

Add code
May 16, 2025
Figure 1 for Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
Figure 2 for Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
Figure 3 for Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
Figure 4 for Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
Viaarxiv icon

Spark: A System for Scientifically Creative Idea Generation

Add code
Apr 25, 2025
Figure 1 for Spark: A System for Scientifically Creative Idea Generation
Figure 2 for Spark: A System for Scientifically Creative Idea Generation
Figure 3 for Spark: A System for Scientifically Creative Idea Generation
Viaarxiv icon