Picture for Qiao Jin

Qiao Jin

Do AI chatbots find what experts would? Effects of model, user role, and sample size on study retrieval for medical questions

Add code
Aug 13, 2026
Viaarxiv icon

Agents' Last Exam

Add code
Jun 03, 2026
Viaarxiv icon

Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models

Add code
May 19, 2026
Viaarxiv icon

Large Language Models Lack Temporal Awareness of Medical Knowledge

Add code
May 13, 2026
Viaarxiv icon

MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering

Add code
May 12, 2026
Viaarxiv icon

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

Add code
Mar 05, 2026
Viaarxiv icon

CT-Bench: A Benchmark for Multimodal Lesion Understanding in Computed Tomography

Add code
Feb 16, 2026
Viaarxiv icon

MedCite: Can Language Models Generate Verifiable Text for Medicine?

Add code
Jun 07, 2025
Viaarxiv icon

Knowledge-guided Contextual Gene Set Analysis Using Large Language Models

Add code
Jun 04, 2025
Figure 1 for Knowledge-guided Contextual Gene Set Analysis Using Large Language Models
Figure 2 for Knowledge-guided Contextual Gene Set Analysis Using Large Language Models
Figure 3 for Knowledge-guided Contextual Gene Set Analysis Using Large Language Models
Figure 4 for Knowledge-guided Contextual Gene Set Analysis Using Large Language Models
Viaarxiv icon

TrialPanorama: Database and Benchmark for Systematic Review and Design of Clinical Trials

Add code
May 22, 2025
Viaarxiv icon