Picture for Vipul Gupta

Vipul Gupta

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

Add code
Jul 30, 2026
Viaarxiv icon

CRAFT: Clustering Rubrics to Diagnose Weak LLM Capabilities and Generate Targeted Fine-Tuning Data

Add code
Jul 17, 2026
Viaarxiv icon

Model Unlearning Objectives Vary for Distinct Language Functions

Add code
May 26, 2026
Viaarxiv icon

HARNESS-LM: A Three-Phase Training Recipe for Harnessing SLMs in Sponsored Search Retrieval

Add code
May 22, 2026
Viaarxiv icon

Physics-Informed Synthetic Dataset and Denoising TIE-Reconstructed Phase Maps in Transient Flows Using Deep Learning

Add code
Apr 12, 2026
Viaarxiv icon

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

Add code
Feb 05, 2026
Viaarxiv icon

PRBench: Large-Scale Expert Rubrics for Evaluating High-Stakes Professional Reasoning

Add code
Nov 14, 2025
Viaarxiv icon

HRScene: How Far Are VLMs from Effective High-Resolution Image Understanding?

Add code
Apr 29, 2025
Figure 1 for HRScene: How Far Are VLMs from Effective High-Resolution Image Understanding?
Figure 2 for HRScene: How Far Are VLMs from Effective High-Resolution Image Understanding?
Figure 3 for HRScene: How Far Are VLMs from Effective High-Resolution Image Understanding?
Figure 4 for HRScene: How Far Are VLMs from Effective High-Resolution Image Understanding?
Viaarxiv icon

Attention Pruning: Automated Fairness Repair of Language Models via Surrogate Simulated Annealing

Add code
Mar 20, 2025
Viaarxiv icon

Improving Consistency in Large Language Models through Chain of Guidance

Add code
Feb 21, 2025
Viaarxiv icon