Picture for Nitay Calderon

Nitay Calderon

LLM Explainability with Counterfactual Chains and Causal Graphs

Add code
Jun 04, 2026
Viaarxiv icon

A Matter of TASTE: Improving Coverage and Difficulty of Agent Benchmarks

Add code
May 27, 2026
Viaarxiv icon

Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality

Add code
Feb 15, 2026
Viaarxiv icon

LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals

Add code
Jan 15, 2026
Viaarxiv icon

Multi-Domain Explainability of Preferences

Add code
May 26, 2025
Figure 1 for Multi-Domain Explainability of Preferences
Figure 2 for Multi-Domain Explainability of Preferences
Figure 3 for Multi-Domain Explainability of Preferences
Figure 4 for Multi-Domain Explainability of Preferences
Viaarxiv icon

Dementia Through Different Eyes: Explainable Modeling of Human and LLM Perceptions for Early Awareness

Add code
May 19, 2025
Figure 1 for Dementia Through Different Eyes: Explainable Modeling of Human and LLM Perceptions for Early Awareness
Figure 2 for Dementia Through Different Eyes: Explainable Modeling of Human and LLM Perceptions for Early Awareness
Figure 3 for Dementia Through Different Eyes: Explainable Modeling of Human and LLM Perceptions for Early Awareness
Figure 4 for Dementia Through Different Eyes: Explainable Modeling of Human and LLM Perceptions for Early Awareness
Viaarxiv icon

AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation

Add code
Mar 25, 2025
Figure 1 for AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
Figure 2 for AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
Figure 3 for AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
Figure 4 for AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
Viaarxiv icon

The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs

Add code
Jan 19, 2025
Figure 1 for The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
Figure 2 for The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
Figure 3 for The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
Figure 4 for The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
Viaarxiv icon

Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance

Add code
Oct 24, 2024
Figure 1 for Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
Figure 2 for Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
Figure 3 for Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
Figure 4 for Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
Viaarxiv icon

NL-Eye: Abductive NLI for Images

Add code
Oct 03, 2024
Figure 1 for NL-Eye: Abductive NLI for Images
Figure 2 for NL-Eye: Abductive NLI for Images
Figure 3 for NL-Eye: Abductive NLI for Images
Figure 4 for NL-Eye: Abductive NLI for Images
Viaarxiv icon