Picture for Amrita Bhattacharjee

Amrita Bhattacharjee

ToolAlignBench: Investigating Alignment Conflicts in Tool-Calling Enabled LLMs

Add code
Jul 15, 2026
Viaarxiv icon

When Does Personality Composition Matter for Multi-Agent LLM Teams?

Add code
Jun 25, 2026
Viaarxiv icon

From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge

Add code
Nov 25, 2024
Figure 1 for From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
Figure 2 for From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
Figure 3 for From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
Figure 4 for From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
Viaarxiv icon

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering

Add code
Nov 19, 2024
Figure 1 for Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
Figure 2 for Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
Figure 3 for Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
Figure 4 for Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
Viaarxiv icon

Towards Inference-time Category-wise Safety Steering for Large Language Models

Add code
Oct 02, 2024
Figure 1 for Towards Inference-time Category-wise Safety Steering for Large Language Models
Figure 2 for Towards Inference-time Category-wise Safety Steering for Large Language Models
Figure 3 for Towards Inference-time Category-wise Safety Steering for Large Language Models
Figure 4 for Towards Inference-time Category-wise Safety Steering for Large Language Models
Viaarxiv icon

Defending Against Social Engineering Attacks in the Age of LLMs

Add code
Jun 18, 2024
Figure 1 for Defending Against Social Engineering Attacks in the Age of LLMs
Figure 2 for Defending Against Social Engineering Attacks in the Age of LLMs
Figure 3 for Defending Against Social Engineering Attacks in the Age of LLMs
Figure 4 for Defending Against Social Engineering Attacks in the Age of LLMs
Viaarxiv icon

Efficient Reinforcement Learning via Large Language Model-based Search

Add code
May 24, 2024
Figure 1 for Efficient Reinforcement Learning via Large Language Model-based Search
Figure 2 for Efficient Reinforcement Learning via Large Language Model-based Search
Figure 3 for Efficient Reinforcement Learning via Large Language Model-based Search
Figure 4 for Efficient Reinforcement Learning via Large Language Model-based Search
Viaarxiv icon

Zero-shot LLM-guided Counterfactual Generation for Text

Add code
May 08, 2024
Figure 1 for Zero-shot LLM-guided Counterfactual Generation for Text
Figure 2 for Zero-shot LLM-guided Counterfactual Generation for Text
Figure 3 for Zero-shot LLM-guided Counterfactual Generation for Text
Figure 4 for Zero-shot LLM-guided Counterfactual Generation for Text
Viaarxiv icon

EAGLE: A Domain Generalization Framework for AI-generated Text Detection

Add code
Mar 23, 2024
Viaarxiv icon

Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales

Add code
Mar 19, 2024
Figure 1 for Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
Figure 2 for Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
Figure 3 for Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
Figure 4 for Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
Viaarxiv icon