Picture for Shauli Ravfogel

Shauli Ravfogel

What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefs

Add code
Jun 26, 2026
Viaarxiv icon

Can LLM Agents Infer World Models? Evidence from Agentic Automata Learning

Add code
Jun 15, 2026
Viaarxiv icon

Can LLMs Introspect? A Reality Check

Add code
May 25, 2026
Viaarxiv icon

Geometric Factual Recall in Transformers

Add code
May 12, 2026
Viaarxiv icon

The Truthfulness Spectrum Hypothesis

Add code
Feb 23, 2026
Viaarxiv icon

Discrete Diffusion Models Exploit Asymmetry to Solve Lookahead Planning Tasks

Add code
Feb 23, 2026
Viaarxiv icon

From Directions to Regions: Decomposing Activations in Language Models via Local Geometry

Add code
Feb 02, 2026
Viaarxiv icon

State over Tokens: Characterizing the Role of Reasoning Tokens

Add code
Dec 14, 2025
Viaarxiv icon

IQ Test for LLMs: An Evaluation Framework for Uncovering Core Skills in LLMs

Add code
Jul 27, 2025
Viaarxiv icon

The Medium Is Not the Message: Deconfounding Text Embeddings via Linear Concept Erasure

Add code
Jul 01, 2025
Figure 1 for The Medium Is Not the Message: Deconfounding Text Embeddings via Linear Concept Erasure
Figure 2 for The Medium Is Not the Message: Deconfounding Text Embeddings via Linear Concept Erasure
Figure 3 for The Medium Is Not the Message: Deconfounding Text Embeddings via Linear Concept Erasure
Figure 4 for The Medium Is Not the Message: Deconfounding Text Embeddings via Linear Concept Erasure
Viaarxiv icon