Picture for Ziyu Yao

Ziyu Yao

George Mason University

Evaluating the Effectiveness of Persona Simulation in Opinion Prediction with GPT-4.1

Add code
Jul 22, 2026
Viaarxiv icon

Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability?

Add code
Jun 23, 2026
Viaarxiv icon

PeerMathDial: A Middle School Dialogue Dataset for Student Collaborative Math Problem Solving

Add code
Jun 19, 2026
Viaarxiv icon

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings

Add code
Apr 09, 2026
Viaarxiv icon

Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective

Add code
Mar 15, 2026
Viaarxiv icon

ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning

Add code
Oct 09, 2025
Figure 1 for ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
Figure 2 for ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
Figure 3 for ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
Figure 4 for ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
Viaarxiv icon

All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens

Add code
Sep 11, 2025
Viaarxiv icon

Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models

Add code
May 21, 2025
Viaarxiv icon

Revisiting Prompt Optimization with Large Reasoning Models-A Case Study on Event Extraction

Add code
Apr 10, 2025
Viaarxiv icon

Can LLMs Simulate Personas with Reversed Performance? A Benchmark for Counterfactual Instruction Following

Add code
Apr 08, 2025
Figure 1 for Can LLMs Simulate Personas with Reversed Performance? A Benchmark for Counterfactual Instruction Following
Figure 2 for Can LLMs Simulate Personas with Reversed Performance? A Benchmark for Counterfactual Instruction Following
Figure 3 for Can LLMs Simulate Personas with Reversed Performance? A Benchmark for Counterfactual Instruction Following
Figure 4 for Can LLMs Simulate Personas with Reversed Performance? A Benchmark for Counterfactual Instruction Following
Viaarxiv icon