Picture for Vincent Siu

Vincent Siu

Agent Security Needs Redefinition through a Holistic Framework

Add code
Jul 24, 2026
Viaarxiv icon

Controlling Tool Use with Heading-Specific Activation Steering

Add code
Jul 07, 2026
Viaarxiv icon

Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

Add code
Jun 29, 2026
Viaarxiv icon

ChainWorld: Composing Long-Horizon Desktop Workloads from Atomic OSWorld Tasks

Add code
Jun 19, 2026
Viaarxiv icon

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

Add code
Jun 09, 2026
Viaarxiv icon

Agents' Last Exam

Add code
Jun 03, 2026
Viaarxiv icon

A Framework for Formalizing LLM Agent Security

Add code
Mar 19, 2026
Viaarxiv icon

RepIt: Representing Isolated Targets to Steer Language Models

Add code
Sep 16, 2025
Viaarxiv icon

AgentXploit: End-to-End Redteaming of Black-Box AI Agents

Add code
May 09, 2025
Figure 1 for AgentXploit: End-to-End Redteaming of Black-Box AI Agents
Figure 2 for AgentXploit: End-to-End Redteaming of Black-Box AI Agents
Figure 3 for AgentXploit: End-to-End Redteaming of Black-Box AI Agents
Figure 4 for AgentXploit: End-to-End Redteaming of Black-Box AI Agents
Viaarxiv icon