Picture for Cheng-Hau Yang

Cheng-Hau Yang

LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks

Add code
Aug 14, 2026
Viaarxiv icon

Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

Add code
Aug 14, 2026
Viaarxiv icon

Precise but Uncoupled: Reviewer Precision Does Not Guarantee Critique Uptake in Multi-Agent Math Reasoning

Add code
Jul 16, 2026
Viaarxiv icon

FlowBench: A Large Scale Benchmark for Flow Simulation over Complex Geometries

Add code
Sep 26, 2024
Figure 1 for FlowBench: A Large Scale Benchmark for Flow Simulation over Complex Geometries
Figure 2 for FlowBench: A Large Scale Benchmark for Flow Simulation over Complex Geometries
Figure 3 for FlowBench: A Large Scale Benchmark for Flow Simulation over Complex Geometries
Figure 4 for FlowBench: A Large Scale Benchmark for Flow Simulation over Complex Geometries
Viaarxiv icon