Picture for Yunbo Lyu

Yunbo Lyu

Second Thought: Reasoning in Parallel as LLM Agents Act and Observe

Add code
Aug 13, 2026
Viaarxiv icon

Fail-Fast, Restart-Smart: Early Failure Prediction and Restart for SWE Agentic Tasks

Add code
Aug 04, 2026
Viaarxiv icon

How Do Practitioners Build SE Agents? Insights from a Mixed-Methods Study

Add code
Jul 12, 2026
Viaarxiv icon

SecureAgentBench: Benchmarking Secure Code Generation under Realistic Vulnerability Scenarios

Add code
Sep 26, 2025
Viaarxiv icon

"My productivity is boosted, but ..." Demystifying Users' Perception on AI Coding Assistants

Add code
Aug 17, 2025
Viaarxiv icon

LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks

Add code
Feb 10, 2025
Figure 1 for LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks
Figure 2 for LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks
Figure 3 for LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks
Figure 4 for LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks
Viaarxiv icon

Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?

Add code
Jan 27, 2025
Figure 1 for Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
Figure 2 for Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
Figure 3 for Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
Figure 4 for Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
Viaarxiv icon