Picture for Lin Tan

Lin Tan

VICBench: A Multi-Language Benchmark for Code Vulnerability Detection

Add code
Aug 12, 2026
Viaarxiv icon

Learning Context-Free Grammars for Grammar-Constrained Decoding via Declarative Agentic Programming with Guarantees

Add code
Aug 06, 2026
Viaarxiv icon

An Empirical Study of LLM-Generated Specifications for VeriFast

Add code
Jun 25, 2026
Viaarxiv icon

Towards Resiliency in Large Language Model Serving with KevlarFlow

Add code
Jan 30, 2026
Viaarxiv icon

Unified Software Engineering agent as AI Software Engineer

Add code
Jun 17, 2025
Viaarxiv icon

Leveraging Interview-Informed LLMs to Model Survey Responses: Comparative Insights from AI-Generated and Human Data

Add code
May 28, 2025
Figure 1 for Leveraging Interview-Informed LLMs to Model Survey Responses: Comparative Insights from AI-Generated and Human Data
Figure 2 for Leveraging Interview-Informed LLMs to Model Survey Responses: Comparative Insights from AI-Generated and Human Data
Figure 3 for Leveraging Interview-Informed LLMs to Model Survey Responses: Comparative Insights from AI-Generated and Human Data
Figure 4 for Leveraging Interview-Informed LLMs to Model Survey Responses: Comparative Insights from AI-Generated and Human Data
Viaarxiv icon

Can Language Models Replace Programmers? REPOCOD Says 'Not Yet'

Add code
Oct 29, 2024
Viaarxiv icon

WAFFLE: Multi-Modal Model for Automated Front-End Development

Add code
Oct 24, 2024
Figure 1 for WAFFLE: Multi-Modal Model for Automated Front-End Development
Figure 2 for WAFFLE: Multi-Modal Model for Automated Front-End Development
Figure 3 for WAFFLE: Multi-Modal Model for Automated Front-End Development
Figure 4 for WAFFLE: Multi-Modal Model for Automated Front-End Development
Viaarxiv icon

Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code

Add code
Oct 13, 2024
Figure 1 for Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code
Figure 2 for Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code
Figure 3 for Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code
Figure 4 for Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code
Viaarxiv icon

SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models

Add code
Sep 28, 2024
Figure 1 for SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
Figure 2 for SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
Figure 3 for SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
Figure 4 for SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
Viaarxiv icon