Picture for Claudio Mayrink Verdun

Claudio Mayrink Verdun

KV Cache Compression Through the Lens of Transform Coding

Add code
Aug 14, 2026
Viaarxiv icon

An Instrument to Evaluate Governance Proposals: AI Policy Analysis at Scale

Add code
Jul 30, 2026
Viaarxiv icon

Reading Between the Dots: Decoding Hidden Computation across Filler Tokens

Add code
Jul 03, 2026
Viaarxiv icon

Inference-Time Reward Hacking in Large Language Models

Add code
Jun 24, 2025
Figure 1 for Inference-Time Reward Hacking in Large Language Models
Figure 2 for Inference-Time Reward Hacking in Large Language Models
Figure 3 for Inference-Time Reward Hacking in Large Language Models
Figure 4 for Inference-Time Reward Hacking in Large Language Models
Viaarxiv icon

HeavyWater and SimplexWater: Watermarking Low-Entropy Text Distributions

Add code
Jun 06, 2025
Figure 1 for HeavyWater and SimplexWater: Watermarking Low-Entropy Text Distributions
Figure 2 for HeavyWater and SimplexWater: Watermarking Low-Entropy Text Distributions
Figure 3 for HeavyWater and SimplexWater: Watermarking Low-Entropy Text Distributions
Figure 4 for HeavyWater and SimplexWater: Watermarking Low-Entropy Text Distributions
Viaarxiv icon

Multi-Group Proportional Representation for Text-to-Image Models

Add code
May 29, 2025
Figure 1 for Multi-Group Proportional Representation for Text-to-Image Models
Figure 2 for Multi-Group Proportional Representation for Text-to-Image Models
Figure 3 for Multi-Group Proportional Representation for Text-to-Image Models
Figure 4 for Multi-Group Proportional Representation for Text-to-Image Models
Viaarxiv icon

GradPCA: Leveraging NTK Alignment for Reliable Out-of-Distribution Detection

Add code
May 21, 2025
Figure 1 for GradPCA: Leveraging NTK Alignment for Reliable Out-of-Distribution Detection
Figure 2 for GradPCA: Leveraging NTK Alignment for Reliable Out-of-Distribution Detection
Figure 3 for GradPCA: Leveraging NTK Alignment for Reliable Out-of-Distribution Detection
Figure 4 for GradPCA: Leveraging NTK Alignment for Reliable Out-of-Distribution Detection
Viaarxiv icon

Optimized Couplings for Watermarking Large Language Models

Add code
May 13, 2025
Viaarxiv icon

Soft Best-of-n Sampling for Model Alignment

Add code
May 06, 2025
Figure 1 for Soft Best-of-n Sampling for Model Alignment
Viaarxiv icon

Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models

Add code
Jul 31, 2024
Figure 1 for Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models
Figure 2 for Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models
Figure 3 for Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models
Figure 4 for Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models
Viaarxiv icon