Picture for Thomas Breuel

Thomas Breuel

From Modalities to Propositions: A Language-Centric Framework for Multimodal Intelligence

Add code
Jul 18, 2026
Viaarxiv icon

Test-Time Coverage: Test-Conditioned Data Curation for Deployment-Aware Learning

Add code
Jul 18, 2026
Viaarxiv icon

A deeper look at depth pruning of LLMs

Add code
Jul 23, 2024
Figure 1 for A deeper look at depth pruning of LLMs
Figure 2 for A deeper look at depth pruning of LLMs
Figure 3 for A deeper look at depth pruning of LLMs
Figure 4 for A deeper look at depth pruning of LLMs
Viaarxiv icon

Investigating the Nature of 3D Generalization in Deep Neural Networks

Add code
Apr 19, 2023
Figure 1 for Investigating the Nature of 3D Generalization in Deep Neural Networks
Figure 2 for Investigating the Nature of 3D Generalization in Deep Neural Networks
Figure 3 for Investigating the Nature of 3D Generalization in Deep Neural Networks
Figure 4 for Investigating the Nature of 3D Generalization in Deep Neural Networks
Viaarxiv icon

GroupViT: Semantic Segmentation Emerges from Text Supervision

Add code
Feb 22, 2022
Figure 1 for GroupViT: Semantic Segmentation Emerges from Text Supervision
Figure 2 for GroupViT: Semantic Segmentation Emerges from Text Supervision
Figure 3 for GroupViT: Semantic Segmentation Emerges from Text Supervision
Figure 4 for GroupViT: Semantic Segmentation Emerges from Text Supervision
Viaarxiv icon

Identifying Layers Susceptible to Adversarial Attacks

Add code
Jul 10, 2021
Figure 1 for Identifying Layers Susceptible to Adversarial Attacks
Figure 2 for Identifying Layers Susceptible to Adversarial Attacks
Figure 3 for Identifying Layers Susceptible to Adversarial Attacks
Figure 4 for Identifying Layers Susceptible to Adversarial Attacks
Viaarxiv icon

Automatic Curation of Large-Scale Datasets for Audio-Visual Representation Learning

Add code
Jan 26, 2021
Figure 1 for Automatic Curation of Large-Scale Datasets for Audio-Visual Representation Learning
Figure 2 for Automatic Curation of Large-Scale Datasets for Audio-Visual Representation Learning
Figure 3 for Automatic Curation of Large-Scale Datasets for Audio-Visual Representation Learning
Figure 4 for Automatic Curation of Large-Scale Datasets for Audio-Visual Representation Learning
Viaarxiv icon

Parameter Efficient Multimodal Transformers for Video Representation Learning

Add code
Dec 08, 2020
Figure 1 for Parameter Efficient Multimodal Transformers for Video Representation Learning
Figure 2 for Parameter Efficient Multimodal Transformers for Video Representation Learning
Figure 3 for Parameter Efficient Multimodal Transformers for Video Representation Learning
Figure 4 for Parameter Efficient Multimodal Transformers for Video Representation Learning
Viaarxiv icon

Displacement-Invariant Cost Computation for Efficient Stereo Matching

Add code
Dec 01, 2020
Figure 1 for Displacement-Invariant Cost Computation for Efficient Stereo Matching
Figure 2 for Displacement-Invariant Cost Computation for Efficient Stereo Matching
Figure 3 for Displacement-Invariant Cost Computation for Efficient Stereo Matching
Figure 4 for Displacement-Invariant Cost Computation for Efficient Stereo Matching
Viaarxiv icon

Discovering Nonlinear Relations with Minimum Predictive Information Regularization

Add code
Jan 07, 2020
Figure 1 for Discovering Nonlinear Relations with Minimum Predictive Information Regularization
Figure 2 for Discovering Nonlinear Relations with Minimum Predictive Information Regularization
Figure 3 for Discovering Nonlinear Relations with Minimum Predictive Information Regularization
Viaarxiv icon