Picture for Daan de Geus

Daan de Geus

Towards Metric-Agnostic Trajectory Forecasting

Add code
Jul 01, 2026
Viaarxiv icon

Surgical Anatomy Recognition with Context Learning using Foundation Representations

Add code
Jun 20, 2026
Viaarxiv icon

SurGe: Improved Surface Geometry in Point Maps

Add code
May 29, 2026
Viaarxiv icon

Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding

Add code
Apr 21, 2026
Viaarxiv icon

A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens

Add code
Apr 06, 2026
Viaarxiv icon

PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders

Add code
Mar 26, 2026
Viaarxiv icon

How Important are Videos for Training Video LLMs?

Add code
Jun 07, 2025
Figure 1 for How Important are Videos for Training Video LLMs?
Figure 2 for How Important are Videos for Training Video LLMs?
Figure 3 for How Important are Videos for Training Video LLMs?
Figure 4 for How Important are Videos for Training Video LLMs?
Viaarxiv icon

DONUT: A Decoder-Only Model for Trajectory Prediction

Add code
Jun 07, 2025
Figure 1 for DONUT: A Decoder-Only Model for Trajectory Prediction
Figure 2 for DONUT: A Decoder-Only Model for Trajectory Prediction
Figure 3 for DONUT: A Decoder-Only Model for Trajectory Prediction
Figure 4 for DONUT: A Decoder-Only Model for Trajectory Prediction
Viaarxiv icon

DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation

Add code
Mar 24, 2025
Figure 1 for DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
Figure 2 for DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
Figure 3 for DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
Figure 4 for DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
Viaarxiv icon

Your ViT is Secretly an Image Segmentation Model

Add code
Mar 24, 2025
Figure 1 for Your ViT is Secretly an Image Segmentation Model
Figure 2 for Your ViT is Secretly an Image Segmentation Model
Figure 3 for Your ViT is Secretly an Image Segmentation Model
Figure 4 for Your ViT is Secretly an Image Segmentation Model
Viaarxiv icon