Picture for Toru Tamaki

Toru Tamaki

The First EgoCross Challenge at EgoVis 2026: Cross-Domain Egocentric Video Question Answering

Add code
Aug 05, 2026
Viaarxiv icon

Reflective Dialogue between Teacher and Solver Agents for Video Question Answering

Add code
May 27, 2026
Viaarxiv icon

BFMD: A Full-Match Badminton Dense Dataset for Dense Shot Captioning

Add code
Mar 26, 2026
Viaarxiv icon

M3DDM+: An improved video outpainting by a modified masking strategy

Add code
Jan 16, 2026
Viaarxiv icon

Action tube generation by person query matching for spatio-temporal action detection

Add code
Mar 17, 2025
Figure 1 for Action tube generation by person query matching for spatio-temporal action detection
Figure 2 for Action tube generation by person query matching for spatio-temporal action detection
Figure 3 for Action tube generation by person query matching for spatio-temporal action detection
Figure 4 for Action tube generation by person query matching for spatio-temporal action detection
Viaarxiv icon

Shift and matching queries for video semantic segmentation

Add code
Oct 10, 2024
Figure 1 for Shift and matching queries for video semantic segmentation
Figure 2 for Shift and matching queries for video semantic segmentation
Figure 3 for Shift and matching queries for video semantic segmentation
Figure 4 for Shift and matching queries for video semantic segmentation
Viaarxiv icon

Query matching for spatio-temporal action detection with query-based object detector

Add code
Sep 27, 2024
Figure 1 for Query matching for spatio-temporal action detection with query-based object detector
Figure 2 for Query matching for spatio-temporal action detection with query-based object detector
Figure 3 for Query matching for spatio-temporal action detection with query-based object detector
Figure 4 for Query matching for spatio-temporal action detection with query-based object detector
Viaarxiv icon

Online pre-training with long-form videos

Add code
Aug 28, 2024
Figure 1 for Online pre-training with long-form videos
Viaarxiv icon

Fine-grained length controllable video captioning with ordinal embeddings

Add code
Aug 27, 2024
Figure 1 for Fine-grained length controllable video captioning with ordinal embeddings
Figure 2 for Fine-grained length controllable video captioning with ordinal embeddings
Figure 3 for Fine-grained length controllable video captioning with ordinal embeddings
Figure 4 for Fine-grained length controllable video captioning with ordinal embeddings
Viaarxiv icon

Multi-model learning by sequential reading of untrimmed videos for action recognition

Add code
Jan 26, 2024
Viaarxiv icon