Picture for Jan Kautz

Jan Kautz

NVIDIA

StreamChat: Chatting with Streaming Video

Add code
Dec 11, 2024
Viaarxiv icon

RADIO Amplified: Improved Baselines for Agglomerative Vision Foundation Models

Add code
Dec 10, 2024
Viaarxiv icon

Gated Delta Networks: Improving Mamba2 with Delta Rule

Add code
Dec 09, 2024
Viaarxiv icon

NVILA: Efficient Frontier Visual Language Models

Add code
Dec 05, 2024
Figure 1 for NVILA: Efficient Frontier Visual Language Models
Figure 2 for NVILA: Efficient Frontier Visual Language Models
Figure 3 for NVILA: Efficient Frontier Visual Language Models
Figure 4 for NVILA: Efficient Frontier Visual Language Models
Viaarxiv icon

NaVILA: Legged Robot Vision-Language-Action Model for Navigation

Add code
Dec 05, 2024
Viaarxiv icon

Hymba: A Hybrid-head Architecture for Small Language Models

Add code
Nov 20, 2024
Figure 1 for Hymba: A Hybrid-head Architecture for Small Language Models
Figure 2 for Hymba: A Hybrid-head Architecture for Small Language Models
Figure 3 for Hymba: A Hybrid-head Architecture for Small Language Models
Figure 4 for Hymba: A Hybrid-head Architecture for Small Language Models
Viaarxiv icon

EoRA: Training-free Compensation for Compressed LLM with Eigenspace Low-Rank Approximation

Add code
Oct 28, 2024
Viaarxiv icon

HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots

Add code
Oct 28, 2024
Figure 1 for HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots
Figure 2 for HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots
Figure 3 for HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots
Figure 4 for HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots
Viaarxiv icon

nvTorchCam: An Open-source Library for Camera-Agnostic Differentiable Geometric Vision

Add code
Oct 15, 2024
Figure 1 for nvTorchCam: An Open-source Library for Camera-Agnostic Differentiable Geometric Vision
Figure 2 for nvTorchCam: An Open-source Library for Camera-Agnostic Differentiable Geometric Vision
Figure 3 for nvTorchCam: An Open-source Library for Camera-Agnostic Differentiable Geometric Vision
Figure 4 for nvTorchCam: An Open-source Library for Camera-Agnostic Differentiable Geometric Vision
Viaarxiv icon

Exploring the design space of deep-learning-based weather forecasting systems

Add code
Oct 09, 2024
Figure 1 for Exploring the design space of deep-learning-based weather forecasting systems
Figure 2 for Exploring the design space of deep-learning-based weather forecasting systems
Figure 3 for Exploring the design space of deep-learning-based weather forecasting systems
Figure 4 for Exploring the design space of deep-learning-based weather forecasting systems
Viaarxiv icon