Picture for Jan Kautz

Jan Kautz

NVIDIA

LLM Pruning and Distillation in Practice: The Minitron Approach

Add code
Aug 21, 2024
Figure 1 for LLM Pruning and Distillation in Practice: The Minitron Approach
Figure 2 for LLM Pruning and Distillation in Practice: The Minitron Approach
Figure 3 for LLM Pruning and Distillation in Practice: The Minitron Approach
Figure 4 for LLM Pruning and Distillation in Practice: The Minitron Approach
Viaarxiv icon

A deeper look at depth pruning of LLMs

Add code
Jul 23, 2024
Figure 1 for A deeper look at depth pruning of LLMs
Figure 2 for A deeper look at depth pruning of LLMs
Figure 3 for A deeper look at depth pruning of LLMs
Figure 4 for A deeper look at depth pruning of LLMs
Viaarxiv icon

Compact Language Models via Pruning and Knowledge Distillation

Add code
Jul 19, 2024
Figure 1 for Compact Language Models via Pruning and Knowledge Distillation
Figure 2 for Compact Language Models via Pruning and Knowledge Distillation
Figure 3 for Compact Language Models via Pruning and Knowledge Distillation
Figure 4 for Compact Language Models via Pruning and Knowledge Distillation
Viaarxiv icon

MambaVision: A Hybrid Mamba-Transformer Vision Backbone

Add code
Jul 10, 2024
Viaarxiv icon

An Empirical Study of Mamba-based Language Models

Add code
Jun 12, 2024
Figure 1 for An Empirical Study of Mamba-based Language Models
Figure 2 for An Empirical Study of Mamba-based Language Models
Figure 3 for An Empirical Study of Mamba-based Language Models
Figure 4 for An Empirical Study of Mamba-based Language Models
Viaarxiv icon

Flextron: Many-in-One Flexible Large Language Model

Add code
Jun 11, 2024
Figure 1 for Flextron: Many-in-One Flexible Large Language Model
Figure 2 for Flextron: Many-in-One Flexible Large Language Model
Figure 3 for Flextron: Many-in-One Flexible Large Language Model
Figure 4 for Flextron: Many-in-One Flexible Large Language Model
Viaarxiv icon

Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

Add code
Jun 11, 2024
Figure 1 for Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation
Figure 2 for Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation
Figure 3 for Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation
Figure 4 for Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation
Viaarxiv icon

CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation

Add code
Jun 04, 2024
Figure 1 for CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
Figure 2 for CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
Figure 3 for CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
Figure 4 for CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
Viaarxiv icon

SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model

Add code
Jun 03, 2024
Figure 1 for SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model
Figure 2 for SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model
Figure 3 for SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model
Figure 4 for SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model
Viaarxiv icon

X-VILA: Cross-Modality Alignment for Large Language Model

Add code
May 29, 2024
Figure 1 for X-VILA: Cross-Modality Alignment for Large Language Model
Figure 2 for X-VILA: Cross-Modality Alignment for Large Language Model
Figure 3 for X-VILA: Cross-Modality Alignment for Large Language Model
Figure 4 for X-VILA: Cross-Modality Alignment for Large Language Model
Viaarxiv icon