Picture for Zhaofeng Sun

Zhaofeng Sun

Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUs

Add code
Oct 21, 2024
Viaarxiv icon