Picture for Zhenfeng Su

Zhenfeng Su

Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUs

Add code
Oct 21, 2024
Viaarxiv icon