Picture for Xiaofei Ding

Xiaofei Ding

Improving Audio-Visual Speech Recognition by Lip-Subword Correlation Based Visual Pre-training and Cross-Modal Fusion Encoder

Add code
Aug 14, 2023
Viaarxiv icon