Picture for Jiaqing Liu

Jiaqing Liu

MinMo: A Multimodal Large Language Model for Seamless Voice Interaction

Add code
Jan 10, 2025
Viaarxiv icon

OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation

Add code
Oct 23, 2024
Figure 1 for OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
Figure 2 for OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
Figure 3 for OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
Figure 4 for OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
Viaarxiv icon

Recording for Eyes, Not Echoing to Ears: Contextualized Spoken-to-Written Conversion of ASR Transcripts

Add code
Aug 19, 2024
Viaarxiv icon

Multimodal Fusion and Coherence Modeling for Video Topic Segmentation

Add code
Aug 01, 2024
Viaarxiv icon

Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers

Add code
Jun 17, 2024
Figure 1 for Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers
Figure 2 for Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers
Figure 3 for Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers
Figure 4 for Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers
Viaarxiv icon

Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token Based ASR

Add code
Nov 08, 2023
Viaarxiv icon

Improving Long Document Topic Segmentation Models With Enhanced Coherence Modeling

Add code
Oct 23, 2023
Viaarxiv icon

Ladder Fine-tuning approach for SAM integrating complementary network

Add code
Jun 22, 2023
Figure 1 for Ladder Fine-tuning approach for SAM integrating complementary network
Figure 2 for Ladder Fine-tuning approach for SAM integrating complementary network
Figure 3 for Ladder Fine-tuning approach for SAM integrating complementary network
Figure 4 for Ladder Fine-tuning approach for SAM integrating complementary network
Viaarxiv icon

Ditto: A Simple and Efficient Approach to Improve Sentence Embeddings

Add code
May 18, 2023
Viaarxiv icon

MUG: A General Meeting Understanding and Generation Benchmark

Add code
Mar 27, 2023
Viaarxiv icon