Picture for Zhiwu Lu

Zhiwu Lu

MMRole: A Comprehensive Framework for Developing and Evaluating Multimodal Role-Playing Agents

Add code
Aug 08, 2024
Viaarxiv icon

CoTBal: Comprehensive Task Balancing for Multi-Task Visual Instruction Tuning

Add code
Mar 07, 2024
Viaarxiv icon

Improvable Gap Balancing for Multi-Task Learning

Add code
Jul 28, 2023
Viaarxiv icon

VDT: An Empirical Study on Video Diffusion with Transformers

Add code
May 22, 2023
Viaarxiv icon

UniAdapter: Unified Parameter-Efficient Transfer Learning for Cross-modal Modeling

Add code
Feb 13, 2023
Viaarxiv icon

TikTalk: A Multi-Modal Dialogue Dataset for Real-World Chitchat

Add code
Jan 14, 2023
Viaarxiv icon

Text2Poster: Laying out Stylized Texts on Retrieved Images

Add code
Jan 06, 2023
Viaarxiv icon

LGDN: Language-Guided Denoising Network for Video-Language Modeling

Add code
Oct 03, 2022
Figure 1 for LGDN: Language-Guided Denoising Network for Video-Language Modeling
Figure 2 for LGDN: Language-Guided Denoising Network for Video-Language Modeling
Figure 3 for LGDN: Language-Guided Denoising Network for Video-Language Modeling
Figure 4 for LGDN: Language-Guided Denoising Network for Video-Language Modeling
Viaarxiv icon

A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language

Add code
Sep 12, 2022
Figure 1 for A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language
Figure 2 for A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language
Figure 3 for A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language
Figure 4 for A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language
Viaarxiv icon

Multimodal foundation models are better simulators of the human brain

Add code
Aug 17, 2022
Figure 1 for Multimodal foundation models are better simulators of the human brain
Figure 2 for Multimodal foundation models are better simulators of the human brain
Figure 3 for Multimodal foundation models are better simulators of the human brain
Figure 4 for Multimodal foundation models are better simulators of the human brain
Viaarxiv icon