Picture for Jing Bi

Jing Bi

Generative AI for Cel-Animation: A Survey

Add code
Jan 08, 2025
Viaarxiv icon

Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach

Add code
Dec 24, 2024
Viaarxiv icon

Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning

Add code
Dec 13, 2024
Figure 1 for Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning
Figure 2 for Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning
Figure 3 for Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning
Figure 4 for Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning
Viaarxiv icon

VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?

Add code
Nov 19, 2024
Figure 1 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 2 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 3 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 4 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Viaarxiv icon

EAGLE: Egocentric AGgregated Language-video Engine

Add code
Sep 26, 2024
Figure 1 for EAGLE: Egocentric AGgregated Language-video Engine
Figure 2 for EAGLE: Egocentric AGgregated Language-video Engine
Figure 3 for EAGLE: Egocentric AGgregated Language-video Engine
Figure 4 for EAGLE: Egocentric AGgregated Language-video Engine
Viaarxiv icon

AVicuna: Audio-Visual LLM with Interleaver and Context-Boundary Alignment for Temporal Referential Dialogue

Add code
Mar 24, 2024
Viaarxiv icon

OSCaR: Object State Captioning and State Change Representation

Add code
Feb 28, 2024
Figure 1 for OSCaR: Object State Captioning and State Change Representation
Figure 2 for OSCaR: Object State Captioning and State Change Representation
Figure 3 for OSCaR: Object State Captioning and State Change Representation
Figure 4 for OSCaR: Object State Captioning and State Change Representation
Viaarxiv icon

Video Understanding with Large Language Models: A Survey

Add code
Jan 04, 2024
Figure 1 for Video Understanding with Large Language Models: A Survey
Figure 2 for Video Understanding with Large Language Models: A Survey
Figure 3 for Video Understanding with Large Language Models: A Survey
Viaarxiv icon

MISAR: A Multimodal Instructional System with Augmented Reality

Add code
Oct 18, 2023
Figure 1 for MISAR: A Multimodal Instructional System with Augmented Reality
Figure 2 for MISAR: A Multimodal Instructional System with Augmented Reality
Figure 3 for MISAR: A Multimodal Instructional System with Augmented Reality
Figure 4 for MISAR: A Multimodal Instructional System with Augmented Reality
Viaarxiv icon

Multi-omics Prediction from High-content Cellular Imaging with Deep Learning

Add code
Jun 19, 2023
Viaarxiv icon