Picture for Jing Bi

Jing Bi

Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning

Add code
Dec 13, 2024
Viaarxiv icon

VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?

Add code
Nov 19, 2024
Figure 1 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 2 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 3 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 4 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Viaarxiv icon

EAGLE: Egocentric AGgregated Language-video Engine

Add code
Sep 26, 2024
Figure 1 for EAGLE: Egocentric AGgregated Language-video Engine
Figure 2 for EAGLE: Egocentric AGgregated Language-video Engine
Figure 3 for EAGLE: Egocentric AGgregated Language-video Engine
Figure 4 for EAGLE: Egocentric AGgregated Language-video Engine
Viaarxiv icon

AVicuna: Audio-Visual LLM with Interleaver and Context-Boundary Alignment for Temporal Referential Dialogue

Add code
Mar 24, 2024
Viaarxiv icon

OSCaR: Object State Captioning and State Change Representation

Add code
Feb 28, 2024
Figure 1 for OSCaR: Object State Captioning and State Change Representation
Figure 2 for OSCaR: Object State Captioning and State Change Representation
Figure 3 for OSCaR: Object State Captioning and State Change Representation
Figure 4 for OSCaR: Object State Captioning and State Change Representation
Viaarxiv icon

Video Understanding with Large Language Models: A Survey

Add code
Jan 04, 2024
Viaarxiv icon

MISAR: A Multimodal Instructional System with Augmented Reality

Add code
Oct 18, 2023
Figure 1 for MISAR: A Multimodal Instructional System with Augmented Reality
Figure 2 for MISAR: A Multimodal Instructional System with Augmented Reality
Figure 3 for MISAR: A Multimodal Instructional System with Augmented Reality
Figure 4 for MISAR: A Multimodal Instructional System with Augmented Reality
Viaarxiv icon

Multi-omics Prediction from High-content Cellular Imaging with Deep Learning

Add code
Jun 19, 2023
Viaarxiv icon

Performances of Symmetric Loss for Private Data from Exponential Mechanism

Add code
Oct 09, 2022
Figure 1 for Performances of Symmetric Loss for Private Data from Exponential Mechanism
Figure 2 for Performances of Symmetric Loss for Private Data from Exponential Mechanism
Figure 3 for Performances of Symmetric Loss for Private Data from Exponential Mechanism
Figure 4 for Performances of Symmetric Loss for Private Data from Exponential Mechanism
Viaarxiv icon

Procedure Planning in Instructional Videos via Contextual Modeling and Model-based Policy Learning

Add code
Oct 08, 2021
Figure 1 for Procedure Planning in Instructional Videos via Contextual Modeling and Model-based Policy Learning
Figure 2 for Procedure Planning in Instructional Videos via Contextual Modeling and Model-based Policy Learning
Figure 3 for Procedure Planning in Instructional Videos via Contextual Modeling and Model-based Policy Learning
Figure 4 for Procedure Planning in Instructional Videos via Contextual Modeling and Model-based Policy Learning
Viaarxiv icon