Picture for Pooyan Fazli

Pooyan Fazli

VidHalluc: Evaluating Temporal Hallucinations in Multimodal Large Language Models for Video Understanding

Add code
Dec 04, 2024
Viaarxiv icon

VideoSAVi: Self-Aligned Video Language Models without Human Supervision

Add code
Dec 01, 2024
Viaarxiv icon

VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?

Add code
Nov 19, 2024
Figure 1 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 2 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 3 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 4 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Viaarxiv icon

OSCaR: Object State Captioning and State Change Representation

Add code
Feb 28, 2024
Figure 1 for OSCaR: Object State Captioning and State Change Representation
Figure 2 for OSCaR: Object State Captioning and State Change Representation
Figure 3 for OSCaR: Object State Captioning and State Change Representation
Figure 4 for OSCaR: Object State Captioning and State Change Representation
Viaarxiv icon

CLearViD: Curriculum Learning for Video Description

Add code
Nov 08, 2023
Viaarxiv icon

Clustering Social Touch Gestures for Human-Robot Interaction

Add code
Apr 03, 2023
Viaarxiv icon

Charting Visual Impression of Robot Hands

Add code
Nov 17, 2022
Viaarxiv icon

NarrationBot and InfoBot: A Hybrid System for Automated Video Description

Add code
Nov 07, 2021
Figure 1 for NarrationBot and InfoBot: A Hybrid System for Automated Video Description
Figure 2 for NarrationBot and InfoBot: A Hybrid System for Automated Video Description
Figure 3 for NarrationBot and InfoBot: A Hybrid System for Automated Video Description
Figure 4 for NarrationBot and InfoBot: A Hybrid System for Automated Video Description
Viaarxiv icon

Real-time Policy Distillation in Deep Reinforcement Learning

Add code
Dec 29, 2019
Figure 1 for Real-time Policy Distillation in Deep Reinforcement Learning
Figure 2 for Real-time Policy Distillation in Deep Reinforcement Learning
Viaarxiv icon

DeepMoTIon: Learning to Navigate Like Humans

Add code
Mar 02, 2019
Figure 1 for DeepMoTIon: Learning to Navigate Like Humans
Viaarxiv icon