Picture for Pooyan Fazli

Pooyan Fazli

VideoPASTA: 7K Preference Pairs That Matter for Video-LLM Alignment

Add code
Apr 18, 2025
Viaarxiv icon

ChartQA-X: Generating Explanations for Charts

Add code
Apr 17, 2025
Viaarxiv icon

VideoA11y: Method and Dataset for Accessible Video Description

Add code
Feb 27, 2025
Viaarxiv icon

VidHalluc: Evaluating Temporal Hallucinations in Multimodal Large Language Models for Video Understanding

Add code
Dec 04, 2024
Viaarxiv icon

VideoSAVi: Self-Aligned Video Language Models without Human Supervision

Add code
Dec 01, 2024
Viaarxiv icon

VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?

Add code
Nov 19, 2024
Figure 1 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 2 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 3 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Figure 4 for VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
Viaarxiv icon

OSCaR: Object State Captioning and State Change Representation

Add code
Feb 28, 2024
Figure 1 for OSCaR: Object State Captioning and State Change Representation
Figure 2 for OSCaR: Object State Captioning and State Change Representation
Figure 3 for OSCaR: Object State Captioning and State Change Representation
Figure 4 for OSCaR: Object State Captioning and State Change Representation
Viaarxiv icon

CLearViD: Curriculum Learning for Video Description

Add code
Nov 08, 2023
Viaarxiv icon

Clustering Social Touch Gestures for Human-Robot Interaction

Add code
Apr 03, 2023
Viaarxiv icon

Charting Visual Impression of Robot Hands

Add code
Nov 17, 2022
Viaarxiv icon