Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Hongcang Jin

Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training

May 09, 2024

Sheng Yan, Xin Du, Zongying Li, Yi Wang, Hongcang Jin, Mengyuan Liu

Figure 1 for Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training

Figure 2 for Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training

Figure 3 for Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training

Figure 4 for Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training

Abstract:Temporal grounding is crucial in multimodal learning, but it poses challenges when applied to animal behavior data due to the sparsity and uniform distribution of moments. To address these challenges, we propose a novel Positional Recovery Training framework (Port), which prompts the model with the start and end times of specific animal behaviors during training. Specifically, Port enhances the baseline model with a Recovering part to predict flipped label sequences and align distributions with a Dual-alignment method. This allows the model to focus on specific temporal regions prompted by ground-truth information. Extensive experiments on the Animal Kingdom dataset demonstrate the effectiveness of Port, achieving an IoU@0.3 of 38.52. It emerges as one of the top performers in the sub-track of MMVRAC in ICME 2024 Grand Challenges.

* Accepted by ICMEW 2024. arXiv admin note: text overlap with arXiv:2404.13657

Via

Access Paper or Ask Questions