Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Yunhan Wang

Benchmarking Data Efficiency and Computational Efficiency of Temporal Action Localization Models

Aug 24, 2023

Jan Warchocki, Teodor Oprescu, Yunhan Wang, Alexandru Damacus, Paul Misterka, Robert-Jan Bruintjes, Attila Lengyel, Ombretta Strafforello, Jan van Gemert

Figure 1 for Benchmarking Data Efficiency and Computational Efficiency of Temporal Action Localization Models

Figure 2 for Benchmarking Data Efficiency and Computational Efficiency of Temporal Action Localization Models

Figure 3 for Benchmarking Data Efficiency and Computational Efficiency of Temporal Action Localization Models

Figure 4 for Benchmarking Data Efficiency and Computational Efficiency of Temporal Action Localization Models

Abstract:In temporal action localization, given an input video, the goal is to predict which actions it contains, where they begin, and where they end. Training and testing current state-of-the-art deep learning models requires access to large amounts of data and computational power. However, gathering such data is challenging and computational resources might be limited. This work explores and measures how current deep temporal action localization models perform in settings constrained by the amount of data or computational power. We measure data efficiency by training each model on a subset of the training set. We find that TemporalMaxer outperforms other models in data-limited settings. Furthermore, we recommend TriDet when training time is limited. To test the efficiency of the models during inference, we pass videos of different lengths through each model. We find that TemporalMaxer requires the least computational resources, likely due to its simple architecture.

* Accepted to the CVEU workshop at ICCV 2023

Via

Access Paper or Ask Questions

Investigation of Architectures and Receptive Fields for Appearance-based Gaze Estimation

Aug 18, 2023

Yunhan Wang, Xiangwei Shi, Shalini De Mello, Hyung Jin Chang, Xucong Zhang

Abstract:With the rapid development of deep learning technology in the past decade, appearance-based gaze estimation has attracted great attention from both computer vision and human-computer interaction research communities. Fascinating methods were proposed with variant mechanisms including soft attention, hard attention, two-eye asymmetry, feature disentanglement, rotation consistency, and contrastive learning. Most of these methods take the single-face or multi-region as input, yet the basic architecture of gaze estimation has not been fully explored. In this paper, we reveal the fact that tuning a few simple parameters of a ResNet architecture can outperform most of the existing state-of-the-art methods for the gaze estimation task on three popular datasets. With our extensive experiments, we conclude that the stride number, input image resolution, and multi-region architecture are critical for the gaze estimation performance while their effectiveness dependent on the quality of the input face image. We obtain the state-of-the-art performances on three datasets with 3.64 on ETH-XGaze, 4.50 on MPIIFaceGaze, and 9.13 on Gaze360 degrees gaze estimation error by taking ResNet-50 as the backbone.

Via

Access Paper or Ask Questions

A Deep Learning Approach to Predicting Ventilator Parameters for Mechanically Ventilated Septic Patients

Feb 21, 2022

Zhijun Zeng, Zhen Hou, Ting Li, Lei Deng, Jianguo Hou, Xinran Huang, Jun Li, Meirou Sun, Yunhan Wang, Qiyu Wu(+3 more)

Figure 1 for A Deep Learning Approach to Predicting Ventilator Parameters for Mechanically Ventilated Septic Patients

Figure 2 for A Deep Learning Approach to Predicting Ventilator Parameters for Mechanically Ventilated Septic Patients

Figure 3 for A Deep Learning Approach to Predicting Ventilator Parameters for Mechanically Ventilated Septic Patients

Figure 4 for A Deep Learning Approach to Predicting Ventilator Parameters for Mechanically Ventilated Septic Patients

Abstract:We develop a deep learning approach to predicting a set of ventilator parameters for a mechanically ventilated septic patient using a long and short term memory (LSTM) recurrent neural network (RNN) model. We focus on short-term predictions of a set of ventilator parameters for the septic patient in emergency intensive care unit (EICU). The short-term predictability of the model provides attending physicians with early warnings to make timely adjustment to the treatment of the patient in the EICU. The patient specific deep learning model can be trained on any given critically ill patient, making it an intelligent aide for physicians to use in emergent medical situations.

Via

Access Paper or Ask Questions