Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

A. S. M. Iftekhar

BLoad: Enhancing Neural Network Training with Efficient Sequential Data Handling

Oct 16, 2023

Raphael Ruschel, A. S. M. Iftekhar, B. S. Manjunath, Suya You

Figure 1 for BLoad: Enhancing Neural Network Training with Efficient Sequential Data Handling

Figure 2 for BLoad: Enhancing Neural Network Training with Efficient Sequential Data Handling

Figure 3 for BLoad: Enhancing Neural Network Training with Efficient Sequential Data Handling

Figure 4 for BLoad: Enhancing Neural Network Training with Efficient Sequential Data Handling

Abstract:The increasing complexity of modern deep neural network models and the expanding sizes of datasets necessitate the development of optimized and scalable training methods. In this white paper, we addressed the challenge of efficiently training neural network models using sequences of varying sizes. To address this challenge, we propose a novel training scheme that enables efficient distributed data-parallel training on sequences of different sizes with minimal overhead. By using this scheme we were able to reduce the padding amount by more than 100$x$ while not deleting a single frame, resulting in an overall increased performance on both training time and Recall in our experiments.

Via

Access Paper or Ask Questions

CNN-Based Prediction of Frame-Level Shot Importance for Video Summarization

Aug 23, 2017

Mohaiminul Al Nahian, A. S. M. Iftekhar, Mohammad Tariqul Islam, S. M. Mahbubur Rahman, Dimitrios Hatzinakos

Figure 1 for CNN-Based Prediction of Frame-Level Shot Importance for Video Summarization

Figure 2 for CNN-Based Prediction of Frame-Level Shot Importance for Video Summarization

Figure 3 for CNN-Based Prediction of Frame-Level Shot Importance for Video Summarization

Figure 4 for CNN-Based Prediction of Frame-Level Shot Importance for Video Summarization

Abstract:In the Internet, ubiquitous presence of redundant, unedited, raw videos has made video summarization an important problem. Traditional methods of video summarization employ a heuristic set of hand-crafted features, which in many cases fail to capture subtle abstraction of a scene. This paper presents a deep learning method that maps the context of a video to the importance of a scene similar to that is perceived by humans. In particular, a convolutional neural network (CNN)-based architecture is proposed to mimic the frame-level shot importance for user-oriented video summarization. The weights and biases of the CNN are trained extensively through off-line processing, so that it can provide the importance of a frame of an unseen video almost instantaneously. Experiments on estimating the shot importance is carried out using the publicly available database TVSum50. It is shown that the performance of the proposed network is substantially better than that of commonly referred feature-based methods for estimating the shot importance in terms of mean absolute error, absolute error variance, and relative F-measure.

* Accepted in International Conference on new Trends in Computer Sciences (ICTCS), Amman-Jordan, 2017

Via

Access Paper or Ask Questions