Using Segmentation to Enhance Frame Prediction in a Multi-Scale Spatial-Temporal Feature Extraction Network

This paper proposes a two-stream convolution network to extract spatial and temporal cues for video based person ReIdentification (ReID). A temporal stream in this network is constructed by inserting several Multi-scale 3D (M3D) convolution layers into a 2D CNN network. The resulting M3D convolution network introduces a fraction of parameters into the 2D CNN, but gains the ability of multi-scale temporal feature learning. With this compact architecture, M3D convolution network is also more efficient and easier to optimize than existing 3D convolution networks. The temporal stream further involves Residual Attention Layers (RAL) to refine the temporal features. By jointly learning spatial-temporal attention masks in a residual manner, RAL identifies the discriminative spatial regions and temporal cues. The other stream in our network is implemented with a 2D CNN for spatial feature extraction. The spatial and temporal features from two streams are finally fused for the video based person ReID. Evaluations on three widely used benchmarks datasets, i.e.,MARS, PRID2011, and iLIDS-VID demonstrate the substantial advantages of our method over existing 3D convolution networks and state-of-art methods.

Download Full-text

On the Impact of Gabor Phase for Spectro-Temporal Feature Extraction in Building an ASR System

2020 11th IEEE Annual Information Technology, Electronics and Mobile Communication Conference (IEMCON) ◽

10.1109/iemcon51383.2020.9284872 ◽

2020 ◽

Author(s):

Anirban Dutta ◽

Gudmalwar Prabhakar ◽

Ch V Rama Rao

Keyword(s):

Feature Extraction ◽

The Impact ◽

Asr System ◽

Temporal Feature

Download Full-text

A Multi-Scale Feature Extraction-Based Normalized Attention Neural Network for Image Denoising

Electronics ◽

10.3390/electronics10030319 ◽

2021 ◽

Vol 10 (3) ◽

pp. 319

Author(s):

Yi Wang ◽

Xiao Song ◽

Guanghong Gong ◽

Ni Li

Keyword(s):

Neural Network ◽

Feature Extraction ◽

Image Denoising ◽

Color Image ◽

Rapid Development ◽

Similarity Index ◽

Structural Similarity ◽

Convolutional Network ◽

Scale Feature ◽

Multi Scale

Due to the rapid development of deep learning and artificial intelligence techniques, denoising via neural networks has drawn great attention due to their flexibility and excellent performances. However, for most convolutional network denoising methods, the convolution kernel is only one layer deep, and features of distinct scales are neglected. Moreover, in the convolution operation, all channels are treated equally; the relationships of channels are not considered. In this paper, we propose a multi-scale feature extraction-based normalized attention neural network (MFENANN) for image denoising. In MFENANN, we define a multi-scale feature extraction block to extract and combine features at distinct scales of the noisy image. In addition, we propose a normalized attention network (NAN) to learn the relationships between channels, which smooths the optimization landscape and speeds up the convergence process for training an attention model. Moreover, we introduce the NAN to convolutional network denoising, in which each channel gets gain; channels can play different roles in the subsequent convolution. To testify the effectiveness of the proposed MFENANN, we used both grayscale and color image sets whose noise levels ranged from 0 to 75 to do the experiments. The experimental results show that compared with some state-of-the-art denoising methods, the restored images of MFENANN have larger peak signal-to-noise ratios (PSNR) and structural similarity index measure (SSIM) values and get better overall appearance.

Download Full-text

Image Compressive Sensing via Multi-scale Feature Extraction and Attention Mechanism

2020 International Conference on Intelligent Computing, Automation and Systems (ICICAS) ◽

10.1109/icicas51530.2020.00061 ◽

2020 ◽

Author(s):

Chuning He

Keyword(s):

Feature Extraction ◽

Compressive Sensing ◽

Attention Mechanism ◽

Scale Feature ◽

Multi Scale

Download Full-text

FP-STE: A Novel Node Failure Prediction Method Based on Spatio-Temporal Feature Extraction in Data Centers

Computer Modeling in Engineering & Sciences ◽

10.32604/cmes.2020.09404 ◽

2020 ◽

Vol 123 (3) ◽

pp. 1015-1031

Author(s):

Yang Yang ◽

Jing Dong ◽

Chao Fang ◽

Ping Xie ◽

Na An

Keyword(s):

Feature Extraction ◽

Data Centers ◽

Failure Prediction ◽

Prediction Method ◽

Node Failure ◽

Spatio Temporal ◽

Temporal Feature

Download Full-text

Design strategies for direct multi-scale and multi-orientation feature extraction in the log-polar domain

Pattern Recognition Letters ◽

10.1016/j.patrec.2011.09.021 ◽

2012 ◽

Vol 33 (1) ◽

pp. 41-51 ◽

Cited By ~ 14

Author(s):

Fabio Solari ◽

Manuela Chessa ◽

Silvio P. Sabatini

Keyword(s):

Feature Extraction ◽

Design Strategies ◽

Multi Scale ◽

Orientation Feature

Download Full-text

Multi-scale feature extraction algorithm of ear image

2011 International Conference on Electric Information and Control Engineering ◽

10.1109/iceice.2011.5777641 ◽

2011 ◽

Cited By ~ 3

Author(s):

Zhi-qin Wang ◽

Xiao-dong Yan

Keyword(s):

Feature Extraction ◽

Scale Feature ◽

Multi Scale ◽

Feature Extraction Algorithm ◽

Extraction Algorithm

Download Full-text

Learning rich features with hybrid loss for brain tumor segmentation

BMC Medical Informatics and Decision Making ◽

10.1186/s12911-021-01431-y ◽

2021 ◽

Vol 21 (S2) ◽

Author(s):

Daobin Huang ◽

Minghui Wang ◽

Ling Zhang ◽

Haichun Li ◽

Minquan Ye ◽

...

Keyword(s):

Feature Extraction ◽

Brain Tumor ◽

Class Imbalance ◽

Feature Representation ◽

Loss Functions ◽

Radiotherapy Planning ◽

Tumor Segmentation ◽

Brain Tumor Segmentation ◽

Scale Feature ◽

Multi Scale

Abstract Background Accurately segment the tumor region of MRI images is important for brain tumor diagnosis and radiotherapy planning. At present, manual segmentation is wildly adopted in clinical and there is a strong need for an automatic and objective system to alleviate the workload of radiologists. Methods We propose a parallel multi-scale feature fusing architecture to generate rich feature representation for accurate brain tumor segmentation. It comprises two parts: (1) Feature Extraction Network (FEN) for brain tumor feature extraction at different levels and (2) Multi-scale Feature Fusing Network (MSFFN) for merge all different scale features in a parallel manner. In addition, we use two hybrid loss functions to optimize the proposed network for the class imbalance issue. Results We validate our method on BRATS 2015, with 0.86, 0.73 and 0.61 in Dice for the three tumor regions (complete, core and enhancing), and the model parameter size is only 6.3 MB. Without any post-processing operations, our method still outperforms published state-of-the-arts methods on the segmentation results of complete tumor regions and obtains competitive performance in another two regions. Conclusions The proposed parallel structure can effectively fuse multi-level features to generate rich feature representation for high-resolution results. Moreover, the hybrid loss functions can alleviate the class imbalance issue and guide the training process. The proposed method can be used in other medical segmentation tasks.

Download Full-text

Progressive Spatio-Temporal Feature Extraction Model For Gait Recognition

10.1109/icip42928.2021.9506490 ◽

2021 ◽

Author(s):

Jingran Su ◽

Yang Zhao ◽

Xuelong Li

Keyword(s):

Feature Extraction ◽

Gait Recognition ◽

Spatio Temporal ◽

Extraction Model ◽

Temporal Feature

Download Full-text