TSSD: Temporal Single-Shot Detector Based on Attention and LSTM for Robotic Intelligent Perception

03/01/2018
by   Xingyu Chen, et al.
0

Temporal object detection has attracted significant attention, but most popular detection methods can not leverage the rich temporal information in video or robotic vision. Although many different algorithms have been developed for video detection task, real-time online approaches are frequently deficient. In this paper, based on attention mechanism and convolutional long short-term memory (ConvLSTM), we propose a temporal single-shot detector (TSSD) for robotic vision. Distinct from previous methods, we take aim at temporally integrating pyramidal feature hierarchy using ConvLSTM, and design a novel structure including a high-level ConvLSTM unit as well as a low-level one (HL-LSTM) for multi-scale feature maps. Moreover, we develop a creative temporal analysis unit, namely ConvLSTM-based attention and attention-based ConvLSTM (A&CL), in which ConvLSTM-based attention is specially tailored for background suppression and scale suppression while attention-based ConvLSTM temporally integrates attention-aware features. Finally, our method is evaluated on ImageNet VID dataset. Extensive comparisons on the detection capability confirm or validate the superiority of the proposed approach. Consequently, the developed TSSD is fairly faster and achieves a considerably enhanced performance in terms of mean average precision. As a temporal, real-time, and online detector, TSSD is applicable to robot's intelligent perception.

READ FULL TEXT

page 2

page 3

page 4

page 5

page 6

research
03/01/2018

TSSD: Temporal Single-Shot Object Detection Based on Attention-Aware LSTM

Temporal object detection has attracted significant attention, but most ...
research
11/17/2017

Mobile Video Object Detection with Temporally-Aware Feature Maps

This paper introduces an online model for object detection in videos des...
research
03/14/2022

Attention based Memory video portrait matting

We proposed a novel trimap free video matting method based on the attent...
research
03/04/2019

TKD: Temporal Knowledge Distillation for Active Perception

Deep neural networks based methods have been proved to achieve outstandi...
research
05/03/2023

Attention Based Feature Fusion For Multi-Agent Collaborative Perception

In the domain of intelligent transportation systems (ITS), collaborative...
research
05/31/2021

Multi-Scale Attention Neural Network for Acoustic Echo Cancellation

Acoustic Echo Cancellation (AEC) plays a key role in speech interaction ...

Please sign up or login with your details

Forgot password? Click here to reset