☆ 3.8 Proceedings Paper

Spatiotemporal Modeling for Crowd Counting in Videos

2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV) (2017)

期刊

2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV)

卷 -, 期 -, 页码 5161-5169

出版社

IEEE

DOI: 10.1109/ICCV.2017.551

关键词

类别

Computer Science, Artificial Intelligence Engineering, Electrical & Electronic

资金

Research Grants Council of Hong Kong [16207316]
Innovation and Technology Commission of Hong Kong [ITS/170/15FP]

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

Region of Interest (ROI) crowd counting can be formulated as a regression problem of learning a mapping from an image or a video frame to a crowd density map. Recently, convolutional neural network (CNN) models have achieved promising results for crowd counting. However, even when dealing with video data, CNN-based methods still consider each video frame independently, ignoring the strong temporal correlation between neighboring frames. To exploit the otherwise very useful temporal information in video sequences, we propose a variant of a recent deep learning model called convolutional LSTM (ConvLSTM) for crowd counting. Unlike the previous CNN-based methods, our method fully captures both spatial and temporal dependencies. Furthermore, we extend the ConvLSTM model to a bidirectional ConvLSTM model which can access long-range information in both directions. Extensive experiments using four publicly available datasets demonstrate the reliability of our approach and the effectiveness of incorporating temporal information to boost the accuracy of crowd counting. In addition, we also conduct some transfer learning experiments to show that once our model is trained on one dataset, its learning experience can be transferred easily to a new dataset which consists of only very few video frames for model adaptation.

Spatiotemporal Modeling for Crowd Counting in Videos

期刊

2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV)

出版社

IEEE

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

Spatiotemporal Modeling for Crowd Counting in Videos

期刊

2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV)

出版社

IEEE

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文