arXiv (Cornell University) · 2017 · 178 citations · 22 references
Pose Estimation ConfidenceEngineeringMachine LearningPose Invariant EmbeddingHuman Pose EstimationBiometricsImage AnalysisPedestrian MisalignmentPattern RecognitionMachine VisionFeature LearningObject DetectionData Re-identificationComputer ScienceDeep LearningPose EstimationComputer VisionHuman IdentificationObject Recognition
Pedestrian misalignment, which mainly arises from detector errors and pose variations, is a critical problem for a robust person re-identification (re-ID) system. With bad alignment, the background noise will significantly compromise the feature learning and matching process. To address this problem, this paper introduces the pose invariant embedding (PIE) as a pedestrian descriptor. First, in order to align pedestrians to a standard pose, the PoseBox structure is introduced, which is generated through pose estimation followed by affine transformations. Second, to reduce the impact of pose estimation errors and information loss during PoseBox construction, we design a PoseBox fusion (PBF) CNN architecture that takes the original image, the PoseBox, and the pose estimation confidence as input. The proposed PIE descriptor is thus defined as the fully connected layer of the PBF network for the retrieval task. Experiments are conducted on the Market-1501, CUHK03, and VIPeR datasets. We show that PoseBox alone yields decent re-ID accuracy and that when integrated in the PBF network, the learned PIE descriptor produces competitive performance compared with the state-of-the-art approaches.
22
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher et al. · 2009 IEEE Conference on Computer Vision and Pattern Recognition · 2009 · 60.2K citations
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren et al. · arXiv (Cornell University) · 2015 · 4.6K citations · Full text
Caffe: Convolutional Architecture for Fast Feature Embedding
Yangqing Jia, Evan Shelhamer, Jeff Donahue et al. · arXiv (Cornell University) · 2014 · 4.3K citations · Full text
Convolutional Neural Network, Engineering, Machine Learning +16
2D Human Pose Estimation: New Benchmark and State of the Art Analysis
Mykhaylo Andriluka, Leonid Pishchulin, Peter Gehler et al. · 2014 · 2.8K citations