Deep Attributes Driven Multi-Camera Person Re-identification
Su, Zhang, Xing, Gao, Tian · cs.CV · 2016-08-09 · 原文
The visual appearance of a person is easily affected by many factors like pose variations, viewpoint changes and camera parameter differences. This makes person Re-Identification (ReID) among multiple cameras a very challenging task. This work is motivated to learn mid-level human attributes which are robust to such visual appearance variations. And we propose a semi-supervised attribute learning framework which progressively boosts the accuracy of attributes only using a limited number of labeled data. Specifically, this framework involves a three-stage training. A deep Convolutional Neural Network (dCNN) is first trained on an independent dataset labeled with attributes. Then it is fine-tuned on another dataset only labeled with person IDs using our defined triplet loss. Finally, the updated dCNN predicts attribute labels for the target dataset, which is combined with the independent dataset for the final round of fine-tuning. The predicted attributes, namely \emph{deep attributes} exhibit superior generalization ability across different datasets. By directly using the deep attributes with simple Cosine distance, we have obtained surprisingly good accuracy on four person ReID dat
讲义
讲义·推断 依据「原文」自动生成的结构化摘要(推断),非原文表述;以原文为准。
1. 人话版
The visual appearance of a person is easily affected by many factors like pose variations, viewpoint changes and camera parameter differences.
This makes person Re-Identification (ReID) among multiple cameras a very challenging task.
2. 领域脉络
本文类目:cs.CV,属于其所在研究脉络的最新进展。
3. 机制拆解
Specifically, this framework involves a three-stage training.
4. 证据与数字
摘要未给出量化结果——留意原文的实验与数据。
5. 反例与边界
This work is motivated to learn mid-level human attributes which are robust to such visual appearance variations.
And we propose a semi-supervised attribute learning framework which progressively boosts the accuracy of attributes only using a limited number of labeled data.
A deep Convolutional Neural Network (dCNN) is first trained on an independent dataset labeled with attributes.
Then it is fine-tuned on another dataset only labeled with person IDs using our defined triplet loss.
6. 跨领域连接与意外收获
思考本文机制能否迁移到你正在跟进的问题。
7. 可复用方法
把本文机制与你手头项目对照,找一个两周内能验证的最小实验。
8. 术语表
精读时把不熟的术语记入此处,作为下次回忆的锚点。