Joint Face Representation Adaptation And Clustering In Videos
COMPUTER VISION - ECCV 2016, PT III(2016)
摘要
Clustering faces in movies or videos is extremely challenging since characters' appearance can vary drastically under different scenes. In addition, the various cinematic styles make it difficult to learn a universal face representation for all videos. Unlike previous methods that assume fixed handcrafted features for face clustering, in this work, we formulate a joint face representation adaptation and clustering approach in a deep learning framework. The proposed method allows face representation to gradually adapt from an external source domain to a target video domain. The adaptation of deep representation is achieved without any strong supervision but through iteratively discovered weak pairwise identity constraints derived from potentially noisy face clustering result. Experiments on three benchmark video datasets demonstrate that our approach generates character clusters with high purity compared to existing video face clustering methods, which are either based on deep face representation (without adaptation) or carefully engineered features.
更多查看译文
关键词
Convolutional network, Transfer learning, Face clustering, Face recognition
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络