“CN-Celeb”版本间的差异

2019年10月29日 (二) 12:11的版本

Deng et al., "RetinaFace: Single-stage Dense Face Localisation in the Wild", 2019. [1]
Deng et al., "ArcFace: Additive Angular Margin Loss for Deep Face Recognition", 2018, [2]
Wang et al., "CosFace: Large Margin Cosine Loss for Deep Face Recognition", 2018, [3]
Liu et al., "SphereFace: Deep Hypersphere Embedding for Face Recognition", 2017[4]
Zhong et al., "GhostVLAD for set-based face recognition", 2018. link
Chung et al., "Out of time: automated lip sync in the wild", 2016.link
Xie et al., "UTTERANCE-LEVEL AGGREGATION FOR SPEAKER RECOGNITION IN THE WILD", 2019. link
Zhang1 et al., "FULLY SUPERVISED SPEAKER DIARIZATION", 2018. link

@@ 第22行： / 第22行： @@
 * Environments: Tensorflow, PyTorch, Keras, MxNet
-* Face detection and tracking based on RetinaFace and ArcFace models.
+* Face detection and tracking: RetinaFace and ArcFace models.
-* Active speaker verification based on SyncNet model.
+* Active speaker verification: SyncNet model.
-* Speaker Diarization based on UIS-RNN model.
+* Speaker Diarization: UIS-RNN model.
-* Double check by speaker recognition based on VGG model.
+* Double check by speaker recognition: VGG model.
 * Input: Pictures and videos of POIs (Persons of Interest).
 * Output: well-labelled videos of POIs (Persons of Interest).