Ana gezinime geç Aramaya geç Ana içeriğe geç

Automatic emotion recognition in the wild using an ensemble of static and dynamic representations

  • Sabanci University

Araştırma çıktısı: Kitap/Rapor/Konferans Bildirisinde BölümKonferans katkısıHakem

7 Atıf (Scopus)

Özet

Automatic emotion recognition in the wild video datasets is a very challenging problem because of the inter-class similarities among different facial expressions and large intraclass variabilities due to the significant changes in illumination, pose, scene, and expression. In this paper, we present our proposed method for video-based emotion recognition in the EmotiW 2016 challenge. The task considers the unconstrained emotion recognition problem by training from short video clips extracted from movies and testing on short movie clips and spontaneous video clips of the reality TV data. Four different methods are employed to extract both static and dynamic emotion representations from the videos. First, local binary patterns of three orthogonal planes are used to describe spatiotemporal features of the video frames. Second, principal component analysis is applied to the image patches in a two-stage convolutional network to learn weights and extract facial features from the aligned faces. Third, the deep convolutional neural network model of VGGFace is deployed to extract deep facial representations from aligned faces. Fourth, a bag of visual words is computed based on dense scale-invariant feature transform descriptors from aligned face images to form hand-crafted representations. Support vector machines are then utilized to train and classify the obtained spatiotemporal representations and facial features. Finally, score-level fusion is applied to combine the classification results and predict the emotion labels of the video clips. The results show that the proposed combined method has outperformed all the utilized techniques with the overall validations and test accuracies of 43.13% and 40.13%, respectively. This system, is relatively a good classifier in Happy and Angry emotion categories and is unsuccessful in detecting Surprise, Disgust, and Fear.

Orijinal dilİngilizce
Ana bilgisayar yayını başlığıICMI 2016 - Proceedings of the 18th ACM International Conference on Multimodal Interaction
EditörlerCatherine Pelachaud, Yukiko I. Nakano, Toyoaki Nishida, Carlos Busso, Louis-Philippe Morency, Elisabeth Andre
YayınlayanAssociation for Computing Machinery, Inc
Sayfalar514-521
Sayfa sayısı8
ISBN (Elektronik)9781450345569
DOI'lar
Yayın durumuYayınlandı - 31 Eki 2016
Etkinlik18th ACM International Conference on Multimodal Interaction, ICMI 2016 - Tokyo, Japan
Süre: 12 Kas 201616 Kas 2016

Yayın serisi

AdıICMI 2016 - Proceedings of the 18th ACM International Conference on Multimodal Interaction

???event.eventtypes.event.conference???

???event.eventtypes.event.conference???18th ACM International Conference on Multimodal Interaction, ICMI 2016
Ülke/BölgeJapan
ŞehirTokyo
Periyot12/11/1616/11/16

Bibliyografik not

Publisher Copyright:
© 2016 ACM.

Parmak izi

Automatic emotion recognition in the wild using an ensemble of static and dynamic representations' araştırma başlıklarına git. Birlikte benzersiz bir parmak izi oluştururlar.

Alıntı Yap