{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,22]],"date-time":"2026-07-22T02:10:54Z","timestamp":1784686254220,"version":"3.55.0"},"reference-count":34,"publisher":"IGI Global Scientific Publishing","issue":"4","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2013,10,1]]},"abstract":"<p>This paper aims at providing general theoretical analysis for the issue of multimodal information fusion and implementing novel information theoretic tools in multimedia application. The most essential issues for information fusion include feature transformation and reduction of feature dimensionality. Most previous solutions are largely based on the second order statistics, which is only optimal for Gaussian-like distribution, while in this paper we describe kernel entropy component analysis (KECA) which utilizes descriptor of information entropy and achieves improved performance by entropy estimation. The authors present a new solution based on the integration of information fusion theory and information theoretic tools in this paper. The proposed method has been applied to audiovisual emotion recognition. Information fusion has been implemented for audio and video channels at feature level and decision level. Experimental results demonstrate that the proposed algorithm achieves improved performance in comparison with the existing methods, especially when the dimension of feature space is substantially reduced.<\/p>","DOI":"10.4018\/ijmdem.2013100101","type":"journal-article","created":{"date-parts":[[2014,3,11]],"date-time":"2014-03-11T14:25:59Z","timestamp":1394547959000},"page":"1-14","source":"Crossref","is-referenced-by-count":16,"title":["Multimodal Information Fusion of Audiovisual Emotion Recognition Using Novel Information Theoretic Tools"],"prefix":"10.4018","volume":"4","author":[{"given":"Zhibing","family":"Xie","sequence":"first","affiliation":[{"name":"Ryerson Multimedia Research Lab, Ryerson University, Toronto, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ling","family":"Guan","sequence":"additional","affiliation":[{"name":"Ryerson Multimedia Research Lab, Ryerson University, Toronto, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"2432","reference":[{"key":"ijmdem.2013100101-0","doi-asserted-by":"publisher","DOI":"10.1007\/s00530-010-0182-0"},{"key":"ijmdem.2013100101-1","doi-asserted-by":"publisher","DOI":"10.1016\/j.engappai.2007.11.010"},{"key":"ijmdem.2013100101-2","doi-asserted-by":"publisher","DOI":"10.1016\/j.specom.2008.03.012"},{"key":"ijmdem.2013100101-3","doi-asserted-by":"publisher","DOI":"10.1109\/79.911197"},{"key":"ijmdem.2013100101-4","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2010.09.020"},{"key":"ijmdem.2013100101-5","doi-asserted-by":"crossref","unstructured":"G\u00f3mez-Chova, L., Jenssen, R., & Camps-Vails, G. (2011, July). Kernel entropy component analysis in remote sensing data clustering. In Proceedings of the 2011 IEEE International Geoscience and Remote Sensing Symposium (IGARSS) (pp. 3728-3731). IEEE.","DOI":"10.1109\/IGARSS.2011.6050035"},{"key":"ijmdem.2013100101-6","doi-asserted-by":"crossref","unstructured":"Guan, L., Wang, Y., & Tie, Y. (2009, June). Toward natural and efficient human computer interaction. In Proceedings of the IEEE International Conference on Multimedia and Expo (ICME 2009) (pp. 1560-1561). IEEE.","DOI":"10.1109\/ICME.2009.5202807"},{"key":"ijmdem.2013100101-7","doi-asserted-by":"publisher","DOI":"10.1504\/IJMIS.2010.035969"},{"key":"ijmdem.2013100101-8","doi-asserted-by":"crossref","unstructured":"Gurban, M., Thiran, J. P., Drugman, T., & Dutoit, T. (2008). Dynamic modality weighting for multi-stream hmms inaudio-visual speech recognition. In Proceedings of the 10th International Conference on Multimodal Interfaces (pp. 237-240). ACM.","DOI":"10.1145\/1452392.1452442"},{"key":"ijmdem.2013100101-9","doi-asserted-by":"crossref","unstructured":"Huahu, X., Jian, Y., & Jue, G. (2010, October). Application of speech emotion recognition in intelligent household robot. In Proceedings of the 2010 International Conference on Artificial Intelligence and Computational Intelligence (AICI) (Vol. 1, pp. 537-541). IEEE.","DOI":"10.1109\/AICI.2010.118"},{"key":"ijmdem.2013100101-10","doi-asserted-by":"crossref","unstructured":"Jenssen, R. (2009). Information theoretic learning and kernel methods. In Information theory and statistical learning (pp. 209-230). Springer US.","DOI":"10.1007\/978-0-387-84816-7_9"},{"key":"ijmdem.2013100101-11","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.100"},{"key":"ijmdem.2013100101-12","doi-asserted-by":"crossref","unstructured":"Jenssen, R. (2011, September). Kernel entropy component analysis: New theory and semi-supervised learning. In Proceedings of the 2011 IEEE International Workshop on Machine Learning for Signal Processing (MLSP) (pp. 1-6). IEEE.","DOI":"10.1109\/MLSP.2011.6064626"},{"key":"ijmdem.2013100101-13","first-page":"633","author":"R.Jenssen","year":"2006","journal-title":"Kernel maximum entropy data transformation and an enhanced spectral clustering algorithm"},{"key":"ijmdem.2013100101-14","doi-asserted-by":"publisher","DOI":"10.1504\/IJBM.2009.027303"},{"key":"ijmdem.2013100101-15","doi-asserted-by":"publisher","DOI":"10.1016\/j.image.2012.04.001"},{"key":"ijmdem.2013100101-16","doi-asserted-by":"crossref","unstructured":"Lin, Y. L., & Wei, G. (2005, August). Speech emotion recognition based on HMM and SVM. In Proceedings of 2005 International Conference on Machine Learning and Cybernetics (Vol. 8, pp. 4898-4901). IEEE.","DOI":"10.1109\/ICMLC.2005.1527805"},{"key":"ijmdem.2013100101-17","doi-asserted-by":"crossref","unstructured":"Martin, O., Kotsia, I., Macq, B., & Pitas, I. (2006). The enterface\u201905 audio-visual emotion database. In Proceedings of the 22nd International Conference on Data Engineering Workshops (pp. 8-8). IEEE.","DOI":"10.1109\/ICDEW.2006.145"},{"key":"ijmdem.2013100101-18","doi-asserted-by":"publisher","DOI":"10.1016\/S0167-6393(03)00099-2"},{"key":"ijmdem.2013100101-19","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177704472"},{"key":"ijmdem.2013100101-20","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2003.817150"},{"key":"ijmdem.2013100101-21","doi-asserted-by":"publisher","DOI":"10.1016\/S0167-8655(03)00079-5"},{"key":"ijmdem.2013100101-22","unstructured":"RRNYI. A. (1961). On measures of entropy and information. In Proceedings of the Fourth Berkeley Symposium on Mathematical Statistics and Probability (pp. 547-561)."},{"key":"ijmdem.2013100101-23","doi-asserted-by":"crossref","unstructured":"Sch\u00f6lkopf, B., Smola, A., & M\u00fcller, K. R. (1997). Kernel principal component analysis. In Proceedings of the International Conference on Artificial Neural Networks (ICANN'97) (pp. 583-588). Springer Berlin Heidelberg.","DOI":"10.1007\/BFb0020217"},{"key":"ijmdem.2013100101-24","doi-asserted-by":"publisher","DOI":"10.1145\/584091.584093"},{"key":"ijmdem.2013100101-25","doi-asserted-by":"publisher","DOI":"10.2478\/v10006-011-0043-9"},{"key":"ijmdem.2013100101-26","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2010.2057231"},{"key":"ijmdem.2013100101-27","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2008.927665"},{"key":"ijmdem.2013100101-28","doi-asserted-by":"crossref","unstructured":"Wang, Y., Guan, L., & Venetsanopoulos, A. N. (2011, May). Kernel cross-modal factor analysis for multimodal information fusion. In Proceedings of the 2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (pp. 2384-2387). IEEE.","DOI":"10.1109\/ICASSP.2011.5946963"},{"key":"ijmdem.2013100101-29","doi-asserted-by":"crossref","unstructured":"Wang, Y., Guan, L., & Venetsanopoulos, A. N. (2011, July). Audiovisual emotion recognition via cross-modal association in kernel space. In Proceedings of the 2011 IEEE International Conference on Multimedia and Expo (ICME) (pp. 1-6). IEEE.","DOI":"10.1109\/ICME.2011.6011949"},{"key":"ijmdem.2013100101-30","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2012.2189550"},{"key":"ijmdem.2013100101-31","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-21596-4_15"},{"key":"ijmdem.2013100101-32","doi-asserted-by":"crossref","unstructured":"Xu, X., & Mu, Z. (2007, August). Feature fusion method based on KCCA for ear and profile face based multimodal recognition. In Proceedings of the 2007 IEEE International Conference on Automation and Logistics (pp. 620-623). IEEE.","DOI":"10.1109\/ICAL.2007.4338638"},{"key":"ijmdem.2013100101-33","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2008.52"}],"container-title":["International Journal of Multimedia Data Engineering and Management"],"original-title":[],"language":"ng","link":[{"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/www.igi-global.com\/viewtitle.aspx?TitleId=103008","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,5,1]],"date-time":"2025-05-01T21:30:56Z","timestamp":1746135056000},"score":1,"resource":{"primary":{"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/services.igi-global.com\/resolvedoi\/resolve.aspx?doi=10.4018\/ijmdem.2013100101"}},"subtitle":[""],"short-title":[],"issued":{"date-parts":[[2013,10,1]]},"references-count":34,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2013,10]]}},"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/doi.org\/10.4018\/ijmdem.2013100101","relation":{},"ISSN":["1947-8534","1947-8542"],"issn-type":[{"value":"1947-8534","type":"print"},{"value":"1947-8542","type":"electronic"}],"subject":[],"published":{"date-parts":[[2013,10,1]]}}}