{"id":9689,"date":"2026-08-05T21:56:55","date_gmt":"2026-08-05T13:56:55","guid":{"rendered":"\/jase\/?post_type=tkuisotope&#038;p=9689"},"modified":"2026-08-21T14:59:15","modified_gmt":"2026-08-21T06:59:15","slug":"jase-202611-34-019","status":"publish","type":"tkuisotope","link":"\/jase\/?tkuisotope=jase-202611-34-019","title":{"rendered":"Virtual Reality Simulations for Immersive Training in Cross- Cultural Communication for Bilingual Broadcasters"},"content":{"rendered":"\n<div class=\"wp-block-tkuwpbs5-bs5-row row article-info\">\n<div class=\"wp-block-tkuwpbs5-bs5-column col-md-3 align-self-start\">\n<p><i class=\"fa fa-folder\" aria-hidden=\"true\"><\/i>&nbsp;<a href=\"\/jase\/?page_id=807\" data-type=\"page\" data-id=\"807\">2026<\/a><\/p>\n<\/div>\n\n\n\n<div class=\"wp-block-tkuwpbs5-bs5-column col-md-3 align-self-start\">\n<p><i class=\"fa fa-folder-open\" aria-hidden=\"true\"><\/i>&nbsp;<a href=\"\/jase\/?page_id=9439\" data-type=\"page\" data-id=\"9439\">Volume 34<\/a><\/p>\n<\/div>\n\n\n\n<div class=\"wp-block-tkuwpbs5-bs5-column col-md-6 align-self-start\">\n<div class=\"wp-block-tkuwpbs5-bs5-div dv_publish\" data-aos=\"normal\"><div class=\"wp-block-post-date\"><time datetime=\"2026-08-05T21:56:55+08:00\">2026-08-05<\/time><\/div><\/div>\n<\/div>\n<\/div>\n\n\n\n<div class=\"wp-block-tkuwpbs5-bs5-row row\">\n<div class=\"wp-block-tkuwpbs5-bs5-column col-md-5 align-self-start\">\n<div class=\"wp-block-tkuwpbs5-bs5-div au-ol\" data-aos=\"normal\">\n<p>Luyang Xu<a href=\"mailto:xuluyang@cucn.edu.cn\"><i class=\"fa fa-envelope\"><\/i><\/a><\/p>\n\n\n\n<p style=\"font-size:14px\">School of International Communication, Communication University of China, Nanjing 211172, Jiangsu, China<\/p>\n<\/div>\n\n\n\n<div class=\"wp-block-tkuwpbs5-bs5-div\" style=\"margin-top:var(--wp--preset--spacing--40)\" data-aos=\"normal\">\n<p>Received: May 20, 2026<br>Accepted:&nbsp;June 27, 2026<br>Publication Date:&nbsp;August 05, 2026<\/p>\n<\/div>\n<\/div>\n\n\n\n<div class=\"wp-block-tkuwpbs5-bs5-column col-md-7 align-self-start clk=\u5716\u7247\"><img decoding=\"async\" src=\"\/jase\/wp-content\/uploads\/2026\/08\/34_019.jpg\" class=\"img-fluid img-fluid mx-auto d-block\" alt=\"\u4e0a\u50b3\u5716\u7247\">\n\n\n<p class=\"has-text-align-center\">Architecture of the Multimodal Transformer for Cross-Cultural Emotion Recognition<\/p>\n<\/div>\n<\/div>\n\n\n\n<p class=\"has-small-font-size\"><i class=\"fab fa-creative-commons\"><\/i>&nbsp;<strong>Copyright&nbsp;<\/strong>The Author(s). This is an open access article distributed under the terms of the&nbsp;<a rel=\"noreferrer noopener\" href=\"https:\/\/creativecommons.org\/licenses\/by\/4.0\/\" target=\"_blank\">Creative Commons Attribution&nbsp;License (CC BY 4.0)<\/a>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are cited.<\/p>\n\n\n\n<p>Download Citation:&nbsp; <a href=\"\/jase\/wp-content\/uploads\/2026\/08\/V34.0019.txt\" data-type=\"attachment\" data-id=\"9755\" target=\"_blank\" rel=\"noreferrer noopener\">BibTeX <\/a>| <a rel=\"noreferrer noopener\" href=\"http:\/\/dx.doi.org\/10.6180\/jase.202611_34.019\" target=\"_blank\">http:\/\/dx.doi.org\/10.6180\/jase.202611_34.019<\/a>&nbsp;&nbsp;<\/p>\n\n\n\n<p class=\"btn btn-primary article-btn\"><a href=\"\/jase\/wp-content\/uploads\/2026\/08\/JASE_2026_1339.pdf\" data-type=\"link\" data-id=\"\/jase\/wp-content\/uploads\/2026\/08\/JASE_2026_1339.pdf\" target=\"_blank\" rel=\"noreferrer noopener\">Download PDF<\/a><\/p>\n\n\n\n<div style=\"height:24px\" aria-hidden=\"true\" class=\"wp-block-spacer\"><\/div>\n\n\n\n<p>Cross-cultural awareness and communication training for bilingual broadcasters should place the emphasis on high-engagement, tailored and transactive learning environments that support multilingual communication, socio-contextual processing, and cultural awareness. The objective is to establish an artificial intelligence (AI)-based virtual reality (VR) solution for immersive bilingual communication by utilizing multimodal emotion<br>analysis and real-time dynamic feedback. The proposed framework combines two publicly available multimodal datasets CAMEO and MELD with a systematically controlled human study dataset at the core collected from 110 bilingual speakers. In general, the procedures for this study are multimodal data acquisition, MFCC based speech feature extraction, multilingual BERT (mBERT) text embedding, bilingual code-switch processing, emotion encoding, and stratified data splitting. A multimodal transformer based on attention to jointly represent text and speech data for contextual emotion recognition and bilingual interaction assessment. The framework includes cross-cultural adaptation and AI-based real-time feedback in the immersive VR communication settings. Simulation results indicate that the proposed model achieves the best performance with an Accuracy of 98.63%,<br>a Precision of 98.62%, a Recall of 98.62%, and an F1-Score of 98.62%. The results suggest that the proposed framework surpass both the CNN, the BiLSTM, the transformer encoder, as well as other existing multimodal methods in the task of emotion recognition. In this work we report on real-time inference with an average response latency of 8.37ms that facilitates natural immersive interaction.<\/p>\n\n\n\n<p><em>Keywords:&nbsp;Virtual Reality, Multimodal Transformer, Mel-Frequency Cepstral Coefficients, Bilingual Broadcasting, Bidirectional Encoder Representations from Transformers<\/em><\/p>\n\n\n\n<div style=\"height:2rem\" aria-hidden=\"true\" class=\"wp-block-spacer\"><\/div>\n\n\n\n<div class=\"wp-block-tkuwpbs5-bs5-div ref_ol\" data-aos=\"normal\">\n<ol>\n<li>[1] Y.-J. Wu, T.-J. Ding, J.-C. Hsu, K.-L. Ou, and W. Tarng, (2025) \u201cExploratory Learning of Amis Indigenous Culture and Local Environments Using Virtual Reality and Drone Technology\u201d ISPRS International Journal of Geo-Information 14(11): 441. DOI: 10.3390\/ijgi14110441.<\/li>\n<li>[2] Z. Yang, (2025) \u201cDesign of a Visual Communication System for Animated Characters in Virtual Reality Using Sobel Edge Detection and Motion Capture\u201d Journal of Applied Science and Engineering 29(1): 235\u2013243. DOI: 10.6180\/jase.202601_29(1).0023.<\/li>\n<li>[3] A. K. Jumani, K. Kumar, and M. A. Chhajro, (2021) \u201cSystematic Analysis of Virtual Reality &amp; Augmented Reality\u201d International Journal of Information Engineering &amp; Electronic Business 13(1): 1\u201312. DOI: 10.5815\/ijieeb.2021.01.04.<\/li>\n<li>[4] A. A. Laghari, V. V. Estrela, H. Li, Y. Shoulin, A. A. Khan, M. S. Anwar, A. Wahab, and K. Bouraqia, (2024) \u201cQuality of Experience Assessment in Virtual\/Augmented Reality Serious Games for Healthcare: A Systematic Literature Review\u201d Technology and Disability 36(1\u20132): 17\u201328. DOI: 10.3233\/TAD-230035.<\/li>\n<li>[5] A. K. Jumani, J. Shi, A. A. Laghari, V. V. Estrela, G. A. Sampedro, A. Almadhor, N. Kryvinska, and A. U. Nabi, (2024) \u201cQuality of Experience That Matters in Gaming Graphics: How to Blend Image Processing and Virtual Reality\u201d Electronics 13(15): 2998. DOI: 10.3390\/electronics13152998.<\/li>\n<li>[6] A. K. Jumani, J. Shi, A. A. Laghari, M. A. Amin, A. U. Nabi, K. Narwani, and Y. Zhang, (2025) \u201cQuality of Experience (QoE) in Cloud Gaming: A Comparative Analysis of Deep Learning Techniques via Facial Emotions in a Virtual Reality Environment\u201d Sensors 25(5): 1594. DOI: 10.3390\/s25051594.<\/li>\n<li>[7] A. A. Laghari, S. Shahid, R. Yadav, S. Karim, A. Khan, H. Li, and Y. Shoulin, (2023) \u201cThe State of Art and Review on Video Streaming\u201d Journal of High Speed Networks 29(3): 211\u2013236. DOI: 10.3233\/JHS-222087.<\/li>\n<li>[8] M. R. Sareddy and V. Kumar. \u201cVirtual Reality Meets AI: Revolutionizing Physical Education with LiDAR and RL\u201d. In: 2025 8th International Conference on Electronics, Materials Engineering &amp; Nano-Technology (IEMEnTech). IEEE, 2025, 1\u20136. DOI: 10.1109\/IEMEnTech65115.2025.10959637.<\/li>\n<li>[9] M. Maciejewski, J. Koco\u0144, W. Oleksy, M. Kope\u0107, and W. Kazienko. amu-cai\/CAMEO Dataset. Accessed: May 9, 2026. 2026.<\/li>\n<li>[10] S. Poria, D. Hazarika, N. Majumder, G. Naik, E. Cambria, and R. Mihalcea. \u201cMELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations\u201d. In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics. 2019. DOI: 10.18653\/v1\/P19-1050.<\/li>\n<li>[11] B. Xie, M. Sidulova, and C. H. Park, (2021) \u201cRobust Multimodal Emotion Recognition from Conversation with Transformer-Based Crossmodality Fusion\u201d Sensors 21(14): 4913. DOI: 10.3390\/s21144913.<\/li>\n<li>[12] C. Qu et al., (2025) \u201cEnhancing Emotion Recognition in Virtual Reality: A Multimodal Dataset and a Temporal Emotion Detector\u201d Frontiers in Psychology 16(2): 1709943. DOI: 10.3389\/fpsyg.2025.1709943.<\/li>\n<li>[13] H. Wang and M. Wang, (2026) \u201cSubject-Independent Multimodal Interaction Modeling for Joint Emotion and Immersion Estimation in Virtual Reality\u201d Symmetry 18(3): 451. DOI: 10.3390\/sym18030451.<\/li>\n<\/ol>\n<\/div>\n\n\n\n<p><\/p>\n","protected":false},"author":3,"template":"wp-custom-template-detail-4-aricles","meta":{"_uag_custom_page_level_css":""},"categories":[12,1682,6],"tags":[1701],"acf":[],"uagb_featured_image_src":[],"uagb_author_info":{"display_name":"\u6797\u923a\u6db5","author_link":"\/jase\/?author=3"},"uagb_comment_info":0,"uagb_excerpt":"&nbsp;Copyright&nbsp;The Author(s). This is an open access article distributed under the terms of the&nbsp;Creative Commons Attribution&nbsp;License (CC BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are cited. Download Citation:&nbsp; BibTeX | http:\/\/dx.doi.org\/10.6180\/jase.202611_34.019&nbsp;&nbsp; Download PDF Cross-cultural awareness and communication training for bilingual broadcasters should place the&hellip;","_links":{"self":[{"href":"\/jase\/index.php?rest_route=\/wp\/v2\/tkuisotope\/9689"}],"collection":[{"href":"\/jase\/index.php?rest_route=\/wp\/v2\/tkuisotope"}],"about":[{"href":"\/jase\/index.php?rest_route=\/wp\/v2\/types\/tkuisotope"}],"author":[{"embeddable":true,"href":"\/jase\/index.php?rest_route=\/wp\/v2\/users\/3"}],"wp:attachment":[{"href":"\/jase\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=9689"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"\/jase\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=9689"},{"taxonomy":"post_tag","embeddable":true,"href":"\/jase\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=9689"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}