Bridging Text and Speech for Emotion Understanding: An Explainable Multimodal Transformer Fusion Framework with Unified Audio–Text Attribution
Abstract
Share and Cite
Pandey, A.; Singh, J.; Kaur, M. Bridging Text and Speech for Emotion Understanding: An Explainable Multimodal Transformer Fusion Framework with Unified Audio–Text Attribution. J. Intell. 2025, 13, 159. https://doi.org/10.3390/jintelligence13120159
Pandey A, Singh J, Kaur M. Bridging Text and Speech for Emotion Understanding: An Explainable Multimodal Transformer Fusion Framework with Unified Audio–Text Attribution. Journal of Intelligence. 2025; 13(12):159. https://doi.org/10.3390/jintelligence13120159
Chicago/Turabian StylePandey, Ashutosh, Jasmeet Singh, and Maninder Kaur. 2025. "Bridging Text and Speech for Emotion Understanding: An Explainable Multimodal Transformer Fusion Framework with Unified Audio–Text Attribution" Journal of Intelligence 13, no. 12: 159. https://doi.org/10.3390/jintelligence13120159
APA StylePandey, A., Singh, J., & Kaur, M. (2025). Bridging Text and Speech for Emotion Understanding: An Explainable Multimodal Transformer Fusion Framework with Unified Audio–Text Attribution. Journal of Intelligence, 13(12), 159. https://doi.org/10.3390/jintelligence13120159

