Publications
2024
Kipyatkova I., Kagirov I., Dolgushin M., Rodionova A. Towards a Livvi-Karelian End-to-End ASR System // Lecture Notes in Computer Science, SPECOM-2024. 2024. vol. 15299. pp. 57-68.
Guseva D., Mitrofanova O., Dolgushin M. Human and Machine Keyphrase Perception in Russian Text and Speech // Lecture Notes in Computer Science, SPECOM-2024. 2024. vol. 15299. pp. 265-280.
Kosulin K., Karpov A. A Survey of Masked Face Recognition Methods and Corpora/Data // Springer Geography. IMS-2022. 2024. pp. 27-37.
Ivanko D., Ryumin D., Markitantov M. End-to-End Visual Speech Recognition for Human-Robot Interaction // In Proc. of the AIP Conference. 2024. vol. 3021. pp. 82-90.
Dvoynikova A. A., Karpov A. A. Method of creating multimodal databases for audiovisual analysis of engagement and emotions of virtual communication participants // Journal of Instrument Engineering. 2024. vol. 67. no. 11. pp. 984–993
2023
Ryumina E., Ryumin D., Markitantov M., Kaya H., Karpov A. Multimodal Personality Traits Assessment (MuPTA) Corpus: The Impact of Spontaneous and Read Speech// In Proc. of the 24th International Conference INTERSPEECH-2023. 2023. pp. 4049–4053.
Karpov A., Samudravijaya K., Deepak K.T., Hegde R.M., Agrawal S.S., Prasanna S.R.M. SPECOM 2023 Preface. Lecture Notes in Computer Science// In Proc. of the 25th International Conference on Speech and Computer SPECOM-2023. LNAI. 2023. vol. 14338/14339.
Ivanko D., Ryumin D., Karpov A. A Review of Recent Advances on Deep Learning Methods for Audio-Visual Speech Recognition // Mathematics. 2023. vol. 11(12). no. 2665.
Kipyatkova I., Kagirov I. Deep Models for Low-Resourced Speech Recognition: Livvi-Karelian Case // Mathematics. 2023. vol. 11(18). no. 3814.
Ryumin D., Ryumina E., Ivanko D. EMOLIPS: Towards Reliable Emotional Speech Lip-Reading // Mathematics. 2023. vol. 11(23). no. 4787.
Ryumina E., Markitantov M., Karpov A. Multi-Corpus Learning for Audio–Visual Emotions and Sentiment Recognition // Mathematics. 2023. vol. 11(16). no. 3519.
Ryumin D., Ivanko D., Ryumina E. Audio-Visual Speech and Gesture Recognition by Sensors of Mobile Devices // Sensors. 2023. vol. 23(4). no. 2284.
Axyonov A.A., Ryumina E.V., Ryumin D.A., Ivanko D.V., Karpov A.A. Neural network-based method for visual recognition of driver’s voice commands using attention mechanism // Scientific and Technical Journal of Information Technologies, Mechanics and Optics. 2023. vol. 23. no. 4. pp. 767–775.
Kipyatkova I., Kagirov I. Automatic speech recognition system for Karelian // Information and Control Systems. 2023. vol. 3. pp. 16-25.
Velichko A., Karpov A. An approach and software system for integral analysis of destructive paralinguistic phenomena in colloquial speech // Information and Control Systems. 2023. vol. 4. pp. 2-11.