Outline

Ingegneria Sismica

Ingegneria Sismica

An Interactive English Listening and Speaking Self-study System Based on Computer Vision

Author(s): Shouren Wu1, Qin Zhang2
1School of Foreign Languages, Shaoyang University, Shaoyang 422000, Hunan, China
2Department of Basic Courses, Hunan Polytechnic of Water Resources and Electric Power, Changsha 410131, Hunan, China
Wu, Shouren. and Zhang, Qin. “An Interactive English Listening and Speaking Self-study System Based on Computer Vision.” Ingegneria Sismica Volume 43 Issue 2: 1-20, doi:10.65102/is2026936.

Abstract

With the development of modern society, many people have become accustomed to life under IoT and machine learning, and now interaction has become an important factor in affecting English listening and speaking skills. The more interaction the learners have among themselves, the more active they will be and the better the results will be. Vision is a way for people to look at the world and know about it. People use their eyes and brains to acquire, process and understand visual information. At present, the series of problems for mobile English listening and speaking learning are still being addressed, such as non-optimal functions and poor operability. Based on Smart Sensing and Communication, we fully utilize computer vision technology to design a set of computer vision modules for an interactive English self-study system, and have completed user demand analysis, overall architecture design, functional applications, etc. The System will be able to play sound and video generally. Add images, mind maps, art words and other multimedia to increase the learners’ enthusiasm and interest in learning, and provide an intuitive and vivid experience for users. Based on the above experiments, it is clear that the system interaction is more convenient and pleasant for English learners. The system can address some deficiencies of the existing system by adding new functions to the current system, improving the user experience, and achieving an effective voice recognition accuracy of over 95 per cent. The studies in this paper provide necessary support for the application of both IoT networks and machine learning.

Keywords
Computer Vision; Interactive English Listening and Speaking; Self-study System; Recognition Rate

Related Articles

Qianwen Xiong1, Yuhong Chen1
1Guangzhou University of Chinese Medicine, School of Pharmaceutical Medicine, Guangzhou,Guangdong,China,510006
Zhihao Jiang1,2, Limi Chen1,2, Jing Yang1
1Hainan Vocational University of Science and Technology, Haikou 571126, China
2Institute for Mathematical Research, Universiti Putra Malaysia, Serdang 43400, Malaysia
Limi Chen1,2, Zhihao Jiang1,2, Jing Yang1
1Hainan Vocational University of Science and Technology, Haikou 571126, China
2Institute for Mathematical Research, Universiti Putra Malaysia, Serdang 43400, Malaysia
Hui Yuan1, Minjie Chai2, Siqing Xu1, Jinsong Li1, Jinwan Zheng1
1Electric Power Research Institute, State Grid Shanxi Electric Power Co., Ltd., Taiyuan, 030001, Shanxi, China
2Jincheng Power Supply Branch, State Grid Shanxi Electric Power Co., Ltd., Jincheng, 048000, Shanxi, China
Yanhan Zhu1,2
1China Academy of Cultural Heritage, Chaoyang District, 100029, Beijing, China
2Beijing University of Civil Engineering and Architecture, Xicheng District, 100044, Beijing, China