查看文章

Visual speech recognition for small scale dataset using VGG16 convolution neural network

作者

Sudarshan Patilkulkarni

发表日期

2021/8

期刊

Multimedia Tools and Applications

卷号

期号

页码范围

28941-28952

出版商

Springer US

简介

Visual speech recognition is a method that comprehends speech from speakers lip movements and the speech is validated only by the shape and lip movement. Implementation of this practice not only helps people with hearing impaired but also can be used for professional lip reading whose application can be seen in crime and forensics. It plays a crucial role in aforementioned domains, as normal person’s speech will be converted to text. Here, it is proposed to enhance the visual speech recognition technique from the video. The dataset was created and the same was used for implementation and verification. The aim of the approach was to recognize words only from the lip movement using video in the absence of audio and this mostly helps to extract words from a video without audio that helps in forensic and crime analysis. The proposed method employs VGG16 pre trained Convolutional Neural Network …

引用总数

被引用次数：31

20212022202320243 9 12 7

学术搜索中的文章

Visual speech recognition for small scale dataset using VGG16 convolution neural network

S Patilkulkarni - Multimedia Tools and Applications, 2021

被引用次数：31 相关文章所有 4 个版本