Enhancing accessibility: deep learning-based image description for individuals with visual impairments
Abstrak
Technological developments in artificial intelligence, namely in the area of deep learning, have created new avenues for enhancing accessibility for those with visual impairments. In order to improve the capacity of people who are blind or visually impaired to understand and interact with visual material, this research investigates the creation and use of deep learning-based image description systems. We provide a comprehensive method that uses recurrent neural networks (RNNs) to generate natural language descriptions and convolutional neural networks (CNNs) and Autoencoders for extracting picture features. Our technology automatically creates comprehensive, context-aware descriptions of photographs by incorporating these models, giving users a better knowledge of their surroundings. We show the accuracy and reliability of the system on a wide range of photos through comprehensive testing. According to our research, deep learning-based picture description systems and converting the description in audio and making a promise to empower people who are visually impaired and foster diversity in the digital sphere.
Kata Kunci
Cari jurnal yang tepat untuk naskah Anda
MatchMind AI mencocokkan abstrak naskah Anda dengan ribuan jurnal terakreditasi dan menampilkan rekomendasi terbaik beserta alasannya.
Coba MatchMindLihat profil lengkap jurnal ini
Waktu review, biaya APC, statistik sitasi, indeksasi Scopus, dan banyak lagi.
Buka Indonesian Journal of Electrical Engineering and Computer ScienceArtikel ini juga tersedia di situs resmi jurnal.
