High Impact Factor : 4.396 icon | Submit Manuscript Online icon |

A Framework for Generating Text from Images Using Deep Learning

Author(s):

Mohit Dalal , SKITM BAHADURGARH; Minakshi Arora, SKITM BAHADURGARH

Keywords:

Deep Learning, CNN, RNN

Abstract

Image captioning is the process of creating a description for a picture. Identification of the key elements, their characteristics, and their interactions within an image are necessary for image captioning. There are a lot of practical uses for this procedure. One important one would be to store a picture's captions so that, in the future, the image may be conveniently accessed based only on this description. Our goal in this survey paper is to provide a thorough analysis of the state-of-the-art deep learning-based picture captioning methods. We go into the theoretical underpinnings of the methods to evaluate their effectiveness, drawbacks, and strengths. Additionally, we go over the evaluation measures and datasets that are frequently employed in deep learning-based automatic image captioning.

Other Details

Paper ID: IJSRDV12I30250
Published in: Volume : 12, Issue : 3
Publication Date: 01/06/2024
Page(s): 260-262

Article Preview

Download Article