A Framework for Generating Text from Images Using Deep Learning |
Author(s): |
| Mohit Dalal , SKITM BAHADURGARH; Minakshi Arora, SKITM BAHADURGARH |
Keywords: |
| Deep Learning, CNN, RNN |
Abstract |
|
Image captioning is the process of creating a description for a picture. Identification of the key elements, their characteristics, and their interactions within an image are necessary for image captioning. There are a lot of practical uses for this procedure. One important one would be to store a picture's captions so that, in the future, the image may be conveniently accessed based only on this description. Our goal in this survey paper is to provide a thorough analysis of the state-of-the-art deep learning-based picture captioning methods. We go into the theoretical underpinnings of the methods to evaluate their effectiveness, drawbacks, and strengths. Additionally, we go over the evaluation measures and datasets that are frequently employed in deep learning-based automatic image captioning. |
Other Details |
|
Paper ID: IJSRDV12I30250 Published in: Volume : 12, Issue : 3 Publication Date: 01/06/2024 Page(s): 260-262 |
Article Preview |
|
|
|
|
