An Empirical Study of Different Deep Learning Models for Image to Text Generation |
Author(s): |
| Mohit Dalal , SKITM BAHADURGARH; Minakshi Arora, SKITM BAHADURGARH |
Keywords: |
| Machine Learning, Deep Learning, Encoder, Decoder |
Abstract |
|
The technique of creating a written description of an image from its objects and their characteristics is known as image captioning. There are a lot of practical uses for this procedure. One crucial one would be to store a picture's captions so that, in the future, the image may be conveniently accessed based only on this description. The authors have demonstrated the various image captioning techniques in this research. The numerous models for improved performance are also examined, in addition to several approaches for image captioning. Since there are many different kinds of photographs from different domains in real life, this can be utilized to find the right method or strategy for captioning images. Finally, the authors have included evaluation criteria, statistics, and future areas for research. |
Other Details |
|
Paper ID: IJSRDV12I30249 Published in: Volume : 12, Issue : 3 Publication Date: 01/06/2024 Page(s): 257-259 |
Article Preview |
|
|
|
|
