Enhancing image captioning performance based on efficientnet B0 model and transformer encoder-decoder

In recent years, improvements in natural language processing and computer vision have come together to provide automatic image caption generation. Image captioning is the process of creating a description for an image. Captioning an image needs the recognition of significant items, their properties,...

Ausführliche Beschreibung

Gespeichert in:

Bibliographische Detailangaben
Hauptverfasser:	Joshi, Abhisht, Alkhayyat, Ahmed, Gunwant, Harsh, Tripathi, Abhay, Sharma, Moolchand
Format:	Tagungsbericht
Sprache:	eng
Schlagworte:	Algorithms Artificial neural networks Computer vision Encoders-Decoders Image enhancement Machine learning Natural language processing Transformers
Online-Zugang:	Volltext
Tags:	Tag hinzufügen Keine Tags, Fügen Sie den ersten Tag hinzu!

Schreiben Sie den ersten Kommentar!