Generated image description in the form of coherent paragraphs using Densecap(CNN for Dense Captioning) and Hierarchical RNN and Model was able to perform as good as state-of-art in terms of metrics inclined towards human-like sentences.
python deep-learning image-captioning scene-description visualgenome-dataset paragraph-generation hrnn densecap
-
Updated
May 21, 2020 - Jupyter Notebook