Piclingo: Multilingual Image Caption Generator
摘要
In an era characterized by the exponential growth of digital content, the demand for effective communication across diverse linguistic landscapes has become increasingly pronounced. Multilingual Caption Generator, is a cutting-edge solution designed to transcend language barriers by automating the generation of image descriptions in multiple languages. Leveraging state-of-the-art techniques, our model incorporates a combination of LSTM, ReLU, RNN, and dense layers to effectively generate descriptive captions for images. Feature extraction is accomplished using the VGG16 model, ensuring a robust representation of visual content. The system is designed to dynamically translate captions from English to multiple languages, enhancing accessibility. Experimental results demonstrate the efficiency of the proposed approach in producing contextually relevant and linguistically accurate captions across different languages. Following the caption generation stage, our proposed model's effectiveness was evaluated using BLEU Scores. The model's versatility and performance make it a promising solution for overcoming linguistic barriers in image understanding and accessibility.