Enhanced Image Caption Generator Using Deep Learning
摘要
The present study introduces a new system named “Enhanced Image Caption Generator Using Deep Learning” that employs a revolutionary method to automatically provide meaningful descriptions for various photos. The suggested approach distinguishes itself from traditional rule-based or shallow learning techniques by integrating Recurrent Neural Networks (RNNs), specifically Long Short-Term Memory (LSTM) networks, for sequence generation and Convolutional Neural Networks (CNNs) for extracting image data. The primary objectives are to enhance comprehension of context, offer a user-friendly interface, and assess performance using the widely accepted BLEU metric.