KEYWORDS: Image retrieval, Image processing, Data modeling, Information technology, Deep learning, Data processing, Performance modeling, Machine learning, Neural networks
The image caption is simply to input an image to the computer model, and the model outputs an accurate text caption according to the image to be described. This paper summarizes image caption methods based on template, retrieval and deep learning. The development history of different methods is introduced respectively. Among these, image labeling methods that use Deep Learning are effective, with the encoding code model being the most commonly used. Then it introduces the data sets and evaluation indicators that are widely used in the field of image caption. Finally, the current challenges of image caption methods are analyzed, and the future development direction is prospected.
Access to the requested content is limited to institutions that have purchased or subscribe to SPIE eBooks.
You are receiving this notice because your organization may not have SPIE eBooks access.*
*Shibboleth/Open Athens users─please
sign in
to access your institution's subscriptions.
To obtain this item, you may purchase the complete book in print or electronic format on
SPIE.org.
INSTITUTIONAL Select your institution to access the SPIE Digital Library.
PERSONAL Sign in with your SPIE account to access your personal subscriptions or to use specific features such as save to my library, sign up for alerts, save searches, etc.