Show and Tell

E899056

Show and Tell is a neural network-based image captioning model developed by Google that automatically generates natural language descriptions for images.

All labels observed (2)

Label Occurrences
Show & Tell 1
Show and Tell canonical 1

How this entity was disambiguated

Statements (47)

Predicate Object
instanceOf deep learning model
image captioning model
neural network
achieves state-of-the-art performance on MSCOCO (2015)
approach end-to-end training
maximum likelihood training
author Alexander Toshev
Dumitru Erhan
Oriol Vinyals
Samy Bengio
basedOn convolutional neural network
encoder-decoder architecture
recurrent neural network
sequence-to-sequence model
captionStyle descriptive sentences
developer Google
Google Research
domain computer vision
multimodal learning
natural language processing
evaluationMetric BLEU
CIDEr
METEOR
featureExtractor Inception CNN
implementedIn TensorFlow (research implementation)
linked to: TensorFlow
influenced Neural image captioning research
Show, Attend and Tell
input image
language English
learningType supervised learning
modality vision-to-language
optimizationAlgorithm stochastic gradient descent
organization Google Brain
output natural language caption
paperTitle Show and Tell: A Neural Image Caption Generator
publicationVenue CVPR
publicationYear 2015
relatedTo encoder-decoder neural networks for machine translation
task automatic image captioning
natural language description generation
trainedOn Flickr30k dataset
linked to: Flickr30k

Flickr8k dataset
linked to: Flickr8k

MSCOCO dataset
uses CNN encoder
LSTM
linked to: LSTM networks

RNN decoder
word embeddings

How these facts were elicited

Referenced by (2)

Full triples — surface form annotated when it differs from this entity's canonical label.

African Giant hasPart Show & Tell
linked to: Show and Tell