Learn Before
Short Answer

Inputs and Targets in a Captioning Model

Question: In a model that turns an image into a sentence, what do x and y stand for?

Sample answer: x is the image given to the model, and y is the caption text the model produces.

Key points:

  • x names the image input.
  • y names the generated caption.

Rubric: The answer is correct if it assigns x to the image and y to the caption or text description.

0

1

Updated 2026-08-12

Contributors are:

Who are from:

Tags

Python Programming Language

Data Science

Machine Learning

Deep Learning

Supervised Learning

Dive into Deep Learning @ D2L

Machine Learning Strategy

Machine Learning Yearning @ DeepLearning.AI