1Cademy - In a common fine-tuning strategy, a prompt and its desired completion are concatenated into a single sequence (e.g., `[prompt_tokens, completion_tokens]`). The language model is then trained on this full sequence, but the training loss is calculated *only* for the models predictions on the completion tokens. What is the most accurate analysis of the primary purpose of this specific loss calculation method?

Learn Before

SFT as Language Model Training on Concatenated Sequences

Multiple Choice

In a common fine-tuning strategy, a prompt and its desired completion are concatenated into a single sequence (e.g., [prompt_tokens, completion_tokens]). The language model is then trained on this full sequence, but the training loss is calculated only for the model's predictions on the completion tokens. What is the most accurate analysis of the primary purpose of this specific loss calculation method?

Updated 2025-09-29

Contributors are:

Who are from:

Learn Before

Related