1Cademy - Training Auxiliary Parameters with a Fixed Transformer Model

Learn Before

The Pre-training and Fine-tuning Paradigm

Activity (Process)

Training Auxiliary Parameters with a Fixed Transformer Model

A training methodology where the parameters of a pre-trained Transformer model are held constant, or 'frozen'. During this process, only a set of newly introduced, learnable parameters are updated. This approach allows for model adaptation while preserving the original, powerful representations of the base Transformer.

Updated 2025-10-10

Contributors are:

Who are from:

References

Reference of Foundations of Large Language Models Course

Learn After

A research team wants to adapt a very large, pre-trained language model (with billions of parameters) to perform a new, specialized task, such as classifying medical reports. The team's primary constraint is a very limited computational budget, which makes it infeasible to update all of the model's original parameters. Which of the following training strategies best resolves this constraint while still effectively adapting the model to the new task?
Evaluating a Model Adaptation Strategy
When adapting a large, pre-trained model by introducing and training only a small set of new parameters, the original weights of the base model are also fine-tuned, but with a much smaller learning rate to prevent drastic changes.

Learn Before

Related

Learn After