Order the steps involved in evaluating a policy response using a rule-based reward model (RBRM).
0
1
Tags
Prep Sessions
Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Ch.3 Model Alignment and Safety - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Model-Assisted Safety and Rule-Based Reward Models - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Related
Match each input processed by a rule-based reward model (RBRM) to its accurate description.
Order the steps involved in evaluating a policy response using a rule-based reward model (RBRM).
Identify the rubric category that the RBRM classifier should assign to this output, and explain what downstream action this classification enables the system to take.