Learn Before
True/False

True or False: On the RealToxicityPrompts benchmark, GPT-4 produces toxic generations at a higher rate than GPT-3.5.

0

1

Updated 2026-09-11

Tags

Prep Sessions

Frontier Model Dynamics: Scaling Laws, Calibration, and Post-Training Alignment @ University of Michigan - Ann Arbor

Ch.2 Post-Training Analysis and Safety - Frontier Model Dynamics: Scaling Laws, Calibration, and Post-Training Alignment @ University of Michigan - Ann Arbor

Safety Metrics and Refusal Behavior - Frontier Model Dynamics: Scaling Laws, Calibration, and Post-Training Alignment @ University of Michigan - Ann Arbor