How did GPT-4's performance compare to GPT-3.5 on the simulated Uniform Bar Examination?
0
1
Tags
Prep Sessions
Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Ch.2 Model Scaling and Capability Evaluation - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Academic and Professional Exam Benchmarks - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Test Set Contamination Analysis - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor
Related
How did GPT-4's performance compare to GPT-3.5 on the simulated Uniform Bar Examination?
Across the majority of academic and professional exams designed for humans, what benchmark level of performance did GPT-4 achieve?
The academic and professional examinations used to evaluate GPT-4 share which design characteristic?
Aside from the Uniform Bar Examination, which four standardized examination programs are explicitly cited as part of GPT-4's evaluation?
Origin of GPT-4 Exam Capabilities in Pre-training