Why two evaluators score the same lesson differently — and how to fix it. A practical guide to calibration and inter-rater reliability in teacher evaluation.

Teacher Evaluation Calibration: How to Get Your Evaluators to Agree (and Why It Matters)

EvalScribe
August 13, 2026
Custom HTML/CSS/JavaScript
teacher evaluation calibrationinter-rater reliability teacher evaluationevaluator calibrationnorming session teacher observationhow to calibrate evaluatorsteacher observation reliabilitywhy do evaluators score differentlyrater drift teacher evaluationrater bias observationhalo effect teacher evaluationleniency severity bias evaluationexact vs adjacent agreementteacher observation normingMET project observation reliabilityhow reliable is a classroom observationdouble scoring observationsevaluator consistencyteacher evaluation defensibilityobservation calibration traininghow to run a norming sessioninter-rater agreement educationteacher evaluation reliabilitycalibrating classroom observationsevaluator agreement
Anthony D. Neely, Ph.D.

Anthony D. Neely, Ph.D.

Anthony Neely is the Founder of EvalScribe, a veteran educator, an AI integration consultant for teaching & learning, researcher, & author.

Back to Blog