top of page

Your Teacher Evaluation Rubric Isn't the Problem

Writer: Kelly Christopher
Kelly Christopher
5 days ago
3 min read

When teacher evaluation results seem inconsistent, the rubric is an easy place to look for the problem. Maybe the indicators need to be clearer. Maybe the performance levels need better descriptions. Maybe observers need more training. Sometimes organizations even begin wondering whether it is time for a different framework altogether.


But what if the rubric isn't really the problem?



The Rubric Is Only Part of the Equation

Well-established teacher evaluation frameworks are designed to describe the professional practices associated with effective teaching and student learning (Danielson, 2013; Klette & Blikstad-Balas, 2018).  The challenge comes when observers have to apply the language to an actual lesson.


A concept such as student engagement, for example, may seem perfectly clear on paper. In reality, it can look very different in the classroom depending on the students, the grade level, the content, the teacher, and what is happening at that particular moment in the lesson. The rubric tells us what matters, but the observer still has to determine what counts as evidence of that practice.


That's where consistency can become difficult.


Same Lesson. Same Rubric. Different Scores.

Two well-trained observers can watch the same lesson, use the same rubric, and still come away with different scores. One observer may notice how many students are participating in a discussion. Another may focus more on the quality of student responses or the level of student cognition required. A third might focus on whether students are making meaningful choices about their learning. They all understand the rubric, but observer bias may impede their ability to consider the same evidence. 


More rubric training alone doesn't necessarily solve that problem. Research on classroom observation suggests that even trained observers can still differ in how they interpret and score classroom practice, making ongoing attention to observer accuracy and reliability important (Hill et al., 2012).   Observers can know the indicators, understand the performance levels, and demonstrate strong agreement during training exercises. The difficulty occurs when they have to recognize those practices while instruction unfolds and connect what they see and hear to the language of the rubric.


Those differences matter because an observation does more than produce a score. The evidence an observer notices influences the feedback a teacher receives and the goals that may come out of a post-observation conversation. Across multiple classrooms and observations, it also becomes part of the data schools and districts use to identify trends, plan professional learning, and make continuous improvement decisions.


If observers continue to recognize the evidence behind those decisions differently, rewriting the rubric may simply carry the same problem into a new framework.


Look at the Evidence Before You Rewrite the Rubric

Before investing time and resources in changing an evaluation system, ask a more useful question: Do our observers share a clear understanding of what the practices described in our rubric actually look like in the classroom?


The challenge is connecting professional expectations to the evidence of those expectations in practice.


Evidence-First™ strengthens that connection without asking organizations to abandon the rubrics and frameworks they already use. By making observable evidence more explicit, the methodology supports greater consistency in what observers recognize, how they connect that evidence to professional practice, and ultimately how they use observation information to support teacher growth.


So, before deciding that your rubric needs an overhaul, take another look at what is happening between the rubric and the score. The real opportunity for improvement may be there.


References

Danielson, C. (2013). The framework for teaching evaluation instrument. The Danielson Group.


Hill, H. C., Charalambous, C. Y., & Kraft, M. A. (2012). When rater reliability is not enough: Teacher observation systems and a case for the generalizability study. Educational Researcher, 41(2), 56–64. https://doi.org/10.3102/0013189X12437203


Klette, K., & Blikstad-Balas, M. (2018). Observation manuals as lenses to classroom teaching: Pitfalls and possibilities. European Educational Research Journal, 17(1), 129–146. https://doi.org/10.1177/1474904117703228


 
 
 

Comments


bottom of page