
How can AI help Missouri evaluators write MEES teacher evaluations?
How can AI help Missouri evaluators write MEES teacher evaluations?

Missouri's MEES — the Missouri Educator Evaluation System — is built on the nine Missouri Teacher Standards with thirty-six quality indicators and a four-level career continuum (Emerging, Developing, Proficient, Distinguished). Missouri districts can adopt MEES, use it as a template for a locally-developed plan, or adopt an alternative like the Network for Educator Effectiveness (NEE) — all of which must align to DESE's Essential Principles of Effective Evaluation. AI can meaningfully help with the documentation burden of MEES evaluations: translating walkthrough notes into rubric-aligned language, mapping evidence to the right quality indicators, and producing draft ratings consistent with the career continuum's developmental framing. AI does not replace evaluator judgment — particularly the Proficient-to-Distinguished call, which marks the transition from career professional to educational leader. Where AI lifts the most weight is the write-up itself.
What Missouri MEES is
The Missouri Educator Evaluation System is DESE's model evaluation system for K-12 teacher evaluation, built on the nine Missouri Teacher Standards. It was developed by Missouri's Department of Elementary and Secondary Education in coordination with educator preparation programs across the state, and is most recently revised in 2023.
Unlike states with a single mandated evaluation framework, Missouri law (Section 168.128 RSMo) requires every district to evaluate certified staff but does not mandate any specific model. School districts can adopt MEES entirely, use it as a template for a locally-developed plan, or adopt an alternative model — the Network for Educator Effectiveness (NEE) from the University of Missouri is widely used; Marzano and Stronge are also common; some districts develop their own plans tailored to local context.
What is required of every Missouri evaluation plan: alignment to DESE's Essential Principles of Effective Evaluation, which include research-based standards as the foundation, observations and ongoing feedback, the use of student growth data, surveys, professional artifacts, and connection to professional learning. Most Missouri evaluation plans — including most district-developed and alternative models — are grounded in the Missouri Teacher Standards, which provide the common professional language for what effective teaching looks like in Missouri.
The nine Missouri Teacher Standards and thirty-six quality indicators
The Missouri Teacher Standards organize teaching practice into nine standards, each with multiple quality indicators (QIs).
Standard 1 (Content Knowledge & Instruction) covers five quality indicators: content knowledge and academic language, student engagement in subject matter, disciplinary research and inquiry methodologies, interdisciplinary instruction, and diverse social and cultural perspectives.
Standard 2 (Student Learning, Growth & Development) covers six QIs — the largest standard — including cognitive, social, emotional and physical development; student goals; theory of learning; differentiated lesson design; prior experiences, multiple intelligences, strengths and needs; and language, culture, family and knowledge of community values.
Standard 3 (Curriculum Implementation) covers three QIs: implementation of curriculum standards, lessons for diverse learners, and instructional goals and differentiated instructional strategies.
Standard 4 (Critical Thinking) covers three QIs: instructional strategies leading to student engagement in problem-solving and critical thinking, appropriate use of instructional resources, and cooperative, small group and independent learning.
Standard 5 (Positive Classroom Environment) covers three QIs: classroom management techniques; management of time, space, transitions, and activities; and classroom, school and community culture.
Standard 6 (Effective Communication) covers four QIs: verbal and nonverbal communication; sensitivity to culture, gender, intellectual and physical differences; learner expression in speaking, writing and other media; and technology and media communication tools.
Standard 7 (Student Assessment & Data Analysis) covers six QIs: effective use of assessments, assessment data to improve learning, student-led assessment strategies, effect of instruction on individual/class learning, communication of student progress and maintaining records, and collaborative data analysis.
Standard 8 (Professionalism) covers three QIs: self-assessment and improvement, professional learning, and professional rights, responsibilities and ethical practices.
Standard 9 (Professional Collaboration) covers three QIs: induction and collegial activities, collaborating to meet student needs, and cooperative partnerships in support of student learning.
That's thirty-six quality indicators total, each carrying its own rubric language and its own evidence requirements.
The four-level career continuum — and the Proficient-to-Distinguished judgment
MEES uses a four-level career continuum: Emerging, Developing, Proficient, Distinguished. Emerging recognizes a teacher entering the profession applying base knowledge and skills — foundational practices present but inconsistent. Developing reflects knowledge and skills continually developing through new classroom, school, and community experiences. Proficient is the career professional standard — consistently advances student growth through effective, well-established practice. Distinguished exceeds proficiency — the teacher serves as an educational leader in the school, district, and profession while consistently advancing student growth.
The framing matters. Missouri's continuum is explicitly developmental rather than deficit-based. Emerging isn't "bad" — it's where the profession starts. Distinguished isn't just "great teaching" — it's the move from individual classroom excellence to school and profession-wide leadership.
The hardest call in MEES — and the one that most divides evaluators in calibration sessions — is Proficient versus Distinguished. Both reflect strong teaching: organized lessons, engaged students, evidence-based practice, productive classroom culture. What separates them is the move from career professional to educational leader. At Proficient, the teacher consistently delivers effective instruction for her students. At Distinguished, that teacher is also mentoring colleagues, leading collaborative curriculum work, championing professional ethics, and building sustained community partnerships. The teacher set the conditions in her own classroom; she's now setting them across the school and district. That distinction is real, and it's the one most likely to get blurred when an evaluator is writing a summative quickly between observations.
After the walkthrough: where the work actually piles up
The observation itself is the part most evaluators trained for. What no one warned them about was the volume of writing that follows.
A single formal observation in a MEES-aligned plan can produce three to seven pages of evaluation documentation: rubric ratings across quality indicators, narrative justifications for each rating, evidence quotations from observation notes, suggested next steps, and language for the post-observation conference. With thirty-six quality indicators distributed across nine standards, MEES has more discrete rating decisions than most state frameworks — and the documentation accumulates fast.
This is where the day actually goes. The walkthrough takes fifteen minutes. The write-up takes forty-five.
Where AI helps with MEES evaluation write-ups
The strongest fit between current AI and Missouri MEES evaluation work is the translation problem. An evaluator captures observation notes in shorthand — half-sentences, abbreviations, fragments — and then has to translate those notes into the formal language of the Missouri Teacher Standards rubric. That translation is mechanical work. It's not where evaluator judgment lives; it's where evaluator time gets spent.
AI handles three pieces of that translation reliably. First, it can take fragmentary observation notes and produce coherent, evidence-grounded prose in the voice of the MEES rubric. Second, it can map specific pieces of evidence to the relevant quality indicators across all nine standards — telling the evaluator that the note about Bloom's higher-order tasks supports Standard 4 QI 1 (Instructional strategies leading to student engagement in problem-solving and critical thinking), while the note about cooperative groups with clear roles supports Standard 4 QI 3 (Cooperative, small group and independent learning). Third, it can produce internally consistent draft language across quality indicators in the same evaluation, so the summative reads as a coherent document rather than thirty-six disconnected paragraphs.
Where AI saves the most time is the third piece. Most evaluators don't struggle to write any one rubric justification. They struggle to write thirty-six of them in a single document while keeping voice and evidence-density consistent — particularly across Standards 1, 2, and 7, where five or six quality indicators each need distinct, evidence-anchored treatment.
Where AI doesn't help — and what stays with the evaluator
The honest scope is narrower than the marketing on most AI tools suggests.
AI cannot make the borderline Proficient-to-Distinguished judgment for an evaluator. That call requires watching the lesson in real time, knowing the teacher's history and context, and judging whether the evidence rises to the standard of educational leadership beyond the classroom. AI can suggest a rating based on the notes provided, but the evaluator owns the final call — particularly because Distinguished evidence often lives outside the observation window (mentoring colleagues, leading PLCs, building community partnerships) and requires institutional context AI doesn't have.
AI cannot navigate the relational work of the post-observation conference. The coaching conversation, the goal-setting around professional growth, the harder conversations when performance is moving toward Developing on the continuum — these are human work. AI can help draft talking points or summary language, but it cannot do the conversation itself.
AI cannot do the contextual reasoning a long-tenured evaluator brings — knowing that this teacher had three new English-learners arrive midyear, that this lesson followed a fire drill, that this is the third week the heat hasn't worked in this hallway. That context shapes evaluator judgment in ways no chatbot will reconstruct from observation notes.
And AI cannot integrate the broader Essential Principles of Effective Evaluation that DESE requires. The evaluator brings the student growth data, the survey feedback, the professional artifacts, and the professional learning history into the summative judgment. AI can draft the rubric portion of the evaluation based on observation evidence; the integration of those other data sources is evaluator work.
What AI does well, it does well. What it doesn't do, it shouldn't pretend to do. Pasting walkthrough notes into ChatGPT works fine as a glorified search engine — but for rubric-aligned MEES documentation, the gap between "sounds polished" and "actually defensible under your district's evaluation plan" is wider than the polish suggests.
Observable evidence: what each Missouri Teacher Standard looks like in practice
Each standard has its own surface features — the things an evaluator actually sees, hears, and notes during an observation.
Standards 1 and 4 (Content Knowledge & Critical Thinking) are most visible in the cognitive demand of the lesson. Look for precise academic vocabulary used and explicitly taught, students positioned as disciplinary thinkers (historians, scientists, mathematicians), tasks that require students to analyze, evaluate, or create — not just recall — and questions scaffolded through Bloom's levels to higher-order thinking.
Standard 2 (Student Learning, Growth & Development) is most visible in differentiation. Look for tiered tasks, flexible grouping, varied entry points, students referencing their own prior knowledge, student goals visible (articulated, posted, or tracked), and cultural, linguistic, and community assets treated as instructional resources rather than barriers.
Standard 3 (Curriculum Implementation) is largely evident before and after the lesson. Look for lesson objectives explicitly tied to Missouri Learning Standards, evidence of long-range or unit planning behind the daily lesson, instructional goals that are stated, measurable, and communicated to students, and differentiated strategies built into the lesson rather than added on.
Standard 5 (Positive Classroom Environment) is most visible in the first five minutes. Look for behavioral expectations that are visible and consistently upheld, transitions that are planned and brisk, physical space organized to support the learning activities in use, students showing self-motivation, and classroom culture reflecting awareness of students' community identities.
Standard 6 (Effective Communication) lives in how the teacher talks and how students get to express their learning. Look for precise, audible language; intentional nonverbal communication (proximity, eye contact); sensitivity to cultural, gender, and ability differences in language and materials; structured opportunities for students to express learning across speaking, writing, and multimedia; and purposeful (not decorative) use of technology and media tools.
Standard 7 (Student Assessment & Data Analysis) is visible in the formative-summative balance and in student-led assessment artifacts. Look for both formative and summative assessments visible in the lesson cycle, observable evidence that data shaped today's instruction, students articulating their learning goals and progress, student-led assessment artifacts (trackers, rubrics, self-reflections), and organized records of student progress.
Standards 8 and 9 (Professionalism & Professional Collaboration) are the standards most likely to be overrated, because most of the evidence lives outside the observation window. Look for documented contributions to PLCs and collaborative planning, professional development visibly applied to practice (not just attended), proactive partnership-oriented family communication, and evidence of collaboration with specialists or support staff around specific student needs.
Common scoring mistakes under MEES
Even experienced evaluators slip into patterns worth naming.
The first is over-Distinguishing Standards 8 and 9. Because these standards live partly outside the lesson, evaluators tend to credit "professional-feeling" teachers with high ratings regardless of evidence. The fix is to require the same evidence rigor for professionalism and collaboration as for instruction — name the PLC contribution, name the specific PD applied, name the specific community partnership.
The second is the Proficient default. When evidence is thin in either direction, evaluators slot Proficient because it feels safe. But MEES requires evidence of Proficient practice — consistently advancing student growth through effective, well-established practice — not just absence of Developing. A quality indicator with insufficient evidence is a coaching opportunity, not a default Proficient.
The third is quality indicator conflation within Standards 1, 2, and 7. Adjacent indicators feel related, and evaluators often justify both with the same evidence. They're distinct claims. In Standard 1, content knowledge (QI 1) and disciplinary inquiry methodologies (QI 3) are different — knowing the content versus engaging students as practitioners of the discipline. In Standard 7, effective use of assessments (QI 1) and student-led assessment strategies (QI 3) are different — the teacher using assessment versus students leading their own assessment.
The fourth is evidence-thin Distinguished. Distinguished in MEES requires evidence of educational leadership beyond the classroom — mentoring colleagues, leading collaborative curriculum work, championing professional ethics, building sustained community partnerships. "The teacher did it well in her classroom" is not Distinguished — it's Proficient. Distinguished requires the move from individual classroom excellence to school and profession-wide leadership.
The fifth is undervaluing Emerging as a developmental label. Some evaluators treat Emerging as a punitive rating equivalent to "Unsatisfactory" in other state frameworks. Missouri's continuum framing is intentionally different: Emerging recognizes the teacher entering the profession, applying base knowledge and skills. A first-year teacher rated Emerging on most quality indicators isn't being marked down — they're being placed accurately on the continuum, with growth toward Developing and Proficient as the explicit trajectory.
The sixth is uniquely Missouri: failing to ground the evaluation in the Essential Principles framework. MEES is one piece of a broader evaluation system that includes student growth data, surveys, professional artifacts, and professional learning. The classroom observation rubric is the most visible piece, but a defensible MEES evaluation integrates all the Essential Principles — and an evaluator who treats the rubric in isolation misses the multi-source design DESE built.
Missouri in context — the district-flexibility model
Missouri's approach to teacher evaluation is genuinely different from many other states. States like Texas (T-TESS), Florida (FEAPs), Georgia (TKES), Arkansas (TESS), and Tennessee (TEAM) operate single statewide rubrics that every district uses. Missouri — like Indiana and Wyoming — provides a state model (MEES) and allows districts to adopt it, modify it, or develop their own evaluation plan as long as it aligns to the Essential Principles of Effective Evaluation.
The Missouri version of this district-flexibility model has a strong alternative ecosystem. NEE (Network for Educator Effectiveness) from the University of Missouri's College of Education is widely adopted across Missouri districts and has its own observation protocol, online platform, and rubric grounded in the same Missouri Teacher Standards. Other districts use Marzano, Stronge, or locally-developed plans. The common thread is the Missouri Teacher Standards as the foundational expectation for what effective teaching looks like.
For Missouri evaluators, this means the specific rubric language and structure of your district's plan is the operational reality you work with. Any tool that supports Missouri evaluation has to be flexible enough to accommodate district variations — MEES adopters, NEE adopters, and locally-developed plan adopters — without losing the standards-aligned logic that makes Missouri evaluation defensible.
How EvalScribe handles Missouri MEES
EvalScribe is built around the actual Missouri Teacher Standards structure — all nine standards, all thirty-six quality indicators, the four-level career continuum — rather than approximating it with generic teaching-evaluation prose. The tool handles the translation problem natively: an evaluator captures observation notes by typing, dictating, or photographing handwritten notes (Smart Scan OCR converts handwriting to text), and EvalScribe maps that evidence to the relevant quality indicators and drafts ratings and comments in the rubric's actual voice.
For districts using a locally-developed plan grounded in the Missouri Teacher Standards, the underlying rubric-aware logic still applies. EvalScribe's MEES template is a starting point; the engine accommodates district-specific variation in language and structure while honoring the standards-aligned foundation Missouri DESE requires. Dedicated NEE support is on our roadmap for a future update — NEE-using districts can reach out to [email protected] to be notified when it ships.
Beta testers report saving 30 to 60 minutes per evaluation versus writing the documentation traditionally. Across a typical evaluator load — say thirty teachers and three to four observations per year — that math adds up to dozens of hours back over an evaluation cycle.
The evaluator stays in control of every rating, every comment, and every piece of mapped evidence. EvalScribe drafts; the evaluator decides. The Distinguished rating call that requires institutional context, the integration with student growth data and other Essential Principles components, and the harder coaching conversations remain with the practitioner who knows the teacher. More detail on the Missouri MEES workflow is available at evalscribe.com/missouri.
Frequently asked questions about Missouri MEES and AI
Where can I read the official Missouri Teacher Standards and MEES framework? The primary sources are maintained by the Missouri Department of Elementary and Secondary Education: the DESE Teacher Evaluation home page, the Missouri Teacher Standards page, and the DESE Model Evaluation System. Statutory basis is Section 168.128 RSMo.
Does AI replace evaluator judgment under MEES? No. AI drafts evidence-mapped evaluations and suggested career-continuum ratings; the evaluator reviews, edits, and finalizes every document. The borderline judgment calls — particularly Proficient-to-Distinguished — stay with the practitioner who observed the lesson and knows the teacher.
My district uses NEE instead of MEES. Will AI tools work for us? Generic AI tools won't, because they have no awareness of the Missouri Teacher Standards or how districts adapt evaluation. EvalScribe currently ships native support for MEES and for locally-developed plans grounded in the Missouri Teacher Standards. Dedicated NEE support is on our roadmap for a future update — if your district uses NEE, reach out to [email protected] and we'll let you know when it ships.
What's the difference between the Missouri Teacher Standards and MEES? The Missouri Teacher Standards are the nine foundational standards that define expectations of Missouri teachers — the content of what's being evaluated. MEES (Missouri Educator Evaluation System) is the evaluation system built around those standards — the process, scoring scale, and protocols. Most Missouri district evaluation plans are grounded in the Missouri Teacher Standards regardless of which evaluation system they use.
What is the four-level career continuum and why does Missouri use it? Missouri's continuum — Emerging, Developing, Proficient, Distinguished — frames teaching as a developmental career trajectory rather than a pass/fail judgment. The framing is intentional. Emerging is where teachers enter the profession; Distinguished represents educational leadership beyond the classroom. This is different from deficit-based scales that treat the lowest rating as "unacceptable."
Does AI work for short observations as well as formal observations? Yes. Walkthroughs, formative observations, and summative reviews all flow through the same capture-and-draft pipeline, so the documentation across observation types stays consistent within your district's evaluation cycle.
How can I see what EvalScribe looks like for Missouri MEES specifically? The Missouri page at evalscribe.com/missouri walks through the nine-standard workflow, the MEES rubric handling, and a sample exported evaluation. EvalScribe is available on iOS and macOS via the App Store, with three free evaluations to start.
If you're evaluating teachers under Missouri MEES or a locally-developed plan grounded in the Missouri Teacher Standards, see how EvalScribe handles the workflow at evalscribe.com/missouri.
References
Missouri Department of Elementary and Secondary Education, Teacher Evaluation
Missouri Department of Elementary and Secondary Education, Missouri Teacher Standards
Missouri Department of Elementary and Secondary Education, Model Evaluation System
Missouri Department of Elementary and Secondary Education, Essential Principles of Effective Evaluation
Section 168.128 RSMo (Missouri Revised Statutes)
Related articles
Page maintained by Anthony D. Neely, Ph.D. — practicing K-12 educator with nearly 20 years in the classroom, 2025–2026 Walker County Distinguished Teacher of the Year, and co-founder of EvalScribe. Framework details verified against current Missouri DESE source documents.
Last reviewed: June 30, 2026.
