
How can AI help Mississippi evaluators write PGS teacher evaluations?
How can AI help Mississippi evaluators write PGS teacher evaluations?

Mississippi's Professional Growth System (PGS) — the successor to the earlier M-STAR framework — is the statewide teacher evaluation instrument used by every public school district in the state. PGS uses a Teacher Growth Rubric organized into four domains and nine standards, with a four-level rating scale (Unsatisfactory, Emerging, Professional, Advanced). AI can meaningfully help with the documentation burden of PGS evaluations: translating observation notes into rubric-aligned language, mapping evidence to the nine standards, and producing draft ratings consistent with MDE's framework. AI does not replace evaluator judgment — particularly the Professional-to-Advanced call that pivots on whether students are taking ownership of learning. Where AI lifts the most weight is the evidence-anchored write-up itself.
What Mississippi PGS is
The Mississippi Educator and Administrator Professional Growth System is the state's current teacher and administrator evaluation framework, administered by the Mississippi Department of Education's Office of Teaching and Leading, Educator Effectiveness office. PGS is the successor to M-STAR, the Mississippi Statewide Teacher Appraisal Rubric that many longer-tenured Mississippi educators still reference by name.
The framework applies to every classroom teacher in every Mississippi public school district. Administrators use a separate Administrator Growth Rubric; counselors, librarians, and specialists use their own PGS-family instruments. The Teacher Growth Rubric — the focus of this piece — organizes classroom teaching practice into four domains and nine standards, grounded in Danielson and Marzano research and customized for Mississippi's standards and student population.
Every teacher receives at least one formal observation annually (two are recommended), with a 30-minute minimum. Every formal observation requires a post-observation conference. Professional Growth Scores are due to MDE no later than June 30 each year.
The four domains and nine standards
PGS organizes teaching practice into four domains, each covering two or three standards.
Domain I: Lesson Design covers two standards. Standard 1 (Alignment & Sequence) asks whether learning outcomes and activities are aligned to the Mississippi College- and Career-Ready Standards (MCCRS) and whether the lesson is part of a coherent, connected sequence. Standard 2 (High Levels of Learning) asks whether the teacher provides scaffolding, tracks each student's progress, differentiates for varied learners, uses student-centered approaches, and connects to students' prior experiences.
Domain II: Student Understanding covers two standards. Standard 3 (Student Responsibility & Monitoring) asks whether the teacher communicates goals accessibly, uses formative assessment to monitor progress, provides self-assessment opportunities, gives specific and timely feedback, and creates opportunities for students to apply that feedback. Standard 4 (Multiple Means of Understanding) asks whether the teacher uses varied explanations, extended productive discussion, effective questioning, cross-disciplinary connections, and real-world application.
Domain III: Culture and Learning Environment covers three standards. Standard 5 (Learning-Focused Community) asks whether the teacher creates routines for safe participation, proactively monitors behavior, provides collaborative learning, and ensures active participation. Standard 6 (Space, Time & Resources) asks whether physical space is maximized, whether students always have something meaningful to do, and whether transitions are efficient. Standard 7 (Respect for All Students) asks whether the teacher communicates respectfully with all students and whether students interact respectfully with each other.
Domain IV: Professional Responsibilities covers two standards. Standard 8 (Professional Learning) asks whether the teacher engages proactively in PLCs and PD, integrates knowledge into practice, applies observer feedback, and shares learning with colleagues. Standard 9 (Family/Guardian Communication) asks whether the teacher partners with families to coordinate learning between home and school and establishes mutual expectations.
Nine standards. Two-plus-two-plus-three-plus-two. That's the surface area PGS asks Mississippi evaluators to document for every teacher, every year.
The rating scale — and the Professional-to-Advanced pivot
PGS uses four performance levels: Unsatisfactory, Emerging, Professional, Advanced.
Unsatisfactory (Level 1) indicates the teacher should receive immediate and comprehensive professional learning and support. Emerging (Level 2) indicates the teacher is making attempts but not fully demonstrating effectiveness — a level with high potential that requires clear, specific, actionable feedback. Professional (Level 3) is the expected standard: the teacher demonstrates effective instructional practices; teacher-directed success is present. Advanced (Level 4) is beyond effective: the teacher demonstrates advanced practices that foster student ownership of the learning and the environment; student-directed success.
The Professional-to-Advanced distinction is the sharpest calibration call in Mississippi PGS. In many state evaluation frameworks, the top-level rating (Distinguished, Exemplary, Highly Effective) describes teacher practice that is "exemplary" or "beyond expected" — often somewhat vaguely defined. In PGS, the top level has a much more specific structural marker: students are directing their own learning.
Look at what Level 4 requires across the nine standards:
Standard 1 (Advanced): Reflects collaboration with other school staff across disciplines to enrich learning.
Standard 2 (Advanced): Provides opportunities for students to choose challenging tasks and instructional materials.
Standard 3 (Advanced): Provides opportunities for students to demonstrate connections between learning and their personal and professional goals.
Standard 4 (Advanced): Moves all — not almost all — students to deeper understanding.
Standard 5 (Advanced): Provides opportunities for students to take on academic leadership roles that promote learning.
Standard 6 (Advanced): Students share responsibility for leading classroom routines and procedures.
Standard 7 (Advanced): Fosters a classroom culture where students give unsolicited praise or encouragement to their peers.
Standard 8 (Advanced): Serves as a critical friend for colleagues — both providing and seeking meaningful feedback on instruction.
Notice the pattern. Across seven of the nine standards, the difference between Professional and Advanced comes down to whether students are doing something specific — leading, choosing, connecting to their own goals, encouraging peers, taking academic leadership. If the teacher is running the lesson beautifully but the students aren't demonstrating that specific ownership, the correct rating is Professional, not Advanced. That's the distinction MDE built into the framework.
The Mississippi observation cycle and Professional Growth Score
PGS runs on an annual cycle. Every teacher receives at least one formal observation annually — two are recommended — with a 30-minute minimum for each formal observation. Pre-observation conferences are optional; post-observation conferences are required after every formal observation.
MDE requires that all evaluators be certified through MDE-approved PGS training before conducting observations. Districts are responsible for maintaining records of evaluator credentials, and using uncertified evaluators can jeopardize the legal defensibility of evaluation results.
Every rating on the rubric must be supported by written evidence — MDE expects evaluators to cite specific, observable teacher behaviors and student responses rather than subjective impressions. This evidence-gathering requirement is where much of the documentation time goes and where AI is most useful.
The year-end Professional Growth Score is due to MDE by June 30. This aggregates evidence collected across the year's observations, informal walkthroughs, and other sources into a summative rating for each of the nine standards.
After the observation: where the work actually piles up
The 30-minute observation is the part most Mississippi evaluators trained for. What no one warned them about was the volume of writing that follows.
A single formal PGS observation can produce three to seven pages of documentation: rubric ratings across the nine standards, evidence citations for each rating, narrative justifications, suggested next steps, and language for the post-observation conference. Multiply that across a caseload of thirty or forty teachers, with two recommended formal observations per year, and layer on the year-end Professional Growth Score aggregation, and the documentation load becomes the biggest single time cost in the evaluator's year.
MDE evaluator training specifically flags the challenge: "I know what effective teaching looks like, but translating a 45-minute observation into rubric-aligned evidence for every indicator takes longer than the observation itself." That translation is the pinch point.
Where AI helps with PGS evaluation write-ups
The strongest fit between current AI and PGS evaluation work is the translation problem. An evaluator captures observation notes in shorthand — half-sentences, abbreviations, fragments — and then has to translate those notes into the formal language of the MDE Teacher Growth Rubric with specific evidence citations for each rating.
AI handles three pieces of that translation reliably. First, it can take fragmentary observation notes and produce coherent, evidence-anchored prose in the voice of the PGS rubric. Second, it can map specific pieces of evidence to the relevant standards across all four domains — telling the evaluator that the note about the teacher's differentiated small-group work supports Standard 2 (High Levels of Learning), while the note about the student who spontaneously encouraged a peer supports Standard 7 (Respect for All Students) at Level 4. Third, it can maintain internal consistency across standards in the same evaluation, so the post-observation write-up reads as a coherent document rather than nine disconnected paragraphs.
The evidence-specificity requirement MDE flags is exactly where AI can meaningfully lift weight. Rather than "students were engaged" (which MDE explicitly names as too vague), AI can generate "students turned to their shoulder partners within seven seconds of the prompt and returned to whole-group discussion with responses that referenced specific text evidence" — the kind of specific, observable evidence PGS actually requires.
Where AI doesn't help — and what stays with the evaluator
The honest scope is narrower than the marketing on most AI tools suggests.
AI cannot make the borderline Professional-to-Advanced judgment for an evaluator. That call requires watching the lesson in real time, knowing whether students are actually taking ownership or whether the teacher is merely running an excellent teacher-directed lesson, and judging whether the specific Level 4 markers (students leading routines, students seeking peer feedback, students connecting learning to personal goals) are present or just approximated. AI can suggest a rating based on the notes provided, but the evaluator owns the final call.
AI cannot replace the relational work of the post-observation conference. The coaching conversation with a novice teacher after their first PGS observation, the harder conversations with a teacher whose ratings are trending toward Emerging — these are human work. AI can help draft talking points or summary language, but it cannot do the conversation itself.
AI cannot navigate MDE evaluator certification. Mississippi requires certified evaluators, and the professional judgment that certification confirms stays with the person who holds it.
And AI cannot do the contextual reasoning a long-tenured evaluator brings — knowing that this teacher had three new English learners arrive midyear, that this lesson followed a fire drill, that this is the third week the heat hasn't worked in this hallway. That context shapes evaluator judgment in ways no chatbot will reconstruct from observation notes.
What AI does well, it does well. What it doesn't do, it shouldn't pretend to do. Pasting observation notes into ChatGPT works fine as a glorified search engine — but for MDE-compliant documentation, the gap between "sounds polished" and "actually defensible under PGS" is wider than the polish suggests.
Observable evidence: what each PGS domain looks like in practice
Each PGS domain has its own surface features — the things an evaluator actually sees, hears, and notes during an observation.
Domain I (Lesson Design) is largely evident before the observation begins. Look for lesson plans that explicitly reference MCCRS standard codes, learning outcomes that are measurable and posted, evidence that scaffolding was planned rather than improvised, and (for Advanced) documentation of cross-disciplinary collaboration in planning. If the plan and materials arrive at the pre-observation conference already showing this, half of Domain I is documented before the lesson starts.
Domain II (Student Understanding) is where formative assessment lives. Look for the learning goal posted and referenced during the lesson, checks for understanding embedded throughout (not just at the end), specific and actionable feedback (rather than "good job" or "try again"), structured time for students to apply feedback and improve their work, and (for Advanced) students articulating how the learning connects to their own goals or life. The last one is unusual to see in a single 30-minute observation, which is why Advanced on Standard 3 is a high bar.
Domain III (Culture and Learning Environment) is most visible in the first ten minutes of the lesson. Look for how the teacher uses students' names, how transitions run, whether all or almost all students are visibly active (versus performing engagement), and whether student-to-student interactions are respectful without teacher prompting. For Advanced ratings — students leading routines, students giving unsolicited peer praise — the evidence must be truly student-initiated, not teacher-prompted.
Domain IV (Professional Responsibilities) is the domain most likely to be over-scored, because much of the evidence lives outside the observation window. Look for evidence of specific PD applied to practice, evidence of feedback loops with colleagues (Standard 8, Advanced), documentation of proactive rather than reactive family communication (Standard 9), and — critically — evidence that outreach isn't only occurring when there's a problem. "Communicates reactively" is the specific language for Standard 9, Level 2 (Emerging).
Common scoring mistakes under PGS
Even experienced evaluators slip into patterns worth naming.
The first is Advanced-inflation on Standard 7 (Respect for All Students). Because respectful communication and positive relationships are what most experienced Mississippi teachers do naturally, evaluators sometimes default to Advanced when the evidence supports Professional. The Level 4 specifier — "fosters a classroom culture where students give unsolicited praise or encouragement to their peers" — is quite specific. Warm, respectful teaching that doesn't include student-initiated peer encouragement is Professional, not Advanced.
The second is the Professional default. When evidence is thin in either direction, evaluators slot Professional because it feels safe. But PGS requires evidence of Professional practice — "teacher demonstrates effective instructional practices" — not just the absence of Emerging. A standard with insufficient evidence is a coaching opportunity, not a default Professional.
The third is Domain IV over-scoring. Because Standards 8 and 9 live partly outside the lesson observation window, evaluators tend to credit "professional-feeling" teachers with high ratings regardless of specific evidence. The fix is to require the same evidence rigor for Domain IV as for Domains I-III — name the specific PD applied, name the specific family outreach documentation, name the specific colleague feedback loop.
The fourth is misreading Standard 2's "connections to students' prior experiences" language. This isn't just prior academic learning — the rubric explicitly names "family, community, culture, language." A lesson that connects to prior lesson content but not to students' out-of-school experiences may only meet Emerging on this specific indicator, not Professional.
The fifth is vague evidence language. MDE explicitly flags "good classroom management" and "students were engaged" as too vague to support ratings. The rubric requires evidence of specific, observable teacher behaviors and student responses. This is the piece where evaluator time gets spent and where AI can meaningfully help — but the standard MDE holds evaluators to is high.
Mississippi in context — the customized-Danielson-and-Marzano model
Mississippi's approach to teacher evaluation is different from both the pure Danielson-adopting states and the district-flexibility states.
Unlike Kentucky (Kentucky Framework for Teaching, based on Danielson), Arkansas (TESS, Danielson-based), or Illinois (Danielson-based), Mississippi did not adopt Danielson's framework whole. PGS is grounded in Danielson and Marzano research but is customized — the four-domain structure, the specific nine standards, the Mississippi-specific rating labels (Unsatisfactory / Emerging / Professional / Advanced rather than Danielson's Unsatisfactory / Basic / Proficient / Distinguished), and the strong Professional-to-Advanced pivot on student ownership are all Mississippi-specific choices.
Unlike Wyoming, Indiana, Missouri, or Massachusetts, Mississippi doesn't leave the framework choice to districts. Every Mississippi district uses PGS. The consistency is real: an evaluator moving from one Mississippi district to another works with the same nine standards, the same rubric, the same rating scale, and the same MDE-required certification.
That structural choice reflects Mississippi's smaller size — roughly 1,000 public schools, compared to Texas or California — and MDE's decision that consistency across districts serves evaluator calibration better than local flexibility.
How EvalScribe handles Mississippi PGS
EvalScribe is built around the actual PGS structure — all four domains, all nine standards, the four-level rating scale with Mississippi's exact labels (Unsatisfactory, Emerging, Professional, Advanced) — rather than approximating it with generic teaching-evaluation prose. The tool handles the translation problem natively. An evaluator captures observation notes by typing, dictating, or photographing handwritten notes (Smart Scan OCR converts handwriting to text), and EvalScribe maps that evidence to the relevant PGS standards and drafts ratings and comments in the rubric's actual voice.
Beta testers report saving 30 to 60 minutes per evaluation versus writing the documentation traditionally. Across a typical Mississippi evaluator's load — say a caseload of thirty teachers with the two recommended formal observations plus informal walkthrough evidence collection for the year-end Professional Growth Score — that math adds up to dozens of hours back over an evaluation cycle.
The evaluator stays in control of every rating, every comment, and every piece of mapped evidence. EvalScribe drafts; the certified evaluator decides. The borderline Professional-to-Advanced judgment call, the harder coaching conversations, and the final Professional Growth Score decisions remain with the practitioner who knows the teacher. More detail on the PGS workflow is available at evalscribe.com/mississippi.
Frequently asked questions about PGS and AI
Where can I read the official Mississippi PGS framework and rubric? The primary sources are maintained by MDE: the Educator and Administrator PGS home page and the Teacher PGS page. Contact MDE Office of Teaching and Leading at 601-359-2330.
Does AI replace evaluator judgment in PGS? No. AI drafts evidence-mapped evaluations and suggested ratings; the certified evaluator reviews, edits, and finalizes every document. The borderline judgment calls — particularly Professional-to-Advanced, which pivots on student ownership — stay with the practitioner who observed the lesson and holds MDE certification.
Is PGS the same as M-STAR? PGS is the current evolution of what many Mississippi educators knew as M-STAR (Mississippi Statewide Teacher Appraisal Rubric). Some older materials still reference M-STAR terminology. Structurally, the current PGS Teacher Growth Rubric organizes practice into four domains, nine standards, and a four-level rating scale.
Does AI work for informal walkthroughs and formal 30-minute observations? Yes. Both flow through the same capture-and-draft pipeline, so documentation across observation types stays consistent throughout the annual evaluation cycle — which supports the evidence collection MDE expects for the year-end Professional Growth Score.
Does AI meet MDE's evidence specificity standard? MDE requires that written evidence reference specific, observable teacher behaviors and student responses rather than subjective impressions. This is exactly where AI can meaningfully help. Rather than "students were engaged" (which MDE explicitly names as too vague), AI can help translate observation notes into the specific evidence PGS requires — provided the notes captured that specificity in the first place. What the evaluator observed still matters more than what the AI writes about it.
How can I see what EvalScribe looks like for PGS specifically? The Mississippi page at evalscribe.com/mississippi walks through the nine-standard workflow, the framework handling, and a sample exported evaluation. EvalScribe is available on iOS and macOS via the App Store, with three free evaluations to start.
If you're evaluating teachers under Mississippi PGS, see how EvalScribe handles the workflow at evalscribe.com/mississippi.
References
Mississippi Department of Education, Educator and Administrator Professional Growth System (PGS)
Mississippi Department of Education, Teacher PGS
Mississippi Department of Education, Office of Teaching and Leading (601-359-2330)
Mississippi College- and Career-Ready Standards (MCCRS)
Related articles
Page maintained by Anthony D. Neely, Ph.D. — practicing K-12 educator with nearly 20 years in the classroom, 2025–2026 Walker County Distinguished Teacher of the Year, and co-founder of EvalScribe. Framework details verified against current MDE source documents.
Last reviewed: July 1, 2026.
