
AI Tools for FEAPs Teacher Evaluations: What Florida Admins Should Know
AI Tools for FEAPs Teacher Evaluations: What Florida Admins Should Know

If you're a Florida administrator wondering whether AI can help with FEAPs-aligned evaluation work, here's the honest answer: yes, but for one specific part of the process. AI can take the observation evidence you gathered and turn it into formal, rubric-aligned documentation across all 6 Accomplished Practices and the elements underneath them. It does not replace your professional judgment as the evaluator, does not replace the broader instructional personnel evaluation system your district has adopted under State Board Rule 6A-5.065, and is not a shortcut around the FEAPs foundation that supports it. Used well, it gives you back the hours you currently spend translating what you saw in a classroom into the specific language each Accomplished Practice and performance level requires. This post walks through where AI genuinely helps with FEAPs, where general tools fall short, and what to look for if you're considering one.
A quick refresher on what FEAPs actually are
The Florida Educator Accomplished Practices serve as Florida's expectations for effective educators under State Board Rule 6A-5.065. They form the foundation for instructional personnel evaluation, professional learning systems, educator preparation programs, and certification requirements. They are not, in themselves, a single statewide rubric — they are the standards on which Florida's evaluation systems are built.
The FEAPs articulate six Accomplished Practices:
FEAP 1: Instructional Design and Lesson Planning
FEAP 2: The Learning Environment
FEAP 3: Instructional Delivery and Facilitation
FEAP 4: Assessment
FEAP 5: Continuous Professional Improvement
FEAP 6: Professional Responsibility and Ethical Conduct
Each practice contains multiple elements that describe what effective practice actually looks like in classrooms. The framework is more granular than most state evaluation systems — FEAP 2 alone contains ten elements covering everything from classroom management to cultural responsiveness to resiliency. That granularity is intentional. It reflects FEAPs' character as a comprehensive standards framework rather than a streamlined checklist.
A few things worth being honest about up front. Florida districts use a range of evaluation instruments that are all FEAPs-aligned by law: many districts use Marzano's Focused Teacher Evaluation Model, others use Danielson-aligned instruments, others use district-developed instruments built directly on FEAPs. They are all anchored to the same foundation, but the rubric an appraiser opens on a given Monday morning may look different depending on which Florida district they work in. EvalScribe supports FEAPs in its native form as well as Marzano and Danielson for districts using those instruments — more on this below.
Performance-level scoring in FEAPs-aligned evaluation work in Florida typically runs on a four-level scale: Unsatisfactory, Needs Improvement, Effective, and Highly Effective. Distinguishing between adjacent levels — particularly Effective from Highly Effective, which often turns on the degree of student ownership and self-direction in the lesson — is where the real interpretive work happens.
One more piece of context worth surfacing because Florida administrators are tracking it: the Florida Department of Education is currently in the process of revising FEAPs following 2025 legislation, with revised practices expected by mid-2026 and State Board consideration anticipated by August 2026. The revision is real and underway. EvalScribe is aware of the coming changes and will update its alignment as the revised practices are finalized and publicly available. More on this below as well.
The labor that takes Florida appraisers away from classrooms isn't the observation itself. It's translating the evidence into rubric-aligned documentation across six practices, many elements, four performance levels — for every teacher, multiple times a year.
Where AI genuinely helps in the FEAPs workflow
The honest, useful application of AI here is narrow and specific: you provide the evidence you gathered, and the tool helps produce FEAPs-aligned narrative language and performance-level-informed scoring rationale across the Accomplished Practices and their elements.
That's not hype. It addresses the actual bottleneck. The cognitive work of observing a classroom and forming a professional judgment is yours and stays yours. What AI can take off your plate is the time-consuming work of rendering that judgment into the formal, framework-aligned documentation FEAPs-based evaluation requires — distinguishing carefully between, say, Needs Improvement and Effective on a given element, or between Effective and Highly Effective, using language that maps to FEAPs' actual character.
The math is what makes it matter. FEAPs is deliberately granular. Producing aligned documentation for each element across a full faculty across multiple observation cycles per year adds up to substantial hours of labor — hours that don't actually require the appraiser's professional judgment, just the translation work that follows from it.
This is worth saying plainly, because a lot of administrators carry guilt about it: using AI to handle the documentation translation is not cutting a corner. It's the right tool for a genuinely time-consuming task, freeing you to be present for the parts of the evaluation process that need a human. The job is still hard. The framework still matters. The tool just handles the part that was never really about your expertise as an evaluator.
Where general generative AI breaks down for FEAPs
Most administrators who try this reach for a general tool — ChatGPT, Claude, Gemini — because they're free, familiar, and right there. They can produce competent-sounding evaluation language. But for FEAPs specifically, four problems show up fast.
The tool doesn't know FEAPs. A general AI platform has no built-in understanding of the 6 Accomplished Practices, the elements underneath them, the 4 performance levels, or the actual FEAPs language. To get aligned output, you have to feed all of that in — every Accomplished Practice, every element, every performance level descriptor — every single time, for every evaluation. The setup work doesn't disappear. It just moves to the front of the process.
The Effective-to-Highly-Effective distinction is specifically hard to handle. This is the failure mode most consequential for FEAPs work. Across most FEAPs elements, the substantive difference between Effective and Highly Effective turns on student ownership and self-direction — students taking responsibility for their own learning, monitoring their own progress, holding each other accountable to high expectations, leading inquiry and discussion. General AI tools default to a generic "good/great" framing that doesn't capture this distinction. The output may sound fine; it may even read as plausible. But applied across a faculty, the Highly Effective ratings will drift in ways that don't reflect what FEAPs actually intends, and the documentation won't hold up under scrutiny.
FEAPs has principles underneath it, and general tools have no awareness of them. The FEAPs framework is built on four essential principles about effective educators — high expectations for all students, deep subject knowledge, professional standards, and the principle that all persons are equal before the law. These principles shape what evidence is supposed to count and how it should be interpreted. A general tool will produce evaluation language that floats free of these principles, which means the documentation doesn't reflect FEAPs as Florida actually intends it.
The scoring isn't consistent. General AI output varies with the wording of your prompt, the day, and the model version. Two appraisers using the same tool to evaluate similar lessons can land on substantively different element-level ratings and substantively different language. In a document that informs an overall evaluation rating with contract weight, that inconsistency is the kind of problem that surfaces in grievance proceedings — particularly in Florida's labor environment, where evaluation processes are increasingly scrutinized.
There's also a quieter question worth raising: where does your teacher observation data go? When you paste classroom observations into a consumer-tier general AI tool, it's worth knowing whether that data is being retained or used to train future models. Policies differ by tool and they change over time, so the honest advice is to check the current data policy of whatever platform you're using before you put personnel-adjacent information into it.
If those limitations sound like dealbreakers, they're exactly why purpose-built tools exist. I wrote a fuller comparison of purpose-built versus general AI for teacher evaluations on the EvalScribe blog if you want to go deeper on that distinction.
What to look for in an AI tool for FEAPs
Whether you end up using EvalScribe or anything else, these are the questions worth asking before you commit to a tool for this work.
Are all 6 Accomplished Practices built into the tool — including the elements underneath each one — or do you have to provide the framework yourself?
Does the tool correctly distinguish between the 4 performance levels, particularly the Effective-to-Highly-Effective distinction that requires understanding what student-centered, student-owned practice actually looks like in the Florida sense?
Does the tool reflect FEAPs' actual character — that effective practice is defined relative to specific principles, not just technical competence — or does it default to generic teaching language?
Is the scoring consistent across appraisers and across teachers, so similar observations land in similar territory regardless of who's running it through the tool?
If your district uses Marzano, Danielson, or another FEAPs-aligned instrument, does the tool support that specific instrument as well, or only FEAPs in its native form?
Where does your observation data live? Is it stored on the vendor's servers? Is it used to train models?
Will the tool update as FEAPs is revised in 2025-2026? If your district is tracking the revision — and Florida districts should be — this matters.
How EvalScribe fits
EvalScribe was built for exactly the part of FEAPs-aligned evaluation work we've been talking about — the documentation step.
All 6 Accomplished Practices are built into the tool, with every element underneath them: 7 elements under Instructional Design and Lesson Planning, 10 elements under The Learning Environment, 10 elements under Instructional Delivery and Facilitation, 6 elements under Assessment, 6 elements under Continuous Professional Improvement, and 3 elements under Professional Responsibility and Ethical Conduct. That depth matters because FEAPs is more granular than most state frameworks — and an AI tool that only handles the practice-level structure without the underlying elements is going to produce thinner documentation than FEAPs actually calls for.
The 4-level performance rubric — Unsatisfactory, Needs Improvement, Effective, Highly Effective — is applied with rubric-aware logic that distinguishes between adjacent performance levels the way FEAPs actually intends. The Effective-to-Highly-Effective distinction is built in deliberately: across most elements, EvalScribe's Highly Effective descriptors center on student ownership and self-direction, because that's what the framework actually means by distinguished practice. That's not surface-level alignment. It's pedagogical fluency in what FEAPs is asking appraisers to assess.
For Florida districts using Marzano's Focused Teacher Evaluation Model or a Danielson-aligned instrument as the FEAPs-aligned evaluation framework, both Marzano and Danielson are also supported in EvalScribe. You can evaluate using the instrument your district has adopted while knowing the underlying FEAPs alignment is structurally sound.
On the data question from earlier: EvalScribe runs on Microsoft Azure with no data storage on our end. Your teachers' observations don't sit on our servers.
And to be straight about scope, since trust matters more than reach here: EvalScribe handles the documentation translation step. It doesn't replace your professional judgment as the evaluator, doesn't conduct pre- or post-observation conferences, and doesn't run any other component of your district's broader instructional personnel evaluation system under State Board Rule 6A-5.065. It translates the evidence you gathered into formal, FEAPs-aligned language. The evaluator remains the evaluator. The tool handles the translation work that was eating your evenings.
On the 2025-2026 FEAPs revision specifically: EvalScribe is tracking the revision process and will update its alignment as the revised Accomplished Practices are finalized and publicly available through the Department of Education and State Board. Florida appraisers using EvalScribe today won't be left behind when the revisions land — the alignment work is already part of how the product is maintained.
A brief note on the framework alignment behind all of this: EvalScribe wasn't built by a vendor looking for a market. It was built by an educator with eighteen years in the classroom and PD-development experience at the national level. The depth of the FEAPs implementation isn't accidental. It reflects an actual reading of what FEAPs requires and what each Accomplished Practice is actually asking appraisers to assess.
What that looks like in practice: some appraisers report completing an entire evaluation in under five minutes. Compared to writing one the traditional way, beta testers report saving thirty to sixty minutes per evaluation. Across a full faculty and multiple observation cycles, that adds up to your fall and spring looking very different.
Try it for yourself
I built this because evaluators who love being in classrooms shouldn't have to dread the documentation that follows — and because the teachers on the other end of these evaluations deserve thoughtful, framework-aligned feedback, not whatever fell out of a rushed Sunday-night session.
Two ways to take the next step, in increasing order of commitment:
See your numbers first. There's a time-saver calculator at evalscribe.com where you can put in your specifics — number of appraisers, number of teachers, observations per year — and see how many hours EvalScribe would give back to your campus or district. Two minutes of work. Real numbers for your actual situation. No marketing claims.
Try the tool. An individual administrator license is $100 a year. That's deliberately set below most districts' procurement thresholds, which means you don't need a committee, a purchase order, or a six-week approval cycle to try it. You can decide for yourself, this week, whether it fits how you work. Download from the App Store or sign up at evalscribe.com.
Talk to me directly. If you're thinking about a school or district license, or if you have specific questions about how EvalScribe handles your district's specific FEAPs-aligned instrument, email me at [email protected]. I read every email.
FAQ
Can I use ChatGPT to write FEAPs-aligned evaluations? You can, but the tool doesn't know FEAPs — you'll need to provide all 6 Accomplished Practices, the elements underneath them, the 4 performance levels, and the actual descriptor language every time. The level-distinctions, particularly Effective-to-Highly-Effective, will vary depending on how you prompt it. For documentation that informs an overall evaluation rating with contract weight in Florida's labor environment, that inconsistency is worth thinking carefully about.
Does AI replace the evaluator in FEAPs-based evaluation systems? No. The observation, the professional judgment, and the element-level rating decisions are yours. AI's legitimate role is helping translate your judgment into formal, framework-aligned documentation — not making the evaluation for you.
What part of FEAPs evaluation can AI actually help with? The documentation step — turning your observation evidence into formal, FEAPs-aligned write-ups across all 6 Accomplished Practices and their elements. AI tools don't replace your district's broader instructional personnel evaluation system, the observation itself, or the pre- and post-observation conferences.
Is AI-generated FEAPs documentation defensible? It depends on the tool. Consistent, framework-aligned output from a purpose-built tool — with meaningful distinctions between the 4 performance levels — holds up better than output that varies by prompt wording or flattens the level distinctions into generic language. Consistency across appraisers is what makes documentation defensible.
Does EvalScribe support FEAPs? Yes. All 6 Accomplished Practices and the elements underneath each one are built in, with the 4-level performance rubric (Unsatisfactory through Highly Effective) and rubric-aware logic that distinguishes between adjacent performance levels the way FEAPs actually requires.
Does EvalScribe also support Marzano or Danielson for Florida districts using those instruments? Yes. Both Marzano's Focused Teacher Evaluation Model and Danielson-aligned frameworks are supported in EvalScribe, in addition to FEAPs in its native form. Florida districts using either instrument as their FEAPs-aligned evaluation framework can use EvalScribe with the instrument they've actually adopted.
What about the 2025-2026 FEAPs revision? EvalScribe is tracking the FDOE revision process and will update its alignment as the revised Accomplished Practices are finalized and publicly available. Florida appraisers using EvalScribe today will not be left behind when the revisions land.
How much time does it save on FEAPs-aligned evaluations? Some appraisers report completing an entire evaluation in under five minutes. Compared to writing them traditionally, beta testers report saving thirty to sixty minutes per evaluation. You can run your own numbers using the time-saver calculator at evalscribe.com.
Is my teacher observation data safe? EvalScribe runs on Microsoft Azure with no data storage on our end — your observations don't sit on our servers. For any general AI tool, check the current data policy before entering personnel-related information.
