P3M.AI, Observation Intelligence

Watching a
lesson is easy.
Changing one is not.

IMPACT Observations turns a classroom visit into a rated, evidenced record — then turns the weak parameters into assigned tasks with owners and dates, and tracks whether anything actually moved by the next visit.

4Observation types
14Rated criteria
AIVideo audit option
AutoTask triggering
YearContinuous tracking

Most observations
end in a drawer.

Schools already observe. The visit happens, a form is filled, a conversation occurs, and then the next term begins with nothing carried forward.

01

A rating without a reason

A 3 out of 4 tells a teacher nothing they can act on. Without the expectation statement and the evidence behind it, the number is a verdict rather than a description.

02

No line to the next visit

The same weakness appears in three consecutive observations because nothing between them was assigned, owned or checked. Observation becomes an annual ritual.

03

Counting is done badly

Nobody in the room can accurately track talk share, wait time or who was actually asked a question, while also watching thirty-six children. Those are the numbers that matter most.

Two ways to observe

A person in the room.
Or a camera at the back.

Most schools use both. The observer judges what only a professional can judge. The video audit counts what no professional can count while teaching attention is elsewhere.

Mode 01 · Manual observation

The observer decides

A structured form on phone or laptop. Each criterion is rated by selecting the expectation statement that matches what was seen, with a comment and evidence attached. Works offline in a classroom with weak signal and syncs later.

  • 14 criteria across 6 parameters
  • Expectation-mapped ratings, not bare numbers
  • Comment required on any rating below expectation
  • Photo, document and audio evidence attached inline
  • Closing questions the rubric cannot capture
Mode 02 · AI video audit

The recording measures

A single fixed camera at the back of the room. The analysis returns a minute-by-minute timeline, talk share, wait time per question, seat-level question distribution and aggregate on-task rate — with a confidence score on every finding.

  • Lesson timeline classified minute by minute
  • Flagged moments with timecodes and confidence
  • Seat positions only — no facial identification
  • Machine ratings stay provisional until an observer confirms
  • Divergence from the observer is logged, not hidden
Where AI wins

Counting. Talk ratio, wait time, question spread, transition length, on-task trend — measured exactly, every lesson, without observer fatigue.

Where people win

Judgement. Whether an explanation was well pitched, whether a question was worth asking, whether a child was quietly lost. A camera cannot see understanding.

Our rule

A machine rating never enters a teacher's record unconfirmed. Where the two disagree, the observer's rating stands and the gap is logged to improve the model.

Four observation types.
One shared rubric.

The same 14 criteria are used by every type, so a self-rating, a peer note and an external audit can sit on one chart and be compared honestly. The differences are who observes and what the result is used for.

EXT
Observer from outside

External

Conducted by a P3M.AI auditor or an empanelled external expert. Used for benchmarking, accreditation evidence and an unfamiliar pair of eyes on established habits.

Typically once or twice a year
INT
Coordinator or head

Internal

Run by the academic coordinator, head of department or principal. The backbone of the cycle — this is where support decisions and development plans are grounded.

Two to three per teacher per year
PEER
Another teacher

Peer

One teacher observes another, both directions, with cover provided. The cheapest intervention a school has, and the one teachers request most often in feedback forms.

Once a term, paired by need
SELF
The teacher

Self

The teacher rates their own lesson on the same rubric, before seeing anyone else's rating. The gap between self and observer is often the most useful thing on the profile.

Before every observed lesson

Why the self-rating matters. A teacher who rates themselves 3.6 where the observer sees 2.4 needs a different conversation from one who rates themselves 2.2 where the observer sees 3.0. The first is a calibration problem; the second is a confidence problem. Both are invisible without the self view.

Inside the form

Context in.
Better output
out.

A rating on its own is thin data. What makes the generated report specific is everything attached to it — the comment explaining the rating, the photograph of the board, the lesson plan, the sample of student work. The engine writes from the evidence, so the more context an observer captures in the room, the less generic the output.

Comments carry the reasoning

Required on any rating below expectation. This is what a teacher reads first and what protects the rating if it is later contested.

Evidence carries the proof

Board photographs, lesson plans, student work samples, short audio clips, worksheets. Attached to the specific criterion, not dumped at the end.

Closing questions carry the rest

Was the objective met? Did anything happen the rubric cannot capture? Is immediate support needed? Free text, and often the most valuable field in the form.

impact.p3m.ai / observation / new
Parameter 04 · Student engagement
Questioning reaches the whole class
EmergingDevelopingProficientExemplary
Expectation selectedQuestioning was directed at a small group of students rather than distributed across the class.
Observer comment · required18 of 23 questions went to four students in the front two rows. Eleven of 36 students spoke at all during the period.
IMG board_1120.jpg
PDF lesson_plan_wk4.pdf
IMG student_work_03.jpg
M4A clip_2140.m4a
Students are given thinking timeProficient · 3
Wait time was given for most questions, though not consistently.
Tasks are adjusted for different levelsDeveloping · 2
No adjustment for different levels was observed during the lesson.
Observation form · criterion rating with expectation, comment and attached evidence

One audit,
or a whole year
of them.

Some schools want a single external read before an accreditation visit. Others want a running record for every teacher. Both run on the same instrument, so a school can start with the first and move to the second without redoing anything.

Engagement 01

One-time audit

A defined window — a week, a department, a grade band. External observers, a full set of individual reports, and one consolidated findings document for the leadership team.

  • Scoped and priced per engagement
  • Individual report per teacher observed
  • Consolidated school findings and priorities
  • No platform commitment required
Engagement 02

Active tracking

A live dashboard across the year. Every observation, of every type, lands on the same teacher profile — so the third visit can open with what the second one asked for.

  • Live dashboard for coordinators and heads
  • Individual profile per teacher, across years
  • Self, peer, internal and external on one chart
  • Tasks, due dates and completion tracked throughout
impact.p3m.ai / observations / dashboard
School dashboard · Term 1, 2026
Observations done96
Feedback returned89%
Tasks overdue7
Mean rating3.0
School mean by parameter
Prof. conduct
3.3
Classroom climate
3.2
Planning
3.1
Delivery
3.0
Engagement
2.9
Assessment
2.8
Needs attention
PHMr. P. Halder2.43 overdue
SNMr. S. Nandi2.91 overdue
RBMs. R. Banerjee3.2On track
Active tracking dashboard · school-wide view for coordinators and heads
impact.p3m.ai / teacher / r-banerjee
Teacher profile · English · 2025–26
Ms. R. Banerjee · 4 observations this year
Current3.2
Since Mar+0.3
Tasks open2
Observation history
SELF12 Mar · self-rating3.6Closed
INT12 Mar · internal2.9Closed
PEER18 Jun · peer3.1Closed
INT29 Jul · internal3.2Signed
VID29 Jul · video audit3.1Provisional
Calibration gapSelf-rating exceeded observer rating by 0.7 in March. Gap narrowed to 0.4 by July.
Individual teacher profile · every observation type on one record

From finding to follow-through

A low score
should create
a task.

This is the part most observation systems are missing. When a parameter falls below its threshold, IMPACT proposes a task tied to that specific criterion — with an owner, a cadence and a due date — and carries it into the next observation as something to re-check.

Step 01

Threshold breached

A criterion is rated below expectation, or a parameter mean falls under the school's set threshold. The trigger is a rule the school configures, not a black box.

Step 02

Task recommended

The engine proposes actions drawn from that criterion's playbook, shaped by the observer's comment and evidence. The coordinator accepts, edits or dismisses each one.

Step 03

Owned and scheduled

Every accepted task gets an owner, a cadence and a due date. Some sit with the teacher, some with the coordinator, some with the department or the school.

Step 04

Re-checked, not repeated

The next observation opens with the open tasks and re-checks only the criteria they address. Progress is measured against the specific thing that was asked for.

impact.p3m.ai / tasks / r-banerjee
Task manager · continuous improvement
Recommended
Question from a seating list

Print the class list, tick names as asked, review weekly.

Trigger · P4 ≤ 2.0
Peer visit · 7C poetry

Observe Ms. Iyer, cover arranged by coordinator.

Requested in feedback
In progress
Writing in the middle of the lessonWk 3/6

Six minutes of written response before discussion, every lesson.

Trigger · P5 ≤ 2.5
Done
Share objective on the boardVerified

Confirmed at the 29 Jul observation. Rated Exemplary.

Closed · from Mar audit
Task manager · recommended, in progress and verified-closed

Rules the school sets

Thresholds, who owns which task type, and how long a task may stay open before it escalates. A school that observes lightly and one that observes intensively need different rules, so they are configuration rather than product.

Recommendations, not instructions

Every proposed task is reviewed by a human before it is assigned. The engine is good at noticing that a criterion has fallen three times; it is not the right thing to decide what a specific teacher needs.

Closed means verified

A task closes when the next observation confirms the change, not when someone ticks a box. That single rule is the difference between a development record and a compliance log.

What comes
back out.

Four report types, all built on the same rubric, all written in plain language for the person receiving them.

impact.p3m.ai / report / IMP-OBS-7B-118
Teacher observation report
3.2 / 4 · Proficient
Planning
3.5
Climate
3.7
Delivery
3.3
Engagement
2.7
Assessment
2.5
SummaryA well-planned lesson that a few students carried. Subject knowledge and climate are strong; the lesson lost ground on distribution and on written work reaching the page.
Observation report · ratings, evidence and recommendations
impact.p3m.ai / report / IMP-VID-7B-118
Video audit · AI analysis
Teacher talk61%
Student talk17%
Wait time2.8s
On-task84%
19:10 · 7m 10s unbroken teacher talk0.96
On-task rate falls from 87% to 61% across the same window.
31:20 · possible check for understanding0.58
Low confidence · observer review required before use.
Video audit report · timeline, metrics and flagged moments

Teacher report

Per observation. Ratings with evidence, strengths, development areas, tasks and sign-off.

Video audit report

Timeline, talk metrics, question distribution, flagged timecodes and stated limitations.

Department rollup

Parameter means by department, weakest criterion, and where practice varies within a team.

School annual

Year-wide teaching profile alongside student outcomes, with priorities by owner.

Trust is the product

An observation system
teachers don't trust
produces compliance.

Every audited teacher receives a feedback form on the process itself. If a rating is contested with evidence, it can be revised — and the revision is recorded, not quietly overwritten.

Feedback

Auditee rates fairness, observer conduct and whether feedback was usable.

Contest

Any rating can be challenged with evidence. Revisions are logged with reasons.

48 hrs

Post-observation conversation target, tracked as a process measure.

Consent

Video audits run with written consent, seat-level anonymity and scheduled deletion.

What schools
ask first.

Is this going to be used to appraise or rank our teachers?

Not by design. Observations are written as development records — the teacher's own feedback is attached to the report, contested ratings are visible, and tasks close on verified change rather than on a manager's sign-off. If a school chooses to use the data in appraisal, that is a school policy decision and we would ask you to state it to your staff before the first cycle rather than after.

Do we need cameras in every classroom?

No. Most schools run manual observations only, and add video audit for a small pilot group or for specific development cases. A single portable camera moved between rooms covers a full department. Video audit requires written consent and carries a deletion schedule.

Who can see a teacher's observation history?

The teacher, their observer, and named leadership roles. Peers see only the observation they conducted. There is no school-wide leaderboard, and department rollups report parameter means rather than named individuals.

What if a teacher disagrees with a rating?

They contest it through the feedback form with their evidence. The observer either revises the rating or explains in writing why it stands, and both appear in the final report. In our reference school, 11 of 96 ratings were contested last year and four were revised — which is the process working, not failing.

How long does one observation take?

The lesson itself plus roughly ten minutes on the form, most of which is the observer's comments. The report is generated within a day. The post-observation conversation is the part that matters and should happen within 48 hours, which the platform tracks.

Can we run a one-time audit before committing?

Yes, and it is the usual starting point. A single window — one department or one grade band — produces individual reports and a consolidated findings document with no platform commitment. If you move to active tracking afterwards, that first audit becomes the baseline rather than being discarded.

Start with
one department.

One window, real reports, your own rubric. Then decide whether it deserves a year.

A 30-minute call with your academic head. We bring a sample observation report built on your rubric.