All case studies
Aga Khan Academy, Hyderabad
Case Study 06Aga Khan Academy, Hyderabad
Programme delivered

Training teachers who haven't taught yet

The Aga Khan Academy runs a Teacher Preparation Programme for trainees entering the profession. Two cohorts, twenty trainee teachers, measured through IMPACT — observed against a shared rubric, given assigned work between visits, and tracked from cohort one into cohort two.

20

Trainee teachers measured across the programme

2

Cohorts on one instrument, compared criterion by criterion

3

Modules used — observation, audit and assigned tasks

At a glance

ClientAga Khan Academy
ProgrammeTeacher Preparation Programme
Participants20 trainee teachers
Structure2 cohorts
Modules usedObservation · Audit · Tasks
Observation typesExternal · Internal · Self
TrackingAcross both cohorts
StagePre-service
Our roleRubric, audits, reporting, tracking
01

The challenge

Preparing a teacher is harder to measure than developing one. An experienced teacher has a track record — results, a reputation, years of observed practice. A trainee has none of that. The programme has to judge readiness on a handful of supervised lessons, and much of that judgement lives in a mentor's head rather than in a record.

That creates three problems the Aga Khan Academy wanted to solve. Feedback varied by whoever happened to observe. Progress across the programme was described rather than demonstrated. And the trainees themselves had no clear picture of where they stood — which matters more for a novice than for anyone else, because a new teacher's biggest risk is not knowing what they don't know.

02

What we put in place

One rubric across the whole programme, and every observation of every trainee scored against it.

  • Expectation-mapped criteria. Observers select the written statement describing what they actually saw, and the rating follows from the statement. A trainee reads why they were rated as they were, not just the number.
  • Evidence attached in the room. Comments, lesson plans, board photographs and samples of student work sit against the specific criterion they support, so feedback points at something rather than describing it.
  • Self-assessment before the observer's. Each trainee rated their own lesson on the same criteria before seeing anyone else's rating.
  • Tasks between observations. A criterion below expectation generated specific work with an owner and a date — and the next observation re-checked that criterion first.
  • Tracking across cohorts. Both cohorts sat on the same instrument, so cohort two could be read against cohort one rather than assessed in isolation.
03

Why self-assessment mattered most

For trainee teachers, the gap between how they rate themselves and how an observer rates them turned out to be the single most useful measurement in the programme.

A trainee who rates themselves well above their observer has a calibration problem — they cannot yet see what a strong lesson looks like, so they will not know when to change course. A trainee who rates themselves well below has a confidence problem — often teaching better than they believe, and at risk of leaving the profession early for the wrong reason.

These need opposite conversations, and neither is visible from an observer's rating alone. Running self and observer assessment on the same criteria makes the gap a number a mentor can act on.

The finding worth carrying forward

Across the programme, watching the self-versus-observer gap narrow was a better indicator of a trainee becoming ready than the observer rating rising on its own. A teacher who can accurately judge their own lesson can keep improving after the programme ends.

04

How the cycle ran

Step 01

Self first

The trainee rates their own lesson against the programme rubric before any external view is shared.

Step 02

Observed

An internal mentor or external auditor scores the same criteria, attaching evidence against each one.

Step 03

Conversation

Both ratings side by side. The gap is the agenda, and the trainee's feedback on the audit is recorded.

Step 04

Tasks, then re-check

Weak criteria become assigned work. The next observation re-checks those criteria before anything else.

05

Two cohorts, compared

Running both cohorts on one instrument turned the second into something more useful than a repeat of the first.

Cohort 01

The baseline

Established what a trainee at this academy actually looks like across each criterion at entry, mid-programme and exit.

  • Criterion-level profile per trainee
  • Programme-wide strengths and weak points
  • Task completion tracked to verified change
Cohort 02

The comparison

Measured against cohort one on identical criteria, so the programme could see whether its own changes had worked.

  • Cohort-to-cohort criterion comparison
  • Evidence for which programme changes landed
  • Earlier identification of trainees needing support
What this gives a training programme

Most teacher preparation improves by intuition — a coordinator senses what worked and adjusts. With two cohorts on one instrument, the programme can test its own changes rather than assume them, and a criterion weak across both cohorts points at the programme itself rather than at the trainees.

06

What the academy received

DeliverableLevel
Observation report per auditIndividual trainee
Running profile across the programmeIndividual trainee
Self versus observer calibration viewIndividual trainee
Task record with verified closureIndividual trainee
Cohort criterion profileCohort
Cohort-to-cohort comparisonProgramme
Programme-level strengths and gapsProgramme

An experienced teacher can be told what to fix. A trainee has to learn to see it themselves — which is why the self-rating is not a formality on this programme.

How the instrument was designed for pre-service
07

What this demonstrates

  • IMPACT works pre-service, not only in-service. The same architecture that measures a practising teacher measures a trainee, which means a school can carry one record from training into employment.
  • Cohort tracking makes a programme testable. Two intakes on one instrument turn programme design from intuition into something with evidence behind it.
  • Small cohorts still benefit. Twenty teachers is not a statistical sample, and we do not report it as one. What it produces is consistency of judgement and a complete record per trainee — both of which matter more than significance at this scale.
Stated plainly

With twenty participants across two cohorts, findings are descriptive rather than statistically tested. We report what was observed and what changed, and we do not claim significance a sample this size cannot support.

Running a teacher training programme?

One cohort is enough to see whether structured observation changes how your trainees develop.