The Tutor Handbook · Volume 0107 · Series ID THB-0107
The Tutor Handbook: Complete series index.
Three learners sit at the same table.
The tutor gives them the same unfamiliar mathematics problem. One learner sees the structure quickly and starts within fifteen seconds. The second reads twice, draws a diagram and begins after a minute. The third is still deciding what the question is asking.
Nothing has gone wrong yet.
Then the room starts comparing.
“She is already on part (b).”
“Why are you so slow today?”
“Look at how she did it.”
If the tutor is not careful, the fastest learner quietly becomes the standard against which the other two are interpreted.
The Peer Baseline Trap occurs when a tutor mistakes another learner’s speed, confidence, method, output or current achievement for the criterion that should determine what a particular learner knows, needs or should do next.
Peer information can be valuable. Students can explain, challenge, compare methods, notice alternatives and learn socially. Small-group tuition can create interaction that one-to-one tutoring does not. The problem is not comparison itself. The problem is comparison without a clear educational job.
Quick Answer
In a three-student tutorial, use peers as additional examples, dialogue partners and sources of alternative reasoning. Do not use the fastest, most confident or highest-scoring learner as the hidden benchmark for everyone else. Judge performance against the learning goal, success criteria, task conditions, prior evidence and the learner’s current route.
Ask three separate questions:
- What does the task require?
- What does this learner’s response show?
- What, if anything, does the peer comparison add?
If the third question starts replacing the first two, the group has entered the peer baseline trap.
1. A Peer Is Evidence About the Group, Not Automatically Evidence About the Learner
Suppose Alicia solves a question in two minutes and Beatrice takes five. The time difference is real. It may be educationally useful. It does not, by itself, prove that Beatrice lacks understanding.
Perhaps Alicia has seen the problem structure before. Perhaps Beatrice uses a slower but robust method. Perhaps Alicia is fluent and Beatrice is still consolidating. Perhaps Beatrice misread one condition. Perhaps Alicia has made a hidden mistake. Perhaps the task is testing accuracy, not speed.
The peer difference opens questions. It does not answer them.
This distinction matters because three-student tuition produces a lot of comparative information naturally. Learners hear each other’s questions, see different completion times and watch different methods emerge. A tutor can use that richness without turning every difference into a rank order.
2. Success Criteria Should Not Move Because Someone Else Is Stronger
A learner’s task should have a success condition that exists before another learner performs it.
If the aim is to solve a linear equation accurately and explain the balance operation, success should not suddenly require matching the fastest peer’s speed. If the aim is to identify evidence for an inference, success should not become “write an answer as sophisticated as Ciara’s”. If the aim is to produce one coherent paragraph, a peer writing two pages does not change the target.
Current assessment guidance from the NSW Department of Education describes success criteria as closely linked to learning intentions and as making clear what students are being judged on. The high-impact formative assessment guidance also frames questioning and success criteria around evidence of learning rather than social rank.
That principle transfers cleanly to tutoring: define the target before looking sideways.
3. The Fastest Learner Can Distort the Tutor’s Sense of Pace
Group pace often drifts toward the learner who finishes first because completion is visible.
The tutor sees one student waiting and feels pressure to move on. The slower learners then experience a compressed finish: hints arrive earlier, explanations shorten, checking disappears, or the next task begins before the current reasoning has stabilised.
This can create a misleading cycle. The fastest learner receives independent practice. The others receive increasing tutor support to keep the group synchronised. Their work then looks more dependent, which seems to confirm that they are “weaker”.
The tutor needs a different definition of group pace: the lesson can contain different local clocks while preserving a shared broad route.
One learner can move to a transfer question while another completes an independent check. The room does not need everyone on the same line at the same second.
4. Speed Is Only a Criterion When Speed Is Part of the Job
Speed can matter. Fluency matters. Examination conditions matter. Timed execution can be a legitimate performance target.
But speed should be introduced deliberately rather than smuggled in through peer comparison.
If the tutor is currently diagnosing conceptual understanding, a slower correct explanation may be better evidence than a rapid answer. If the learner already understands the method and the active job is fluency, response time becomes more relevant. If the active job is performance under examination conditions, then speed must be assessed against the actual time demand, not against whichever peer happens to be quickest.
The Timed Set already owns the deeper distinction between adding a clock and proving learning. The peer baseline trap occurs when speed enters the judgement without that explicit decision.
5. Confidence Is Not a Ranking Instrument
Some learners answer quickly and publicly. Others think quietly. A confident learner can make a group feel as though one method is settled before the tutor has seen independent responses from everyone.
This creates two risks.
First, the tutor may mistake confidence for correctness. Second, quieter learners may revise their thinking before the tutor sees what they originally believed. The room loses diagnostic evidence.
The solution is not to silence confident students. Use response routines that preserve first attempts: short written starts, individual think time, mini-whiteboard responses, or deliberate sequencing in which everyone commits before discussion.
Then confidence becomes one communication characteristic rather than the standard for competence.
6. Peer Explanation Helps When It Does Not Replace the Learner’s Operation
Peer learning is valuable partly because learners can hear alternative reasoning in language different from the tutor’s. A peer may expose a useful contrast or make a hidden assumption visible.
But timing matters.
If Denise explains the full method before Emily has attempted the problem, Emily’s later success may partly reflect imitation. The explanation can still teach. It no longer provides clean evidence of independent method selection.
The Peer Answer Leakage Boundary owns this evidence-contamination problem in detail. The present volume adds another question: after peer discussion, does the tutor begin treating the stronger peer’s method as the only respectable destination?
A valid alternative method should remain valid if it satisfies the task. Peer explanation should expand reasoning, not impose social conformity.
7. The Strongest Learner Can Also Be Harmed by the Trap
The peer baseline trap does not only harm learners who are slower.
The fastest learner can become the permanent explainer, model answer or informal assistant. This may look flattering while quietly reducing their own learning.
If Alicia always finishes first and is immediately asked to help others, she may receive less Frontier work, fewer changed-condition problems and less opportunity to encounter productive difficulty. Her speed becomes a service to the group instead of evidence that her route may need extension.
Peer explanation can deepen learning when it requires genuine reconstruction and responds to a meaningful question. It should not become unpaid teaching duty attached to being ahead.
The tutor owes every learner an appropriate next job.
8. Three Learners Can Share a Task Without Sharing a Criterion of Progress
Consider one algebra task used with three learners.
- Alicia’s active job is speed after demonstrated accuracy.
- Beatrice’s active job is selecting the correct operation without a topic label.
- Ciara’s active job is keeping sign control stable across a longer chain.
They can solve the same equation. The tutor is looking for different evidence.
Alicia finishing first is relevant to her fluency job. It is not automatically evidence that Beatrice or Ciara should also be judged by speed. Beatrice choosing the correct first step after thoughtful delay may be the important success. Ciara completing slowly with stable signs may be progress relative to her active weakness.
This is one reason small-group tutoring requires more than putting learners at the same table. The tutor has to hold multiple learner jobs inside a shared activity.
9. Compare to Criteria First, Then to Prior Self, Then to Peers When Useful
A practical comparison order helps.
First: criterion comparison. Did the learner meet the actual success condition under the declared support conditions?
Second: longitudinal comparison. Has the learner changed relative to earlier relevant attempts? Is accuracy more stable? Is less prompting needed? Does the method survive a changed surface?
Third: peer comparison. Does seeing another learner reveal a useful alternative method, a pacing issue, a grouping mismatch or an opportunity for collaborative reasoning?
This order does not ban peer comparison. It prevents peers from defining the target.
10. Grouping by Current Skill Is Not Permanent Ranking
National Student Support Accelerator guidance notes that small-group tutoring may group learners intentionally based on current skill and that regrouping may be needed because relative skill levels change over time. That is an important word: current.
A current grouping decision is not a permanent ability label.
The Grouping Gate owns the decision about who should share a tutorial. The peer baseline trap occurs after grouping: even a sensible current grouping can become unhealthy if relative positions are treated as identities.
The learner who is fastest in algebra may be slowest in geometry. The learner who needs support today may lead the group next month. Relative performance is task-specific and time-sensitive.
11. A Constructed Case: Beatrice Is “Slow”
This is a fictional composite teaching case.
Beatrice is regularly the last of three learners to finish mathematics word problems. After several weeks, “slow” has become the room’s shorthand.
The tutor tests the label instead of accepting it.
On direct computation questions, Beatrice works at roughly the same pace as the other two. On word problems, she spends much longer building a representation but then makes fewer method-selection errors. When the tutor removes the need to write full sentences during planning, her start time improves. On unfamiliar changed-condition problems, she transfers more reliably than the learner who usually finishes first.
“Slow” was too coarse. The useful description is more specific: Beatrice invests more time in representation, and some of that time is productive. The tutor can now decide whether representation needs streamlining without destroying the behaviour that supports transfer.
Peer comparison surfaced the difference. Criterion-based diagnosis explained it.
12. A Constructed Case: Alicia Is “The Strong One”
Alicia finishes most routine tasks first. The group begins asking her for confirmation before checking their own work.
The tutor notices that Alicia’s role is distorting everyone.
For the other learners, her answers have become an unofficial answer key. For Alicia, routine work is consuming time that could test deeper transfer. The tutor changes the orchestration. Everyone commits to an answer before discussion. Alicia receives a changed-condition extension after demonstrating the routine method. Peer explanation happens only after first attempts are preserved.
Alicia remains a valuable peer. She stops being the group’s measurement instrument.
13. A Constructed Case: Ciara’s Different Method Looks “Wrong” Because It Is Not the Group’s Method
Ciara solves a percentage question using a unitary approach. The other two learners use a multiplier. Their method is faster.
The tutor could push Ciara immediately toward the faster method because it has become the local norm. Instead, the tutor checks whether Ciara’s approach is mathematically valid, whether she understands why it works, and whether the method remains efficient enough for the present stage.
If the multiplier method is an important curriculum goal, it can be taught as an additional representation. But Ciara’s valid reasoning should not be marked as conceptually weak merely because peers chose differently.
Peer diversity is useful precisely because it can reveal multiple pathways. Standardising every pathway too early destroys that benefit.
14. The Tutor Needs a Private Learner Model, Not a Public Ranking
To teach three learners well, the tutor may need to remember different weak links, supports and goals. Those distinctions should guide instruction without becoming public status labels.
“Alicia is working on transfer.” “Beatrice is working on representation speed.” “Ciara is checking sign stability.” These are temporary instructional statements.
“Alicia is the smart one.” “Beatrice is slow.” “Ciara is weak.” These are identity claims that exceed the evidence and can change how everyone behaves.
The Expectation Reset explains why old labels should not decide what fresh work means. In a group, fresh peer comparison can create new labels just as easily as historical marks can.
15. Keep Some Evidence Private Until Everyone Has Committed
Good small-group orchestration often controls the order in which information becomes public.
For diagnostic questions, collect individual responses first. For opinion or interpretation tasks, allow think time before the most confident learner speaks. For method comparison, preserve each first method before showing alternatives. For estimation, record three estimates before revealing the calculation.
This preserves independent evidence and makes later comparison more meaningful. The group can then ask: why did our methods differ? Which assumptions changed? Which method is more efficient under these conditions?
Peer comparison becomes an object of reasoning rather than a race.
16. Use the Fast Learner to Expand the Question, Not Close It
When one learner finishes first, the tutor has several better options than “wait” or “teach the others”.
- Ask for a second method.
- Change one condition and ask whether the method still works.
- Ask the learner to identify the first place someone could reasonably go wrong.
- Ask for a counterexample to an overgeneralised rule.
- Ask the learner to predict which surface changes would not affect the deep structure.
- Move to a short Frontier task while preserving the group’s shared topic.
This keeps the stronger learner learning without turning them into the answer dispenser.
17. Use the Slower Learner’s Process as Evidence, Not an Embarrassment
A learner who takes longer may provide more observable process. The tutor can see where representation changes, what is checked, which step is uncertain and whether a self-correction occurs.
That does not mean slowness is always desirable. It means the extra time contains information.
If the active job is eventually fluent performance, the tutor can measure whether the process becomes more efficient over repeated correct attempts. The intervention should respond to actual bottlenecks rather than to the social discomfort of finishing after a peer.
18. Peer Assessment Should Return to Criteria
Peer assessment can work well when learners judge work against explicit criteria rather than against personal preference or status.
The NSW Department of Education’s peer-assessment guidance repeatedly anchors peer feedback to success criteria. That principle helps in tuition too.
Instead of “Alicia’s answer is better”, ask “which answer states the claim, gives the relevant evidence and explains the relationship?” Instead of “Beatrice’s method is too long”, ask “which steps are necessary, and where could efficiency improve without losing validity?”
Criteria turn peer comparison into analysis.
19. Three Modes, Three Different Peer Risks
The three established tuition modes remain Repair, Alignment and Frontier.
In Repair, the peer baseline trap can make a learner feel deficient because another learner does not share the same weak link. The tutor should isolate and repair the target without turning difference into rank.
In Alignment, peer comparison can reveal school-pace or execution differences, but success should still be defined by curriculum demands and current learner performance, not the quickest student.
In Frontier, the strongest learner can be held back if the tutor uses them mainly to support peers. Extension should deepen their own capability while preserving group cohesion.
The mode changes what the peer difference means.
20. Tutor Function Also Changes the Meaning of Comparison
The Tutor Classification Model describes functions rather than permanent human ranks.
A Class 2 Drill Builder may use peer pacing to notice fluency differences. A Class 3 Diagnostic Tutor uses individual errors to discriminate causes. A Class 4 Route Designer may notice that two learners no longer belong in the same sequence. A Class 5 Performance Coach may use timed comparison to identify execution variability. A Class 6 Learning Architect may examine how group roles are affecting independence.
The comparison is only useful when it changes the correct tutoring function.
21. What Current Tutoring Research Says About Small Groups
The National Student Support Accelerator’s 2026 review of five years of tutoring research describes trade-offs between one-to-one and small-group formats. One-to-one tutoring often provides more individualised instruction and relationship-building, while small groups can remain effective and efficient in some settings.
An NSSA summary of a pilot online middle-school mathematics experiment found suggestive evidence favouring one-to-one over three-to-one tutoring in that specific setting, while other studies in other subjects and ages have found small groups competitive. These differences are a reminder not to turn one ratio into a universal law.
NSSA’s small-group facilitation guidance emphasises understanding group members, managing participation and planning intentionally. The value of a small group therefore depends partly on what the tutor does with the interactions the format creates.
This volume does not claim that three learners are always better than one, or vice versa. It addresses one operational risk once learners are together: peer information should enrich personalisation, not replace it.
22. Feedback Should Be Relative to Learning Goals
The Education Endowment Foundation’s feedback evidence defines feedback in relation to learning goals or outcomes. That framing is important here.
“You are slower than Alicia” is comparison, not useful feedback unless relative speed is itself the defined performance issue and the comparison helps specify a next action. “Your method is accurate; the next goal is to reduce the repeated checking step while preserving accuracy” connects evidence to a learner goal.
Peer difference may alert the tutor to investigate. Feedback should return to what this learner needs to improve.
23. The Peer Baseline Card
- Target: What is this learner supposed to demonstrate?
- Criterion: What counts as success independent of peer performance?
- Support condition: What help, examples or accommodations were available?
- Own history: How does this attempt compare with relevant earlier attempts?
- Peer signal: What exactly did the peer difference reveal: speed, method, confidence, accuracy, representation or prior familiarity?
- Alternative explanation: What else could explain that difference?
- Public/private: Does the comparison need to be discussed in front of the group?
- Next move: Does the peer information justify changing this learner’s route?
- Strong learner: Is the fastest learner still receiving an appropriate learning job?
- Group health: Are learners becoming dependent on one peer’s answer or approval?
If the tutor cannot name what the peer comparison adds, it probably should not drive the decision.
24. Language Matters More Than It Seems
Group identities are built through repeated small phrases.
“You are always the fast one.” “Ask her; she knows.” “You are the careful one.” “He is weak at word problems.” These statements can freeze temporary patterns into social roles.
Use task-specific language instead: “You selected the structure quickly on these three questions.” “Your checking caught two sign errors today.” “This representation took longer, but your transfer was stable.” “On this topic, we need to repair the first step.”
Specific language preserves the possibility of change.
25. Parent Communication Can Recreate the Same Trap
A parent may ask, “How is my child compared with the other two?”
Sometimes relative information matters, but privacy and educational usefulness place limits on what should be discussed. The tutor does not need to reveal another learner’s marks or route to explain the child’s progress.
A stronger response is criterion- and evidence-based: “She can now solve these independently and her error rate has fallen, but method selection is still unstable when the topic label disappears. The next step is mixed practice and a delayed check.”
This tells the parent far more than “middle of the group”.
26. When Peer Comparison Is Genuinely Useful
There are legitimate uses.
- Comparing methods can expose deep structure and trade-offs.
- Comparing explanations can clarify success criteria.
- Comparing pacing can reveal when fluency deserves attention.
- Comparing error patterns can show whether a task is generally confusing rather than learner-specific.
- Comparing approaches can normalise productive struggle and multiple routes.
- Comparing current skill profiles can inform grouping and regrouping.
- Peer dialogue can improve reasoning when first attempts are preserved and participation is managed.
The rule is not “never compare”. It is “know what the comparison is evidence for”.
27. Common Peer-Baseline Failures
- The fastest-is-best fallacy: speed is treated as the whole construct.
- The public rank: temporary differences become stable group identities.
- The pacing pull: the whole room advances because one learner is waiting.
- The answer-key peer: one student becomes the source of confirmation for everyone.
- The helper tax: the strongest learner loses extension because they are always asked to teach peers.
- The method conformity trap: a valid different method is downgraded because it is not the group’s dominant method.
- The equal-time mistake: fairness is confused with identical tutor attention rather than appropriate opportunity.
- The comparison report: parents receive social rank instead of usable evidence about the learner.
- The fixed group identity: current grouping is mistaken for permanent ability.
- The contaminated diagnosis: peers reveal answers before independent evidence is collected.
28. The Human Standard
A three-student tutorial is valuable because three minds create more than three piles of work. Learners can hear alternatives, notice differences, challenge explanations and discover that a difficult question can be approached in several ways.
That value disappears if the group becomes a small ranking machine.
The tutor’s job is to preserve the educational signal inside the comparison. One learner’s speed can reveal a fluency contrast. Another’s careful representation can reveal a robust method. A third learner’s question can expose an ambiguity everyone missed. These differences are useful because they are different.
A peer can show another possible way to learn. A peer should not quietly become the ruler with which every learner is measured.
Return to the task. Return to the criterion. Return to the learner’s own evidence. Then use the group to widen what can be seen.
That is how a three-student room remains personalised.
That is the Peer Baseline Trap.
That is Tutor Handbook Volume 0107.
Connected Reading and Sources
- The Tutor Handbook | Complete Series Index
- The Tutor Handbook Vol No.0077 | The Three-Learner Orchestration
- The Tutor Handbook Vol No.0081 | The Grouping Gate
- The Tutor Handbook Vol No.0094 | The Peer Answer Leakage Boundary
- The Tutor Handbook Vol No.0101 | The Expectation Reset
- National Student Support Accelerator | Five Years of Tutoring Research
- National Student Support Accelerator | Effective Facilitation Guidelines: Small Group Tutoring
- National Student Support Accelerator | Student-Tutor Ratios Pilot Evidence
- Education Endowment Foundation | Feedback
- NSW Department of Education | High-Impact Formative Assessment Practice
- NSW Department of Education | Strategies for Student Peer Assessment