15.5 Using Assessment to Guide Instruction, Goal Setting, and Self- and Peer Assessment
Key Takeaways
- Assessment data guide four distinct decisions: what to reteach, how to group and differentiate, what to communicate to students and families, and how to revise the unit and program.
- Feedback that is specific, criterion-referenced, and delivered while the student can still act on it changes performance; a score attached to a final grade does not.
- Students who use assessment data to set their own goals develop the self-management capability that lifelong physical activity requires.
- Self-assessment must be taught -- students need the criteria, a model of accurate self-rating, and practice -- or their ratings will reflect confidence rather than performance.
- Peer assessment is reliable when students use a specific criteria card focused on observable components, and it should inform feedback rather than determine grades.
Assessment as a Decision Tool
Assessment that ends in a gradebook entry has done a quarter of its job. Its value lies in the decisions it informs.
| Decision | Question the data answer |
|---|---|
| Reteach or advance | Did enough students meet the objective to move on? |
| Group and differentiate | Who needs support, who needs extension, and how should stations be configured? |
| Communicate progress | What do the student and family need to know, and in what terms? |
| Revise the unit and program | Was the time allocation right, the sequence sound, the assessment aligned? |
A useful analytic discipline is to look at three things after any assessment: the class pattern (what most students missed), the spread (how wide the range is), and the outliers (who is far outside the group, in either direction, and why).
Feedback That Improves Performance
The research consensus is consistent: feedback changes performance when it is
- Specific -- naming what was done and what to change
- Criterion-referenced -- tied to the stated criteria, not to other students
- Timely -- delivered while the student can still act on it
- Actionable -- describing a next step the student can take
- Focused -- one or two points rather than a full audit
And it does not change performance when it is a score alone, a comparison to classmates, or arrives after the work is finished and graded. The practical consequence for physical education is that assessment must be built into the unit with revision opportunities, not saved for a final performance day.
Communicating progress
- Report in terms of criteria and standards, not rank
- Show growth from the student's own baseline
- Be specific about what the student can now do and what comes next
- Report behavior and responsibility separately from learning
- For families, translate the standards into plain language and give a concrete way to help
Guiding Personal Goal Setting
Assessment data become a student's own when they use them to set goals. The sequence:
- Interpret the result with the student -- what does this actually mean?
- Identify one target the student cares about
- Write a SMART goal -- specific, measurable, achievable, relevant, time-bound
- Build a plan using FITT for fitness goals or a practice progression for skill goals
- Anticipate barriers with an if-then plan
- Monitor with a specified method
- Reassess and revise
Grade the goal, the plan, the monitoring, and the reflection -- never the attainment. A student whose asthma or growth spurt prevented the outcome should not be penalized for a well-designed plan they executed faithfully.
The long-term purpose is self-management. The graduate who can assess their own fitness, set a goal, design a plan, and adjust it is the outcome the subject exists to produce.
Self-Assessment
Self-assessment builds the error-detection capability that separates a dependent performer from an independent one, and it is directly connected to the fading of teacher feedback.
It has to be taught. Untrained self-rating measures confidence, and it is systematically inaccurate in both directions -- low-skilled students often overrate, and some capable students underrate.
Teaching sequence:
- Give the criteria in student-accessible language
- Model self-rating aloud against a video or a live performance
- Calibrate -- students rate a common performance, then compare their ratings with the teacher's and discuss the differences
- Practice on their own performance, ideally with video so they can see what they did
- Verify by comparing student self-ratings with teacher ratings on a sample
Instruments: criteria checklists, video self-review, effort and responsibility self-ratings, goal-progress logs, and exit reflections.
Peer Assessment
Peer assessment doubles the feedback available in a class of 35 and builds the caring-and-helping responsibility level, but it fails predictably when students are asked to evaluate without tools.
Conditions for reliability:
- A specific criteria card listing observable components -- "steps with the opposite foot: yes / not yet" rather than "good form"
- Explicit instruction in what to look for and where to stand to see it
- Focus on the behavior, not the person -- students report against criteria, they do not judge classmates
- Feedback language taught -- what to say and how
- A limited number of criteria, typically two or three
- Rotating partners so the same pair does not always work together
- Teacher verification of a sample, which also communicates that the task is taken seriously
Peer assessment informs feedback; it should not determine grades. Using peer scores as grades introduces friendship effects, retaliation, and social pressure into an evaluation students are not qualified to make.
The reciprocal teaching style is the standard structure: partners alternate as doer and observer, the observer uses the criteria card, and the teacher circulates supporting the observers rather than the performers -- which is a deliberate shift, because improving the observers improves everyone's feedback.
Closing the Loop
A complete assessment cycle in a unit looks like this:
- Preassess to establish baseline and expose misconceptions
- Share the criteria with students before instruction
- Formatively assess during practice, with feedback and revision
- Have students self- and peer-assess against the same criteria
- Summatively assess against the criteria the students have had all along
- Report learning against standards, with growth from baseline
- Use the data to reteach, to guide student goals, and to revise the unit
The criteria are the same document throughout. That consistency is what makes the assessment fair and the instruction coherent.
Common Praxis Traps
- Trap 1: Assessment that ends in a grade. It should drive four decisions.
- Trap 2: Feedback after the work is finished. Build in revision opportunities.
- Trap 3: Grading goal attainment. Grade the goal, plan, monitoring, and reflection.
- Trap 4: Assuming students can self-assess without instruction. Teach and calibrate it.
- Trap 5: Peer assessment without a criteria card. It becomes judgment of the person.
- Trap 6: Using peer scores as grades. Peer assessment informs feedback.
A student sets a goal to improve their PACER score, follows the training plan faithfully, monitors weekly, and writes a thoughtful reflection, but their score does not improve because of an asthma flare-up. How should this be graded?
A teacher asks students to rate their own serving technique with no further guidance, and finds the ratings bear little relationship to actual performance. What is missing?
Which condition most improves the reliability of peer assessment in physical education?
After a formative assessment shows that most students misapply a tactical concept, which use of the data is most important?