Assessment and Feedback
Assessment Methods for Educators: Options, Evidence, and Review

Why compare assessment methods
Assessment methods collect different kinds of evidence. Keep conventional selected-response tests when they fit the construct, and add or replace them only when another method supplies more appropriate evidence for the stated purpose.
Alternative assessment formats can provide different evidence than a conventional selected-response test. Their validity, reliability, accessibility, workload, and fit depend on the assessment purpose and implementation.
This guide compares assessment options by the evidence they elicit, scoring requirements, access demands, and limitations.
Formative assessment: evidence during learning
Formative assessment collects evidence while there is still time to adjust instruction. It is useful only when the prompt is aligned to the objective, responses are interpreted cautiously, and the teacher takes a feasible next step; an immediate response is not a complete account of what a student knows.
Short formative checks
- Exit tickets - Ask students to answer one question connected to the learning target before leaving. Use the response to identify a stated point of confusion or a next instructional step.
- Digital polls - Use an approved, accessible tool for a question tied to the learning target, protect response privacy, and provide a non-digital option.
- Quick writes - Give students 3-5 minutes to explain a concept in writing
- Private response cards - Let students indicate a response privately or use another accessible route; a public hand signal can expose uncertainty and may not show the reason for an answer.
The linked formative-assessment examples may offer useful prompts, but a quick response does not automatically create a continuous feedback loop or reveal which concept needs reinforcement. Interpret the response against the objective and choose a specific next step.
Performance-Based Assessments - Real-World Applications
Selected-response tests can efficiently sample recall, recognition, and some forms of reasoning. Performance tasks may be preferable when the objective requires students to produce, explain, design, or perform, but they also add scoring, access, and workload demands.
Performance-based assessments ask students to apply knowledge or skill in a product, explanation, demonstration, or decision. Use a letter, design task, or model only when it elicits the intended construct and can be scored fairly; a real-world appearance does not make a task valid or meaningful by itself.
Bringing Authenticity to Assessment
The University of Maryland overview of authentic assessment approaches can help define a task, audience, and criteria. Relevance beyond the classroom does not guarantee deeper engagement, so review participation and work evidence directly.
- Choose an appropriate problem - Use a simulated or locally approved issue that fits the objective; do not require students to disclose personal circumstances or act outside the school's role.
- Create clear rubrics - Co-create assessment criteria with students so expectations are transparent
- Build in reflection - Have students analyze their process, not just their final product
- Incorporate revision cycles - Allow students to improve based on feedback, just like in professional settings
For example, a community-garden design task can elicit measurement, science, and written-explanation evidence when the learning criteria are explicit. Compare that evidence with the intended construct rather than claiming a real classroom result or assuming an audience caused deeper learning.
Computer-adaptive assessments
Computer-adaptive assessments select later items partly from earlier responses. Before adoption, review the intended population, item bank, scoring model, accessibility, privacy, technical requirements, and independent validity evidence for the proposed use.
Changing item difficulty can alter which questions a student receives. It does not by itself create a personalized learning experience or more accurately measure a student's abilities; those claims require reliability, validity, fairness, and comparability evidence for the specific use.
Claims to verify before adoption
Vendors may claim that adaptive testing improves efficiency, precision, and reporting. Require independent technical documentation for the intended population and decision, and do not treat a vendor summary as evidence of student outcomes.
- Efficiency - Test length and time may change with the item bank and stopping rules; verify the effect for the intended population.
- Precision - Review conditional standard errors, reliability, validity, and fairness rather than assuming more accurate measurement across a broad ability range.
- Challenge level: Adaptive delivery changes item difficulty, but do not infer reduced anxiety, confidence, or validity without relevant evidence.
- Reporting - Inspect what each reported score or category means, what uncertainty it carries, and whether teachers can use it without labeling students or exposing sensitive data.
Peer and self-assessment: roles and limits
Peer and self-assessment can give students practice applying stated criteria to sample or draft work. It may add rather than reduce teacher workload, and it does not establish durable metacognition or later-life benefits; teachers remain responsible for consequential judgments.
Self- and peer-assessment can provide practice applying stated criteria. Model the criteria, protect privacy and relationships, retain teacher responsibility for consequential judgments, and examine the feedback produced rather than inferring durable judgment or future outcomes.
Implementing Student-Centered Assessment
For a bounded peer- or self-assessment activity:
- Teach feedback skills explicitly - Model constructive, specific comments
- Use clear rubrics - Provide concrete criteria students can apply consistently
- Structure reflection prompts - Guide students to identify specific strengths and next steps
- Use a feedback protocol - Structure comments around the stated criteria and a feasible next step; do not assume a slogan balances encouragement or produces growth.
Review before expanding an assessment method
Adding formative checks, performance tasks, adaptive technologies, or student review expands the kinds of evidence available. It does not automatically make the evidence richer or more accurate; compare validity, reliability, accessibility, privacy, scoring consistency, and workload for the stated use.
No assessment format is inherently better. Review whether the task elicits the intended construct, supports access without changing that construct, and produces defensible scoring evidence.
Pilot one assessment strategy with one learning target. Review the evidence quality, student access, scoring effort, and next instructional decision before expanding it; teacher enthusiasm is not evidence that the method improved learning.
Changing an assessment format can change the evidence collected, but it does not by itself change how students learn or grow. Keep, revise, or remove the method based on the stated construct and observed results.