What Is Item Discrimination? A Quiz Quality Metric
- 1.How it's calculated (simplified)
- 2.Why it matters
- 3.How to find low-discrimination questions
- 4.Related concepts
- 5.A worked example
- 6.Sample sizes for reliable discrimination
- 7.Related reading
- 8.How to compute item discrimination
- 9.What discrimination scores mean
- 10.Why items have low or negative discrimination
- 11.Discrimination vs. difficulty
- 12.When discrimination doesn't apply
- 13.Using discrimination to improve your bank over time
Short answer. Item discrimination is a statistical measure of how well a quiz question distinguishes between students who know the material and students who don't. A good question is one that strong students get right and weak students get wrong; a bad question is one where the pattern is reversed or random.
How it's calculated (simplified)
For each question, you compute:
D ranges from -1 to +1:
Why it matters
A high-difficulty question can still be useful if it has good discrimination. A low-difficulty question with poor discrimination is worse than useless — it makes the quiz feel rigorous without testing anything.
The classic "broken question" pattern: question has poor discrimination because one of the distractors is *actually defensible*. Strong students see the ambiguity and pick the "wrong" answer; weak students don't notice and pick the keyed answer. Fix: rewrite the distractor.
How to find low-discrimination questions
Most LMSs and quiz tools surface this:
After a quiz, glance at the discrimination indices. The 3-5 lowest-scoring items are candidates for rewriting before the next administration.
Related concepts
A worked example
20 students take a quiz. On Question 7:
On Question 8:
On Question 9:
Sample sizes for reliable discrimination
Ready to create your first quiz?
Use AI to generate quizzes from your own study materials in seconds.
Create a Free Quiz — Sign UpThe discrimination index is noisy on small classes. Rough guidance:
For high-stakes item banks (AP, MCAT, SAT), each item is piloted across thousands of students before going live.
Related reading
How to compute item discrimination
The point-biserial correlation between scoring this item correctly and total exam score:
Worked example: 100 students take a 50-question quiz. Top 27 average 85% on item 12; bottom 27 average 35%. D for item 12 = 0.85 - 0.35 = 0.50. Strong discrimination.
Most assessment platforms compute this automatically. If yours doesn't, export to a spreadsheet.
What discrimination scores mean
Why items have low or negative discrimination
Common causes:
A negative-discrimination item should be removed or rewritten before being used again.
Discrimination vs. difficulty
These two metrics together identify item health:
The most-improved exams use a portfolio of items spanning the upper-right quadrant (decent difficulty, strong discrimination) plus a few high-difficulty discriminators to spread the top of the curve.
When discrimination doesn't apply
A few situations where item discrimination isn't the right metric:
Using discrimination to improve your bank over time
The point of computing discrimination isn't to grade items academically; it's to maintain a quality bank:
Get weekly study & quiz tips
Join teachers and students who get practical tips on quizzing, active recall, and AI-powered learning.
Sarah Mitchell
Curriculum Designer & Former High School Teacher
More articles by Sarah →
Practice with AI-generated quizzes
Ready to create your first quiz?
Use AI to generate quizzes from your own study materials in seconds.
Create a Free Quiz — Sign Up