Items / Reference templates / BP009
REF-BP009-0
C04 probability rules · C04.2 complement rule · target medium, CD2 · split test · reference items carry their template's native demand and difficulty
Pending
In a clinic record review, each patient independently reports side effects with probability 0.08. 6 patients are selected at random. What is the probability that at least one of them reports side effects?
Explanation (as written by the item's author)
P(at least one) = 1 - (1 - 0.08)^6 = 0.3936.
| Option | Code | Catalogue meaning | Author's rationale |
|---|---|---|---|
| A | M-ADD-NOT-MULT | Adds probabilities of independent events where multiplication is required (or vice versa). | Adds the individual probabilities, which can exceed 1 and ignores overlap. |
| B | M-BINOM-AT-MOST | Confuses P(X = k) with P(X <= k) or P(X >= k). | Computes P(all n) instead of P(at least one). |
| C | M-COMPLEMENT | Forgets to take a complement, or takes 1 - p when p already answered the question. | Computes P(none) and stops, forgetting to subtract from 1. |
| D (key) | KEY | Complement of 'none', which is the efficient route for 'at least one'. |
Quality flags · 1 fired
- inconsistent numeric format (style) firedNumeric options mix fractions, decimals and percentages, or use differing decimal precision.
- numeric middle key (info) firedOptions are numeric and the key is one of the two middle values. Informational; tested at set level against the 0.5 chance rate.
- duplicate options (integrity) not firedTwo options are identical after normalisation.
- numerically equal options (integrity) not firedTwo options denote the same number (e.g. 0.5 and 1/2 and 50%).
- missing options (integrity) not firedFewer than four non-empty options.
- key out of range (integrity) not firedThe keyed answer is a probability outside [0, 1] (or a percentage > 100).
- all none of above (cue) not firedAn option is 'all/none of the above' or 'both A and B'.
- key uniquely longest (cue) not firedThe key is the longest option by at least 20% (length cue).
- key uniquely shortest (cue) not firedThe key is the shortest option by at least 20%.
- distractor out of range (cue) not firedA distractor is a probability outside [0, 1]; it can be eliminated without the construct.
- absolute terms in distractors (cue) not firedDistractors (only) use absolute terms such as always/never/must/proves, a known cue that they are wrong.
- stem option overlap cue (cue) not firedThe key alone repeats a content word from the stem.
- grammatical mismatch (cue) not firedThe stem ends in 'a'/'an' and some options do not agree with it.
- key is hedged only (cue) not firedOnly the key contains hedging language (may/might/plausibly/likely).
- negative stem (style) not firedStem uses NOT/EXCEPT/LEAST phrasing.
- overlong stem (style) not firedStem exceeds 110 words.
- unit mismatch (style) not firedOptions carry different units.
- option length imbalance (style) not firedLongest option is more than 2.5x the shortest.
- rationale value mismatch (metadata) not firedA distractor's stated rationale computes a value ('= x') that is not the option it is attached to: the distractor does not actually execute the misconception it claims.
Validator output
PoT not yet run for this item.
Answering probes
Probes not yet run for this item.
Shortcut results
| Heuristic | P(key) |
|---|---|
| longest option | 0.50 |
| shortest option | 0.00 |
| convergence | 0.25 |
| stem overlap | 0.25 |
| numeric middle | 0.50 |
| hedged option | 0.25 |
| position prior | 0.00 |
Provenance and generation trace
- Method
- reference_template / —
- Model
- none
- Source
- parameterised template T-C04.2-atleastone, written by the AssessAI build agent (Claude); no external item bank; not human-reviewed
- Prompt hash
—- LLM calls · cost
- 0 · $0.00000 · 0.0 s
- Parser notes
- none