Awesome Psychology Tasks / measurement map

Specifications

N-back

A continuous stream of stimuli is presented one at a time; on each item the participant signals whether it matches the item presented n positions earlier. Holding and continuously updating the last n items taxes working memory, and detection sensitivity falls as the load n rises. Performance is scored with signal-detection theory: a 'match' to a true n-back item is a hit, a 'match' to a non-matching item is a false alarm, and the bias-corrected sensitivity d' is the primary index. The load-dependent decline in d' — not mere match detection — is the paradigm's signature.

Defining contrast
d-prime (load-response)
Family
n-back
Human review
sourced — unreviewed
Executable test
verify:task — an implementation can be driven and judged against this spec
Schema
spec version 4
Catalog entry
N-back in the paradigm catalog

Where this spec stands

This record is sourced. The ladder below is the human-review axis only — a passing executable test never moves a record up it.

Implementations driven through this spec

ImplementationPlatformVerdictNote
demos/n_backPavlovia / PsychoJSconformsmouse-driven; oracle policy recovered, d' ≈ 3.08 — an n-back the survey's static probe could only call undeterminable, shown conformant when driven

Origin reference

Review this specification

Three separable claims: does the contrast identify the paradigm, is it a defensible measure of the named construct, and are the psychometric numbers correctly conditioned? Attest one, contest another — a split outcome is a real result.