Awesome Psychology Tasks / measurement map

Open call

Review one spec. Three claims. An afternoon.

Every paradigm specification in this catalog is currently sourced — literature-linked, drafted with LLM assistance, and unreviewed by a named expert. We are looking for the first reviewers to change that, one spec at a time.

What we are asking for

Not an endorsement of the project. Not a paper. One spec, read against three separable claims, each signed or contested by name. The mechanical work is done: the spec is drafted, the references are DOI-verified, the executable contract runs, and a dossier lays out exactly what the machine can and cannot show you for each claim.

The three claims

Claim 1

Contrast identification

Does the defining contrast correctly identify the paradigm — did we draft the right operationalisation?

Whether this contrast is the defining one, versus the alternatives, is the part no test can settle. You are reviewing our authorship.

Typically a cognitive psychologist (attention / control)

Claim 2

Construct operationalisation

Is the contrast a defensible measure of the construct it is named for?

The instrumentation layer is silent here by construction: a task can be wired exactly as specified and still not measure what the name promises.

Typically a construct-validity / individual-differences view

Claim 3

Psychometric conditioning

Are the reported numbers correctly conditioned — population, version, trial count — and not over-generalised?

Effect magnitudes and reliability estimates hold under conditions. The spec has to say which, or it is quietly wrong.

Typically a methodologist / psychometrician

How a sign-off changes the record

expert-annotated is not a badge anyone can award themselves. It is derived from the three claim-level attestations: all three attested by named reviewers and the record derives expert-annotated; any one contested and it derives contested; otherwise it stays where it is. The validator enforces the agreement, so no machine check can ever mint human review.

Specs open for review

All of them, today. Pick the one closest to your work — the Eriksen Flanker Task has the most supporting material assembled and is the intended first vertical.

SpecificationFamilyCurrent state
Color Delayed Estimation (Visual Working Memory)continuous-reportsourced — unreviewed
Eriksen Flanker Taskcongruencysourced — unreviewed
Go/No-Gogo-no-gosourced — unreviewed
Lexical Decision Taskspeeded-classificationsourced — unreviewed
Motion Coherence Discrimination (Random-Dot Kinematogram)psychophysicssourced — unreviewed
N-backn-backsourced — unreviewed
Sequential Congruency (Gratton effect)sequential-congruencysourced — unreviewed
Simon Taskcongruencysourced — unreviewed
Stop-Signal Taskstop-signalsourced — unreviewed
Stroop Taskcongruencysourced — unreviewed
Task Switchingtask-switchingsourced — unreviewed

To volunteer

Open an issue on the repository saying which spec and which claim(s) you would take, and we will send the attestation dossier for that spec — the spec, its references, the evidence the harness can show for each claim, an explicit list of what it cannot show, and a scope template for your sign-off.

Prefer not to use GitHub? Reach the maintainer through the maintainer's GitHub profile.