Methodology
What we measure, and how honestly we can do it
MINDLY.help measures five cognitive skills, trains those same skills, and re-measures them monthly on forms you have not trained. It is not a medical device. This page states exactly what the battery measures, where each task comes from, which parameters we run against the published ones, how raw performance becomes a score, and where the measurement breaks down.
The tasks and their sources
Sustained attention
Sustained attention to response task (SART)
A digit appears, you press for every digit, and you hold on the 3. Responding turns automatic within seconds, which is the design: the rare withhold is where a slip of attention shows itself. Core metrics are commission errors (presses on the 3), omission errors and reaction-time variability. Robertson runs 225 trials with the digit on screen for 250 ms and a 900 ms mask. We run 81 trials with the digit on screen for 500 ms and a 1500 ms trial: at the published rate a first-time participant believes the response window has closed the moment the digit disappears, and abandons trials that were still open. The daily training levels compress the trial from 1500 ms towards the published rate, so the canonical pace is where the training goes rather than where it starts.
- Robertson, Manly, Andrade, Baddeley & Yiend, 1997 · Neuropsychologia, 35(6), 747–758 ↗
- Manly, Robertson, Galloway & Hawkins, 1999 · Neuropsychologia, 37, 661–670 ↗
Inhibitory control
Stroop color–word task
Name the ink color of a word whose meaning conflicts with it. The interference effect (incongruent minus congruent RT) indexes how well you suppress an automatic response. Parameters follow common computerised versions: 48 trials at 50% congruent, an 800 ms response-stimulus interval and a 3 s deadline. The fourth colour is purple rather than yellow, because yellow is unreadable on a white background and an easier-to-read colour would bias the interference estimate.
Selective attention
Eriksen flanker task
Respond to the central arrow while flanking arrows point the same or the opposite way. The flanker effect (incongruent minus congruent RT) measures interference from irrelevant stimuli. Parameters follow the Attention Network Test convention: 40 trials balanced across direction and congruency, a jittered 400–800 ms fixation so the onset cannot be anticipated, and a 1500 ms response window.
Working memory
Digit span, forward and backward
Recall digit sequences of growing length, forward then backward. Administration follows the Wechsler convention: sequences start at two digits, digits are presented at one per second, two trials at each length, and the block ends after failing both trials at the same length. Chosen over n-back because its test–retest reliability is far better documented.
Processing speed
Symbol–digit substitution
Match symbols to digits against an on-screen key for 90 seconds. Correct-per-minute is among the most age- and state-sensitive psychometric measures, in the same family as the only training line with durable effects (speed-of-processing / UFOV).
What we claim — and what we refuse to claim
We claim to measure specific cognitive skills — attention, working memory, cognitive control, processing speed — using standard tasks from the research literature, and to show your personal trend over time. Our methodology and sources are open.
We also claim to train those same specific skills, and we mean that narrowly. Practising a cognitive task reliably improves performance on that task and on tasks built the same way — near transfer. That is a real, repeatedly replicated effect, and it is the whole of what the daily session offers. It is also why the monthly retest runs on freshly generated forms you have not been practising: without that, a rising score would only prove you had learned one screen.
We do not claim to raise IQ, improve job or academic performance, prevent or treat any medical condition, or reverse any effect of short-form video or AI tools on cognition. Controlled studies of brain training consistently find improvement on the trained tasks with little to no transfer to untrained abilities (“far transfer”). In 2016 the US Federal Trade Commission fined Lumosity $2 million for making such claims without evidence (FTC v. Lumos Labs, 2016). We take that ruling as a design constraint.
How scoring works
Each task produces raw metrics (error rates, reaction-time costs, span length, throughput). We convert the primary metric of each task into a percentile against reference parameters — population-typical means and spreads for healthy adults on the standard versions of these tasks. These parameters are preliminary: they are not clinical norms and they are not calibrated to your hardware. Treat a percentile as “roughly where this lands for a typical adult,” not as a diagnosis.
The limits of browser timing
Reaction time in a browser is affected by display refresh, input polling and operating-system scheduling — on the order of 5–10 ms of jitter, plus a device-specific constant offset. Two consequences:
- Within-subject comparisons are solid. Comparing you to yourself, on the same device and browser, cancels most of the hardware offset. Your trend over repeated baselines is the metric we stand behind.
- Between-subject comparisons are noisy. Your absolute RT versus another person on different hardware is polluted by both devices. That is why our population percentiles are labeled preliminary.
A fresh sequence every run
Every battery draws a new random seed, and every task derives its stimuli from it: the order of go and no-go trials, the colour–word pairings, the arrow arrays, the digit sequences and the symbol–digit key are all regenerated. You cannot memorise your way through a second run. What you can do is get better at the kind of task — that is the practice effect, it is strongest between your first and second sessions, and it is why your first three runs are marked as calibration on your history page.
Data and privacy
Without an account, nothing reaches our server. Your latest result, the session seed and basic device information (screen size, hardware concurrency) are stored in your browser's local storage under the key mindly:baseline. Clearing browser storage deletes them permanently.
If you create an account, each completed run is additionally stored on the server — the same numbers you see on your results page, plus the device information — so that repeated runs can be plotted as a trend. Registration is an email address and a password, with no confirmation step; the password is kept as an scrypt hash and the session cookie holds a random token whose digest is all we store. We do not send email, and there is no third-party analytics on this site.
The one-line version
We don't claim to make you smarter. We measure specific skills with the same tasks researchers use, train those same skills, and check the result on forms you never trained — then show you the data either way.