Boring Websites The Blog

Human Benchmark: the website that measures your mind

08 Oct 2026 · By the Boring Websites editors

Human Benchmark is a free website that puts a stopwatch on the mind. It hosts eight plain, timed tests, reaction time, sequence memory, aim, number memory, verbal memory, the chimp test, visual memory, and typing, and returns each one as a cold number with a percentile attached. It is worth a look because it takes the dry apparatus of a psychology lab, strips out the lab, and hands anyone with a browser the single most addictive thing a test can offer: a score to beat.

The site lives at humanbenchmark.com and wastes no time explaining itself. There is a grid of tests, each one a card, and clicking any card drops the visitor straight into it. No account, no tutorial, no preamble about cognitive science. The whole pitch is a reflex away, and the first result usually arrives before a person has decided whether they meant to start.

In brief

  • Human Benchmark is a free browser suite of eight timed tests covering reaction speed, several kinds of memory, aim, and typing.
  • Each test returns a single number plus a percentile, so a score is always framed against everyone else who has taken it.
  • The reaction-time test partly measures a person's screen and input hardware, not just their nerves, a point its own community raises often.
  • The design is deliberately bare: no accounts, no lessons, no gamified rewards beyond the number itself.

Eight tests, one tab

The front page is a contents list rather than a home page. Each of the eight tests sits in its own card with a one-line description, and the set is chosen to sample different corners of performance rather than drill one skill. Reaction Time measures raw visual reflex. Aim Trainer measures speed and accuracy together. The four memory tests each isolate a different flavor of recall, and Typing measures words per minute. The spread is the point: no single number claims to be a measure of intelligence, only of one narrow thing done quickly.

What unites them is format. Every test is short, self-contained, and over in under a minute, and every test ends on a number. That brevity is why the site works as a time sink despite having no levels, no story, and no progression. A person sits down to check one reflex and surfaces twenty minutes later having run the whole grid twice, because each test is cheap enough to retry and the retry is where the hook lives.

The presentation is almost aggressively plain. Flat colors, a sans-serif font, a large target or prompt in the middle, and the result underneath. There is nothing to look at while a test loads and nothing to admire once it ends, which leaves the visitor with only the number and the quiet question of whether they could do better. The restraint is a design choice, not a limitation, and it is most of why the site has outlived flashier competitors.

The reaction-time test, and what it really measures

The reaction-time test is the site's front door and its most argued-over feature. The instructions are trivial: a red box turns green, click as fast as possible, and the page reports the gap in milliseconds across five tries. The average that most adults land on sits somewhere in the low-to-mid 200s, which lines up with the research baseline. Wikipedia's summary of mental chronometry notes that human responses on simple reaction-time tasks are usually on the order of 200 milliseconds, with the mean to detect a visual stimulus around 190 milliseconds for college-age subjects.

That agreement with the literature is part of what makes the test feel legitimate. A person who scores 230 is not getting a made-up number; they are getting a figure in the same neighborhood psychologists have measured for over a century. The test quietly teaches the shape of the distribution, too. Scores below 200 are rare and hard to repeat, scores around 250 are ordinary, and a single lapse in attention can add a hundred milliseconds that no amount of trying claws back on the next try.

Why the number is partly about your hardware

The honest caveat is that the reaction-time test does not measure reaction time alone. It measures a person plus their screen plus their mouse plus the browser, and the machinery adds a floor that the nerves cannot beat. The site's own audience has worked this out in detail. On the 2015 Hacker News submission of the test, Measure Human Reaction Time, commenters compared it against other tools and found gaps of 40 milliseconds or more between them, and one user reported hitting exactly the same figure three times in a row, a suspiciously clean result they traced to input being quantized by the display or the browser.

The mechanism is plain once named. A screen refreshing at 60 hertz can only show the green box at 60 moments a second, so the signal a person reacts to is itself delayed by up to one frame, about 16 milliseconds, before their eyes ever see it. Add mouse polling and browser scheduling and the hardware can account for a sizeable slice of the score. The practical lesson is to read the number as relative rather than absolute: useful for beating a personal best on the same setup, shaky as a universal measure of a nervous system.

The memory tests, from digits to colors

Four of the eight tests probe memory, and each one isolates a different kind. Number Memory flashes a digit string that grows by one digit each round and asks the visitor to type it back, a direct test of how long a sequence a person can hold. Most people stall between seven and nine digits, which echoes the long-standing rule of thumb that working memory holds roughly seven items. The test makes that abstract claim concrete in about ninety seconds.

Sequence Memory lights up tiles in a growing pattern that the visitor must repeat, the old handheld-game mechanic rebuilt as a benchmark. Visual Memory flashes a grid of squares to remember and reproduce, scaling the board up as a person succeeds. Verbal Memory shows words one at a time and asks whether each has appeared before, a test of recognition that punishes a wandering mind more than a weak memory. The four together make a small, pointed argument: memory is not one faculty but several, and a person can be sharp at one and ordinary at the rest.

None of these tests dresses itself up. There is a prompt, a response, and a score, and the score is always framed as a percentile so the visitor knows where they stand against the crowd. That framing is the quiet engine of the whole site. A raw number means little, but a number that says a person is faster than sixty percent of takers and slower than forty is an invitation, and the invitation is always to try once more.

The chimp test, and the research behind it

The strangest card on the grid is the Chimp Test, and it is strange for a good reason. Numbered squares flash on screen and vanish, and the visitor must tap the hidden positions in numerical order, with one more square added each round. The test is named for a famous finding in primate cognition: young chimpanzees, in work led by Tetsuro Matsuzawa at Kyoto University, proved startlingly good at exactly this task, recalling the layout of briefly flashed numbers faster and more accurately than adult humans. The test's existence is a small homage to that result, and a quiet challenge to live up to it.

The card drew its own Hacker News thread, Are You Smarter Than a Chimpanzee?, where the answer, for most people, turned out to be no. Commenters reported topping out well below the span a trained chimp like Ayumu could manage, and the discussion drifted into why. Part of the gap is that humans process the digits as language, reading them, while the chimps appear to hold a purely spatial snapshot. The test cannot settle that debate, but it does something rarer: it lets a person feel the limit of their own visual memory in a way a paragraph about the research never could.

Why a single number is so hard to put down

The compulsive pull of Human Benchmark comes from the gap between how cheap a retry is and how much a score seems to matter. A test takes seconds, costs nothing, and ends on a figure that feels like a verdict. The percentile turns that verdict social even though no one else is watching. A person is not really competing against the millions of anonymous takers behind the percentile; they are competing against the number they got last time, and the site makes that rematch available instantly.

There is also the matter of variance. Reaction time and memory both swing from attempt to attempt, so a person who scores badly can always tell themselves the next try is the real one. That built-in noise is catnip. A deterministic test a person could only fail would get closed; a noisy one that occasionally flashes a great result keeps them clicking, chasing a peak they have already glimpsed once. The site never designed a reward loop on purpose, yet the bare structure of a timed test with a visible best produces one anyway.

That quiet, repeatable self-measurement is a close cousin of the small rituals the Boring Websites network tends to collect. The Bureau of Tiny Approvals, for instance, hands out small official-feeling approvals for no reason beyond the pleasure of receiving one, and Human Benchmark works on the same frequency: a tiny, low-stakes transaction that a person returns to not because it leads anywhere but because the loop itself is satisfying. Both sites understand that a trivial reward delivered reliably can hold attention better than an elaborate one delivered rarely.

The science the site quietly stands on

Human Benchmark did not invent the timed cognitive test; it repackaged a tradition more than a century old. The study of reaction time as a window into the mind, mental chronometry, goes back to the nineteenth century. The Wikipedia overview credits Franciscus Donders with the foundational insight that simple reaction time is shorter than recognition time, which is shorter than choice time, and with devising a subtraction method to estimate how long individual mental steps take. The site's reaction test is a direct, if simplified, descendant of Donders's apparatus.

The idea of measuring individual differences and plotting them against a population came slightly later, from Francis Galton, whom the same overview describes as the founder of differential psychology and the first to run rigorous reaction-time tests specifically to map the average and spread across people. That is, almost exactly, what Human Benchmark does at scale: it gathers reaction times from a vast anonymous population and hands each new taker back a percentile against the lot. The percentile that makes the site so moreish is Galton's project, rebuilt in JavaScript and left running for anyone who wanders in.

Framing the site this way also clarifies what it is not. It is not a diagnostic tool, and it does not claim to be. The tests are too short, too dependent on hardware, and too easy to practice to say anything clinical about a person. What they offer is the same thing early psychophysics offered: a crude but real handle on how fast and how much a mind can do under a stopwatch, with all the caveats that come from measuring something slippery with something simple.

Where it fits, and what it is honest about

Human Benchmark belongs to a specific category of website: the single-purpose instrument that does one measurable thing and resists the urge to become a platform. It has no feed, no social graph, and no lessons to sell. A visitor who wants to improve their reaction time will find the site happy to time them a thousand times and entirely uninterested in coaching them, which is both its limit and its charm. It is an instrument, not a trainer, and it never pretends otherwise.

The site is also refreshingly candid, through its community if not its copy, about how much its numbers should be trusted. The hardware caveat on the reaction test, the practice effect on the memory tests, and the language-versus-spatial wrinkle in the chimp test are all discussed openly in the threads around it. A person who reads those discussions comes away with a better score and a healthier skepticism about what the score means, which is a rare combination for a site built on the promise of a number.

In the end, the appeal is the honesty of the exercise. Human Benchmark does not promise to reveal anything profound. It promises a red box that turns green, a grid that flashes and vanishes, and a figure at the end that a person can either accept or try to beat. Most people try to beat it, discover that attention is harder to hold than they thought, and close the tab a few milliseconds wiser. That is the whole offer, and the site keeps it with a straight face.

FAQ

What tests does Human Benchmark include?

The site at humanbenchmark.com lists eight: Reaction Time, Sequence Memory, Aim Trainer, Number Memory, Verbal Memory, the Chimp Test, Visual Memory, and Typing. Each is short, self-contained, and returns a score with a percentile.

Is a good reaction-time score really about reflexes?

Only partly. As commenters on the test's Hacker News thread point out, the measured number includes delays from the screen's refresh rate, the mouse, and the browser, so it reflects a person's hardware as much as their nerves. It is best used to beat a personal best on the same setup rather than as an absolute figure.

What is the chimp test based on?

It is named for research led by Tetsuro Matsuzawa at Kyoto University, which found that young chimpanzees could recall the positions of briefly flashed numbers faster and more accurately than adult humans. The test recreates that task and, for most people, confirms the chimps' edge.

Is Human Benchmark free, and does it need an account?

Yes, it is free, and no account is required to take any test. A visitor can open the site, click a test, and have a score in seconds, which is much of why it works as a quick, repeatable curiosity rather than a committed app.