Skip to content
AbilityBench

Free mixed drill that finds your slowest format

Twenty mixed items that end in a link, not a score

Four question formats an employer’s battery will draw from — verbal analogies, numerical work, abstract series and mechanical reasoning — five items each, interleaved so that no two consecutive questions are the same kind, inside ten minutes. It is free, needs no account, and offers four untimed warm-up items first. The output is not a total: it is the comparison across the four, and a link to the page on this site that drills whichever one took your time or your accuracy.

  • 100% free
  • No signup
  • 20 items
  • 10 minutes
  • 4 formats interleaved

Twenty items, ten minutes, four formats interleaved so that no two consecutive questions are the same kind. The point is not the total at the end — five items cannot establish an ability — it is the comparison between the four, and the link it produces to whichever one cost you the most.

Verbal 5 items
Analogies: name the relation between two words, then find the pair that shares it.
Numerical 5 items
Percentage change, ratio splits and unit rates — arithmetic under a clock.
Abstract 5 items
Number and letter series where the step itself is the thing that changes.
Mechanical 5 items
Gears, levers and pulleys, described in sentences rather than drawn.

The warm-up is one item per format with no clock on it and an answer shown after each; none of it reaches the result. The ten minutes start at the first scored item, not when you press a button.

How to use a mixed set properly

Warm up untimed, run the ten minutes, then follow the link it gives you.

  1. Take the four warm-up items first unless you have done this before

    One item from each format, with no clock on it and the answer shown afterwards, so that the first time you meet a gear train is not also the first time you are losing seconds to it. None of the four reaches any figure in the result. The ten minutes begin at the first scored item, so a slow warm-up costs you nothing.

  2. Answer on the number keys, and let the pace marker do the worrying

    Keys 1 to 4 pick an option and every option is also a button. Beside the countdown is a line saying how many items ahead of or behind a flat pace you are — a flat pace being 30 seconds an item, which is what reaching all twenty requires. Watch that line rather than the clock: the clock tells you time is passing and the line tells you whether it matters.

  3. Read the strand table across, then click the link

    Each format shows what you got right out of what you reached and your median seconds on it. Where accuracy is level, the slow format is the one worth an evening, because on a real paper the slow format is the one that stops you reaching the end. The panel names one and links to it; the table links all four, so ignore the recommendation if you know better.

Technical specifications

Items and clock20 scored items in 10 minutes — a flat pace of 30 seconds each — plus 4 optional warm-up items with no clock and no effect on any figure
Format mix5 verbal, 5 numerical, 5 abstract and 5 mechanical, interleaved so that no two consecutive scored items share a format, with the rotation order redrawn every round
Item generationEvery item is built at run time from a seed: the analogies draw from five relation families, the numerical items from percentage change, ratio splits and unit rates, the abstract items from multiplicative, alternating and letter series, and the mechanical items from gear trains, gear ratios, levers and pulleys
DistractorsThree per item, and each one is the right arithmetic on the wrong quantity — dividing the increase by the finishing figure rather than the starting one, reading the gear ratio the wrong way round. Where two traps collide on the same value, a fourth option is generated so the chance level stays at 25%
Mechanical itemsWorded rather than drawn, which is a genuine departure from the published format. It isolates the physical reasoning from the diagram reading and keeps the strand usable with a screen reader; the drawn version is a separate page on this site
Routing ruleLowest accuracy wins, with the slowest median as the tie-break, and a format the clock never reached is excluded rather than counted as a weakness
What is discardedAn item during which this tab went to the background, and the item on screen when the ten minutes ran out. Both are counted and named in the result rather than folded into an average
What leaves the pageNothing. Items are generated in the tab, answers stay there, and the copy button writes the four-format breakdown to your own clipboard

Frequently asked questions

Why does the format change on almost every question?

Because settling into one is exactly what a real battery denies you, and because a set grouped by format would measure something different. Given five numerical items in a row, most people are quicker on the fifth than on the first — they have the arithmetic loaded and the format in their hands — so a grouped set flatters whichever format happened to come last. Interleaving costs everybody the same switching penalty on every item, which is what makes the four medians comparable with each other. It is also closer to what a graduate battery feels like, where sections arrive back to back with no warning about what is in the next one.

Five items per format is tiny. What can that possibly establish?

Nothing about how good you are, and something quite useful about where your time goes. Accuracy over five items is extremely unstable — one item either way is a 20-point swing, and a different set of five would routinely produce a different order. Median seconds per item is far steadier, because it is a property of how you work rather than of which five questions turned up, and it is the quantity that decides whether you reach the end of a real paper. That is why the routing weighs both and why the panel tells you to read the table across rather than down.

Why are the mechanical questions written out instead of drawn?

So that the strand measures the physics rather than the picture, and so that it works for anybody reading with a screen reader. A drawn pulley system asks two things at once: can you read the diagram, and do you know that the load is shared between the supporting rope sections. Separating them is genuinely informative — plenty of people who find the drawn version hard have no trouble at all once the arrangement is described in a sentence. It is a real departure from what an employer will send you, and the pictorial version lives on the mechanical page next door for exactly that reason.

The ten minutes beat me. What does that actually tell me?

That your pace is under two items a minute on this mix, which is a fact about pace and not about ability. Most published aptitude batteries are built so that finishing is difficult, because a test everybody completes cannot separate the people taking it — so running out of time is the designed outcome rather than the failure mode. What is worth looking at is where the minutes went: if one format's median is double another's, that difference is where the whole set was decided, and it is a much more actionable thing to know than a total.

Should I guess when I have no idea?

On this drill, yes, and on most employer batteries, yes — but check, because the exception is real. Nothing here subtracts for a wrong answer, so a guess among four options is worth a quarter of a mark and a blank is worth nothing. A small number of assessments do apply a correction for guessing, and they say so in the instructions; where they do, a guess among four is worth nothing on average and the calculation changes. The instruction screen is the only place that answers this, and it is worth the thirty seconds of reading it takes.

Can the routing send me to a format I am already good at?

It can, and the commonest way is a format the clock never reached, which is why those are excluded from the decision entirely rather than scored as zero. The other way is a run where all four came out level and the tie-break picked on a median separated by two seconds — a difference that means nothing. The table under the recommendation exists so you can overrule it: all four formats are linked, with their own numbers beside them, and if you already know which one you dread then that is better evidence than twenty items.

Does a mixed set work on a phone?

Yes, and the mechanical strand is the reason it works better here than a pictorial set would. Every option is a button, the items are text so nothing has to be pinched to read, and the answers you give by tapping are counted separately and reported. What a phone does cost you is a little time per item, because aiming a thumb is slower than moving a finger already resting on a number key — enough to matter if you compare a phone run against a keyboard one, and not enough to change which of the four formats is your slowest.

What an aptitude battery is measuring, and the speed trap inside it

“Aptitude test” is an umbrella rather than an instrument. What arrives in the email is a battery — several short sub-tests, each aimed at a different kind of reasoning, scored separately and then combined by a rule the employer chose. Employers build batteries rather than one long test for a reason worth knowing: two sub-tests aimed at genuinely different things add more to a prediction than twice as many items of one kind would, because the second one covers ground the first cannot reach. That is also why the mix varies by role. A finance graduate scheme leans numerical, an engineering one leans mechanical and spatial, and a general management one leans verbal — so the first question to ask about a battery is not how to get better at it, but which of its parts is carrying the weight for the job you applied for.

The trap inside every one of them is the distinction between a speeded test and a power test. A speeded test is made of items almost everybody could solve, with more of them than anybody can finish — it measures rate. A power test allows enough time and lets the items get progressively harder — it measures ceiling. Most commercial batteries are deliberately somewhere in between and do not announce which way they lean, and candidates almost universally prepare as though they were sitting a power test: practising harder items, when the paper in front of them is going to be decided by how many easy ones they reach. The pace marker in the drill above exists to make that visible while it is happening rather than afterwards. If the format under the clock is drawn rather than written, the mechanical page and the non-verbal set are where the pictures are; if it is words, the verbal measure goes deeper than five analogies can.

One thing this page will not tell you is what a good result looks like, and the omission is deliberate rather than coy. Employers set their own thresholds against their own comparison groups, those groups differ by role at the same employer, and the number is not published by anybody — so any figure quoted here for what is “good” would be invented, and inventing it is the one thing on this site that could cost somebody an application. What preparation legitimately buys is narrower and still worth an evening: on a paper where reaching the end is difficult, the seconds you do not spend working out what an unfamiliar question type wants are seconds spent answering. Two of the best-documented timed formats are on this site at their real tempo — fifty items in twelve minutes and fifty in fifteen — and the map of which family your process is actually using is on the assessment guide.

A reaction time here is the interval between the frame that painted the stimulus and the timestamp the browser attached to your key, both read from the same monotonic clock. What neither can see is the display pipeline behind it, so on a 60 Hz screen roughly 16 ms of every figure below is the machine rather than you. That is the timing floor: two numbers closer together than that are the same number, and this page reports no precision it cannot support.

This is a measurement exercise, not a clinical assessment. It reports what you did on this page against a stated reference and nothing more — it cannot establish a learning difficulty, dyscalculia, or a reason to ask for extra time. Only a qualified professional, working with more than a browser, can make that judgment.

Where twenty generated items live

Every number on this page is worked out by JavaScript running in the tab you are reading it in. Your answers, your reaction times and your score are never uploaded, logged or kept — which is also why the test carries on working after you disconnect from the network, and why nothing here can be held back behind an email address.

The items themselves are built in this tab from a seed rather than fetched, which is why the drill keeps working with the network off and why no server ever learns which questions you were asked, let alone how you answered them. Reloading loses a finished run entirely; the copy button is the way to keep one.