Are online tests accurate?
Are online tests accurate? Some are, within limits. An online timing test is accurate enough to show how your own results change and roughly where you stand beside a study, but not to give you an exact figure to the millisecond, because a browser cannot see the delay of your screen, your mouse or your keyboard. An online questionnaire adds up your answers exactly, but what the total means is only as good as the questions, your mood on the day and the group it is compared with. This page says where the limits lie, so that you can use a result for what it is.
Are online reaction time tests accurate?
A test that runs in a browser can only be as exact as the browser. What a good test does is take the time of your press from the moment the browser recorded it and not from the moment the page got round to reading it, and take the start of the signal from the moment the picture was drawn on the screen and not from the moment the page asked for it. The gap between the two stamps is your reaction. Done like this, the clock itself is not the weak point.
What the clock cannot see is the rest of the chain. Your screen takes a moment to light up, your mouse or touch screen takes a moment to report your press, and your computer takes a moment to pass it on. Together these add a delay that is hidden from the page and that differs from device to device.
The study that this site compares reaction times with measured the delay of its own set-up and reports a second, lower figure with that delay taken out[1]. This site cannot measure your device, so it allows for the delays that robot tests of everyday devices found: 44.1 ms to 115.1 ms more than the study's equipment with a keyboard or mouse, and 39.8 ms to 52 ms with a touch screen[2][3]. No study pressed a mouse, so a mouse gets the keyboard's figures. That is why a result is shown beside the study as a range and not as one exact rank, and why, on a computer, where the range is wide, it is only a rough guide.
What your screen does
A screen shows a new picture only so many times a second. A 60 Hz screen draws a new picture sixty times a second, and a 144 Hz screen more than twice as often, so on the slower screen a signal can appear a little later than the moment it was planned, and the next picture comes later still. A test can notice this, and the tests here do: if the screen stalls while a signal is showing, that try does not count and you do it again.
The refresh rate test shows how often your browser really gets a new picture, which is the first thing to know about a timing result from your own device.
What your mouse, keyboard and touch screen do
- A mouse reports its position and its clicks a certain number of times a second. A 1000 Hz mouse reports far more often than a 125 Hz one, and a wireless or Bluetooth mouse may add a delay of its own. The polling rate test shows what yours does.
- A keyboard waits a very short time after a press to be sure it was a real press and not a flicker of the switch. That wait is part of every key press, and it differs between keyboards. The keyboard test shows whether each key registers once.
- A touch screen has to notice your finger, decide that it is a tap and pass it on, which takes a moment. For that reason a result from a phone is a different thing from a result from a mouse, and we say less about a result from a phone than about one from a mouse.
How the tests here keep a result fair
- The first try is for practice and does not count.
- A press that comes before the signal, or so soon after it that it must have been a guess, does not count, and you do that try again.
- If you switch to another tab or app during a try, that try does not count.
- Your result is the middle of several tries, so one lucky or unlucky try does not decide it.
- A run that looks as if a machine made it is kept but not compared with anyone.
The methodology page sets out the same rules and what is still unverified.
Are online personality tests accurate?
A questionnaire is scored by arithmetic, so the sum is exact. What is open to question is what the sum means. The answers are how you see yourself on that day, and they depend on how the statements are worded, whether you understand them as the authors meant, and how you feel. That is why a result moves a little when you take a test again, and why a longer questionnaire is steadier than a short one.
The Big Five test on this site uses the IPIP-NEO-120, a public-domain questionnaire of the psychologist John A. Johnson[4], with many more statements than the quick tests. A result says where you sit between the two ends of each trait. Where a questionnaire has no published group for the language you use, the result says so and does not compare you with other people.
Tests that sort people into a few types are rougher still, because a person near the middle of a trait can land in either box on another day. A questionnaire is a mirror to think about yourself with. It is not a verdict, and it is no basis for choosing a person for a job.
How to get the most accurate result you can
- Use the same device, and the same mouse or keyboard, every time you compare results.
- Close other tabs and programs, so that the page is not interrupted in the middle of a try.
- Take timing tests when you are rested, and at about the same time of day.
- Take a timing test several times and look at the middle of your results, not the best one.
- Read a questionnaire's statements slowly, answer for the person you usually are, and take it again later to see what stays the same.
Questions
Are online tests accurate?
Accurate enough to show your own trend and a rough place beside a study, but not exact. A browser cannot see the delay of your screen and your mouse or keyboard, and a questionnaire is only as good as its questions and the way you answer them.
Why is my online reaction time slower than in a laboratory?
Your screen, your mouse or touch screen and your browser each add a delay that a web page cannot see. Tiredness and attention add to it as well.
Does my device change my result?
Yes. A phone, a laptop with a trackpad and a computer with a mouse each add their own delay, so compare your results only with results from the same device.
Are online personality tests accurate?
They add up your answers exactly, but they describe how you see yourself on the day, so a result is a starting point for thinking about yourself and not a verdict.
Sources
- Woods, D. L., Wyma, J. M., Yund, E. W., Herron, T. J., & Reed, B. (2015). Factors influencing the latency of simple reaction time. Frontiers in Human Neuroscience, 9, 131.doi:10.3389/fnhum.2015.00131
n = 1,469; Community volunteers in Rotorua, New Zealand, aged 18 to 65 (mean age 45.8, 40% men) - Pronk, T., Wiers, R. W., Molenkamp, B., & Murre, J. (2020). Mental chronometry in the pocket? Timing accuracy of web applications on touchscreen and keyboard devices. Behavior Research Methods, 52(3), 1371–1382.doi:10.3758/s13428-019-01321-2
Four devices from 2015 and 2016: a MacBook Pro, an ASUS laptop, a Samsung Galaxy S7 and an iPhone 6S, pressed by a robotAbstract: "In controlled circumstances, as can be realized in a lab setting, very accurate stimulus timing and moderately accurate RT measurements could be achieved on both touchscreen and keyboard devices, though RTs were consistently overestimated." Table 4 "Descriptives of RT overestimations (in milliseconds) per device and browser" (OS, Web Browser, Minimum, Maximum, Mean, SD): "Android Chrome 46.0 103.5 69.8 7.4", "iOS Safari 48.3 96.3 57.6 6.5", "MacOS Safari 93.0 163.7 132.9 8.1", "Windows Chrome 64.7 70.6 68.5 1.7", "Windows Firefox 49.8 84.9 61.9 5.7".
(read on 2026-10-06: https://www.ebi.ac.uk/europepmc/webservices/rest/PMC7280355/fullTextXML) - Anwyl-Irvine, A., Dalmaijer, E. S., Hodges, N., & Evershed, J. K. (2021). Realistic precision and accuracy of online experiment platforms, web browsers, and devices. Behavior Research Methods, 53(4), 1407–1425.doi:10.3758/s13428-020-01501-5
Desktop and laptop computers with Windows 10 and macOS, pressed by a robot; and the equipment of 202,600 online participantsAbstract: "We then employed a robot actuator in realistic set-ups to measure response recording across the aforementioned platforms, and between different keyboard types (desktop and integrated laptop)." "We found that modern web platforms provide reasonable accuracy and precision for display duration and manual response time". Table 2 "RT delay is calculated as the difference between known and recorded RT." Browser means: "Chrome 78.81", "Edge 80.10", "Firefox 82.30", "Safari 76.50"; device means: "macOS-Desktop 85.35", "Windows-Desktop 76.24", "Windows-Laptop 73.65". Results: "We found that 77% of these devices were desktop or laptop computers, whereas only 20% were mobile devices" "Based on a sample of 202,600 participants."
(read on 2026-10-06: https://www.ebi.ac.uk/europepmc/webservices/rest/PMC8367876/fullTextXML) - Johnson, J. A. (2014). Measuring thirty facets of the Five Factor Model with a 120-item public domain inventory: Development of the IPIP-NEO-120. Journal of Research in Personality, 51, 78-89.doi:10.1016/j.jrp.2014.05.003