ChatGPT and IQ: Test Scores by How Often People Use AI
We asked 30,491 people, before they saw their score:
How often do you use ChatGPT or another AI assistant?
Score compared with people their age and country
Each dot is about 4 people (adults). The higher the dot, the better they scored compared with people their age and country. The line marks the group’s average.
MyIQtested data · Dataset 2026-09-29.1 ·
IQ test scores from 30,000+ people, grouped by how often they use ChatGPT or another AI assistant.
The picture above shows adults (12,137 of the 30,491 answers), adjusted for age and country, the same view Figure 1 opens on. Every figure below can show all ages.
The question
The question people were asked about using AI
How often do you use ChatGPT or another AI assistant?
- 1 Every day
- 2 A few times a week
- 3 Rarely
- 4 Never
Options appeared in this order, one tap to answer.
How it was asked
- 33 puzzles The full reasoning test, scored in the browser.
- Age group A single tap. It can be skipped.
- This question One question, chosen at random from those live that week. It can be skipped.
- Score revealed On the results page, after the answer.
The answer was given before the score was shown.
This is data from people who chose to take a free online test. It is not a controlled comparison, and it cannot show cause.
Fig 01 · Test score by AI use
Test scores by AI use. Points versus all respondents.
Each dot is a group’s average score as points above or below all respondents in the current filter, with its 95% confidence interval. With Adjust on, groups are compared within age band and country first.
All ages and all countries, adjusted for age and country: Every day −3.8, A few times a week +2.7, Rarely +0.6, Never −2.9 (points versus all respondents).
Fig 02 · Personality by AI use
Personality by AI use. Percentile points versus all respondents.
Big Five percentiles of 1,685 SeeMyPersonality takers of all ages who answered the same question. Each dot is a group’s average as points above or below all of them, with its 95% confidence interval, adjusted for age and country when Adjust is on.
Table view of this figure
| Answer | People | Openness | Conscientiousness | Extraversion | Agreeableness | Emotional sensitivity |
|---|---|---|---|---|---|---|
| Every day | 347 | −2.4 (−5.5 to 0.8) | −0.7 (−3.7 to 2.2) | −2.6 (−6.0 to 0.7) | −3.4 (−6.4 to −0.4) | −0.5 (−3.5 to 2.5) |
| A few times a week | 617 | −1.4 (−3.3 to 0.4) | −0.1 (−1.8 to 1.7) | +2.2 (0.4 to 4.0) | −0.5 (−2.2 to 1.3) | +0.2 (−1.6 to 1.9) |
| Rarely | 472 | −0.2 (−2.4 to 2.1) | +1.8 (−0.4 to 3.9) | +0.6 (−1.6 to 2.9) | +1.4 (−0.9 to 3.6) | −2.6 (−4.7 to −0.4) |
| Never | 249 | +6.1 (1.7 to 10.6) | −2.5 (−6.7 to 1.6) | −1.6 (−6.0 to 2.7) | −3.2 (−7.5 to 1.1) | +6.5 (2.5 to 10.5) |
Fig 03 · Each country and age group
The same comparison in each country and age group. One shared scale.
- Every day
- A few times a week
- Rarely
- Never
By country 200 or more answers, current age filter, most answers first
By age group current country filter
Figure 1 for every country with 200 or more answers and every age group, on one shared scale, each with its 95% confidence interval; grey rows have fewer than 30 people. Press a panel to filter the page to it.
Table view of this figure
| Group | People | Every day | A few times a week | Rarely | Never |
|---|---|---|---|---|---|
| 🇮🇩 Indonesia | 6,281 | −0.8 (−2.1 to 0.5) | +4.6 (3.9 to 5.4) | −0.2 (−0.7 to 0.3) | −4.8 (−5.9 to −3.7) |
| 🇯🇵 Japan | 1,533 | +0.4 (−0.9 to 1.7) | +1.8 (1.0 to 2.5) | −0.5 (−1.7 to 0.6) | −3.6 (−5.2 to −2.0) |
| 🇺🇸 United States | 1,100 | −8.2 (−12.2 to −4.3) | +1.3 (−0.2 to 2.7) | +2.1 (0.8 to 3.3) | +0.2 (−1.1 to 1.4) |
| 🇮🇳 India | 799 | −2.1 (−3.8 to −0.4) | +2.7 (1.6 to 3.9) | +1.7 (−0.1 to 3.5) | −4.7 (−8.2 to −1.2) |
| 🇬🇧 United Kingdom | 338 | −1.3 (−5.0 to 2.4) | +0.3 (−1.8 to 2.3) | +1.6 (0.0 to 3.3) | −1.7 (−3.5 to 0.1) |
| 🇫🇷 France | 300 | −1.0 (−5.7 to 3.7) | +0.8 (−1.2 to 2.9) | +1.7 (−0.2 to 3.6) | −3.9 (−7.9 to 0.0) |
| Under 13 | 6,290 | −8.1 (−9.8 to −6.4) | +1.7 (0.7 to 2.6) | +1.5 (1.0 to 2.0) | −0.9 (−1.8 to −0.1) |
| 13 to 15 | 6,807 | −4.0 (−5.3 to −2.7) | +2.2 (1.6 to 2.8) | +0.4 (−0.1 to 0.8) | −1.5 (−2.6 to −0.4) |
| 16 to 17 | 4,197 | −2.7 (−4.1 to −1.3) | +3.1 (2.4 to 3.8) | +0.3 (−0.4 to 1.0) | −5.2 (−6.8 to −3.7) |
| 18 to 24 | 5,773 | −0.7 (−1.9 to 0.6) | +3.4 (2.8 to 4.1) | −0.2 (−0.8 to 0.3) | −4.7 (−5.9 to −3.6) |
| 25 to 34 | 2,816 | −1.2 (−3.1 to 0.6) | +3.5 (2.5 to 4.4) | +0.2 (−0.6 to 1.0) | −4.2 (−5.7 to −2.8) |
| 35 to 44 | 1,510 | −4.9 (−7.3 to −2.5) | +3.9 (2.4 to 5.5) | +1.3 (0.2 to 2.3) | −3.0 (−4.5 to −1.5) |
| 45 to 54 | 985 | +1.0 (−1.0 to 3.0) | +1.0 (−0.4 to 2.5) | +0.1 (−1.2 to 1.4) | −2.0 (−3.5 to −0.5) |
| 55 to 64 | 520 | −0.1 (−3.4 to 3.2) | +2.1 (−0.1 to 4.3) | +2.2 (0.2 to 4.1) | −3.7 (−5.9 to −1.4) |
| 65 and over | 533 | −10.9 (−15.6 to −6.1) | +4.0 (0.6 to 7.5) | +3.9 (1.4 to 6.4) | −0.3 (−3.0 to 2.4) |
| Age not given | 1,060 | −5.8 (−9.0 to −2.7) | +4.9 (2.5 to 7.2) | +1.3 (−0.1 to 2.6) | −2.4 (−4.6 to −0.2) |
Data table
Figure 1 as a table. For the current filter.
| How often they use AI | People | Share | Gap, points | 95% interval | SD of scores | d vs the rest | Median percentile | Population covered | Differs at 95% (Holm) from |
|---|---|---|---|---|---|---|---|---|---|
| Every day | 2,006 | 16.5% | −1.60 | −2.44 to −0.76 | 20.6 | −0.11 | 52nd | 97% | every other group |
| A few times a week | 3,375 | 27.8% | +3.28 | 2.80 to 3.75 | 14.6 | 0.25 | 57th | 99% | every other group |
| Rarely | 4,401 | 36.3% | +0.36 | −0.03 to 0.75 | 17.5 | 0.02 | 49th | 100% | every other group |
| Never | 2,355 | 19.4% | −3.94 | −4.64 to −3.24 | 19.5 | −0.27 | 39th | 97% | every other group |
Gap and interval are in points versus all respondents; d is Cohen’s d against everyone else. The last column lists the groups each one differs from at 95% after correcting for the number of comparisons (Holm). No absolute score levels are published.
Methods
How the data was collected, and how every figure is computed.
01How the data was collected
Everyone took the same free online reasoning test on MyIQtested. After the last puzzle, and before the score appeared, the site showed one survey card. The order was: the 33 puzzles, the score calculated in the browser, an age card, one survey card, then the results page, where the score is first shown.
The question is chosen at random for each taker from the 4 questions live at the same time, among those their age group qualifies for. Each question stays live for 7 days, then the next question in the queue takes its place. The card can be skipped, and a row is only written when someone taps an answer or Skip: a card closed without a tap writes nothing.
This question was live from 22 Sep to 29 Sep 2026. It was shown 33,677 times and answered 30,491 times (90.5%).
02The test
The test has 33 multiple-choice puzzles, each with four options and one correct answer: 20 matrix reasoning puzzles (ICAR-based, from the International Cognitive Ability Resource framework), 6 series and logic, 4 verbal and 3 numerical. Every puzzle must be answered and there is no time limit; the median first attempt takes about 9 minutes.
The score is a fixed conversion of the number of correct answers: 51 + 3 × correct, with a floor of 55 and a ceiling of 150, so scores move in steps of 3. It is calculated in the taker’s browser and is not an empirical norm. That is why this page reports differences and ranks, never a level.
The verbal puzzles are re-authored for each language; the other 29 keep the English structure and answer key. The language versions are therefore not exactly the same instrument.
03Who answered
- Shown the question
- 33,677
- Answered
- 30,491 (90.5%)
- Skipped
- 3,186 (9.5%)
- Window
- 22 Sep to 29 Sep 2026
Age
Of the 30,491 answers, 57% came from people under 18 (including 21% under 13), 40% from adults and 3% from people who did not give an age. This question had no age limit, so under-13s were shown it.
Country
The largest groups were Indonesia (60%), United States (9%), Japan (8%), India (6%). 13 countries have their own column; the rest, 8% of answers, are pooled as Other countries.
Language
The page the question appeared on was in 12 languages; the largest were Indonesian 58%, English 30%, Japanese 8%, Russian 2%.
04Exclusions and small cells
- No row is filtered out when the figures are built: no bot, duplicate-visitor or score-range filter is applied. The same rows feed our internal report.
- Automated browsers and local development are kept out at the source: the survey is never sent from localhost, or from a browser reporting itself as automated. The answer service rejects unknown questions, answers that are not one of the options, and age-gated rows.
- Repeat takers are kept. 2,547 of 33,677 rows (7.6%) come from a visitor with an earlier row for this question, each a separate attempt. Removing them moves the group averages by between 0.3 and 0.8 points.
- People who skipped the question scored 3.8 points lower on average than people who answered it (not adjusted). Skipped cards are not in any figure except the answer-rate bars.
- Countries with fewer than 150 answers are pooled into Other countries, as are rows with no country. Age not given pools “prefer not to say”, a skipped age card and missing values.
- Cells with fewer than 5 people are pooled into Other countries before publication (138 cells, 339 people), so no cell can identify one person’s score. The aggregate CSV suppresses those cells instead.
- Any group or subgroup with fewer than 30 people is shown greyed as “n too small” and no interval is drawn.
05How the figures are computed
Every figure is relative. Nothing on this page states how high anyone or any group scored on the scale, only how groups differ from each other and where people rank among people of their own age and country.
Every range is a 95% confidence interval, and every difference we point out holds at 95% confidence after correcting for the number of comparisons (Holm method).
Strata and the standard population
A stratum is one age band within one country. Ten age bands and 14 country columns make the strata. With Adjust for age and country on, the standard population is the current filter’s own stratum mix: ws is the share of everyone in the filter who is in stratum s. With it off, everyone is pooled.
Gap versus all respondents
w̃s = ws / coverage(a), over the strata where answer a has at least one person
Adjust off: gap(a) = mean(a) − mean(everyone in the filter)
Coverage is the share of the standard population that the answer has people in; strata where nobody gave the answer are dropped and the rest re-weighted. Coverage is shown in the table when Adjust is on.
Intervals
pas = nas / ns; 95% interval = gap ± 1.96 × SE
The factor (1 − p) is there because the reference group, everyone, contains the answer. A variance estimated from fewer than 5 people is replaced by the variance of everyone in the filter. With Adjust off there is one stratum and this is the ordinary two-sample formula.
Effect size
Cohen’s d compares an answer with everyone else: d = (mean of the answer − mean of the rest) / pooled SD, using the same strata weights when Adjust is on. It is in units of the spread of scores, not in points.
Percentile within own age and country
Scores are grouped in 3-point bins and each bin is treated as a single score value. Within a stratum, everyone who shares a bin is tied and shares the range of ranks that bin covers, so the average of their percentiles is the mid-rank percentile. A person’s percentile is their rank among everyone who answered in the same age band and country, whatever they answered. Strata with fewer than 30 people are left out of the percentile figures.
The line in each hero column is the group’s average percentile, and the median in the table comes from the same distribution. With Adjust on, each answer’s distribution is the stratum-weighted average of its per-stratum distributions.
Personality figure
This question also ran on SeeMyPersonality, after its Big Five test: it was shown 1,847 times and answered 1,685 times. Each trait is the percentile (0 to 100) the person saw on their results, against SeeMyPersonality’s own empirical norms. The gap is computed as for Figure 1, on those percentiles, for all ages and countries. With Adjust on, the standard population is the mix of five age groups (under 18, 18 to 24, 25 to 34, 35 and over, not given) and 3 countries plus Other among everyone who answered. Only the gaps and their intervals are published, not the cells they come from, and a group of fewer than 30 people shows no numbers.
Answer rate
Answered divided by shown, by age band, with a Wilson 95% interval.
06Limitations
- The people are volunteers who chose to take a free online test. They are not a sample of any population, and country and age mixes are those of the site’s visitors.
- Each answer is a single self-report item with four options. People may read the options differently.
- The score comes from one unsupervised sitting of a 33-puzzle test. It is not a clinical measure, and its scale is a fixed conversion, not a norm.
- 57% of answers are from people under 18, including 21% under 13, who may read the question and the test differently from adults.
- Country is recorded by the site from the visit; it may not be the person’s nationality or where they grew up.
- People who skipped the question differ from those who answered, and the answer rate for age not given (48%) shows that the group who answered is not everyone shown.
- Differences are shown with intervals, but an interval covers sampling noise only, not the ways the sample was selected.
- The personality figure comes from a different group of volunteers, people who took the SeeMyPersonality Big Five test, and from 1,685 answers, so its intervals are wider.
07Versioning
- Dataset version
- 2026-09-29.1
- Generated
- 2026-09-30
- Responses
- 22 Sep to 29 Sep 2026
- Licence
- CC BY 4.0
Change log: 2026-09-29.1, first release. If the data or the method changes, the version changes and this line says what moved. Cite the version you used: myiqtested.com/research/chatgpt-use-and-iq.
