Intelligence supports learning, reasoning and problem solving. Wisdom applies knowledge and reasoning to uncertain life problems involving values, other people and long term consequences. Research finds only a small positive relationship, so high IQ helps but never guarantees wise judgment.
Intelligence helps solve problems. Wisdom also asks which problem, values and consequences should guide the solution.
0 The Short Answer
Intelligence and wisdom are distinct constructs, and the distinction is measured rather than merely asserted. Psychometric intelligence has had a working instrument for over a century. Wisdom spent most of that century as a subject for proverbs, until a research group at the Max Planck Institute for Human Development in Berlin decided to treat it as something you could put in front of a participant, record, transcribe and score. That decision is what makes this comparison worth writing about, because it replaced opinion with numbers that can disagree with us.
Here is the headline those numbers produce. In the meta-analysis by Mengxi Dong, Nic Weststrate and Marc Fournier, published online in Perspectives on Psychological Science in 2022 and printed in the journal's 2023 volume, the pooled correlation between wisdom and intelligence across 17 samples and 56 effect sizes was r = .115, with a confidence interval running from .061 to .170. That is roughly one percent of shared variance. Knowing someone's IQ tells you almost nothing about how they will score on a wisdom instrument, and the reverse holds too.
The interesting part is not the small number itself but what it splits into. The association climbs to r = .21 for crystallized ability and falls to r = .07 for fluid reasoning. It reaches r = .22 for the Berlin performance protocols and drops to r = .05 for Igor Grossmann's wise reasoning interviews. It sits at essentially zero for wisdom measured by self-report. A single correlation between two words hides all of that. This page unpacks it, in the order the research actually developed.
Where ACIS stands
ACIS is an online cognitive self-assessment covering six CHC domains. It does not measure wisdom, and neither does any other cognitive battery, because wisdom research relies on rated protocols, diaries and interviews rather than keyed items. Nothing on this page should be read as a claim that a test score describes a person's judgment.
1 What an IQ Test Is Built to Measure
Before comparing intelligence with anything, it helps to be exact about what a battery samples. A structured cognitive test presents a series of tasks with defensible correct answers, administered under controlled instructions and usually under time limits, and scores the results against a norm sample so that any individual can be placed on a distribution with a mean of 100 and a standard deviation of 15. A score of 130 sits at roughly the top two percent, about one person in 44. The tasks cluster into recognized domains: verbal comprehension, fluid reasoning, visual spatial processing, working memory, processing speed and quantitative reasoning. Our page on what an IQ test actually measures walks through each of them. Screening instruments belong to a different family altogether, a distinction that does real work in the discussion of Donald Trump's IQ, where the cognitive test he cites is the Montreal Cognitive Assessment, a brief dementia screen scored out of 30 with a low ceiling and no IQ norms behind it.
Every design choice in that paragraph is a constraint. A keyed answer is required, because without one there is no way to score reliably. Standard conditions are required, because a score is only interpretable if everyone faced the same task. Time limits are common, because processing efficiency is part of what the instrument samples. Items are stripped of personal stakes on purpose, so that a participant's family history or religion or grief does not contaminate the measurement.
Those constraints are exactly what excludes wisdom. Real life dilemmas have no answer key. They arrive with stakes attached, at unpredictable moments, with incomplete information and with values in conflict. A test cannot score how you handle a decision whose right answer is contested without smuggling in the scorer's own values, and psychometricians have generally preferred to leave that territory alone rather than pretend to objectivity they do not have.
This is not a weakness disguised as modesty. A battery does one job well: it estimates how efficiently a person handles well-posed problems across separate abilities, and it reports the result as a profile with confidence intervals rather than a single verdict. The distinction between the fast reasoning it samples and the accumulated knowledge it also samples is covered in fluid versus crystallized intelligence, and that distinction turns out to matter for the wisdom question later on this page.
2 The Berlin Paradigm: Wisdom as an Expert Knowledge System
The most serious early attempt to operationalize wisdom came from Paul Baltes and Ursula Staudinger, whose summary of the program appeared in American Psychologist in 2000. Their move was to define wisdom narrowly enough to measure: an expert knowledge system concerning what they called the fundamental pragmatics of life, meaning knowledge and judgment about the meaning and conduct of a life, oriented toward both personal and collective well being.
Defining it as expertise had a practical consequence. Expertise can be probed the way psychologists probe expertise in chess or medicine, by handing someone a problem and listening to how they work through it. So the Berlin group built dilemmas covering three areas: life planning, life management and life review. Participants think aloud while working through them, and the resulting transcript, not a score sheet, is the raw data.
Trained raters then evaluate each protocol against a family of five criteria, which Baltes and Staudinger grouped into two basic and three metacriteria:
Rich factual knowledge. How much the person actually knows about human development, relationships, institutions and the ordinary machinery of life.
Rich procedural knowledge. Strategies for giving advice, weighing options, handling conflict and deciding when to act rather than wait.
Lifespan contextualism. Placing the dilemma inside the many contexts a life runs through, family, work, health, culture, and across time rather than at a single moment.
Relativism of values and life priorities. Recognizing that people legitimately want different things, without collapsing into the position that nothing matters.
Recognition and management of uncertainty. Acknowledging the limits of what can be known about a future life and still producing a usable course of action.
Read that list next to a test item and the gap becomes obvious. Four of the five criteria reward the participant for complicating the question. A reasoning item rewards the participant for closing it.
3 What the Berlin Program Found, and Where It Ran Out
Three findings from that program deserve to survive into any modern discussion. The first is that high wisdom related performance is genuinely rare. Protocols that satisfy all five criteria at a high level are uncommon in ordinary samples, which is precisely what you would expect from something defined as expert level knowledge rather than as a pleasant personality trait. Popular writing makes the mirror image error about ability, reading everyday traits such as being a night owl or keeping a messy desk as evidence of intelligence when those particular markers are weakly supported or simply myths, and the ones that do hold up are weak correlations rather than verdicts.
The second is the one that ruined a comfortable assumption. Staudinger's 1999 review in the International Journal of Behavioral Development integrated the age results across the Berlin studies and concluded that between roughly 20 and 75 years of age, wisdom related knowledge and judgment shows essentially no relation to age. Not a weak positive relation. A flat line. Section 11 returns to why.
The third is a methodological warning that the Berlin group made themselves. The paradigm measures wisdom related knowledge and judgment as expressed in a protocol. It does not measure whether the person lives that way. Someone can articulate the tradeoffs in a stranger's marital crisis with textbook subtlety on Tuesday and handle their own divorce badly on Wednesday, and the instrument would never see the second event.
There is also a quieter problem, visible only when the paradigm is compared with everything that came after it. Of all the wisdom instruments in the Dong meta-analysis, the Berlin protocols show the strongest relationship with intelligence, r = .224. A skeptic reads that number as partly a verbal fluency effect: producing a long, articulate, well organized spoken argument is itself a task that rewards crystallized ability, whatever the content happens to be. That suspicion motivated much of the work described next.
4 Grossmann's Wise Reasoning: A Narrower and Harder Target
Igor Grossmann, working first at Michigan and later at Waterloo, took a different route. Rather than treating wisdom as a body of knowledge, he treated it as a way of reasoning that can be observed while it happens, and he narrowed the target to a short list of behaviors: intellectual humility, meaning explicit recognition of the limits of one's own knowledge; recognition that circumstances change and outcomes are uncertain; consideration of perspectives other than one's own; and the search for compromise or integration between conflicting positions.
The first major test appeared in PNAS in 2010, with Jinkyung Na, Michael Varnum, Denise Park, Shinobu Kitayama and Richard Nisbett. Participants from a representative community sample read stories about intergroup and interpersonal conflicts and said aloud how they thought each situation would unfold. Responses were coded for those higher order reasoning schemes, and the coding scheme itself was validated against judgments from professional counselors and wisdom researchers rather than being trusted on the authors' say so.
The result cut against the standard aging story. Older participants used those schemes more than young and middle aged participants did, and they did so despite the decline in fluid reasoning that the same cohorts show on cognitive tests. Whatever the schemes are tracking, it is not raw processing capacity, because processing capacity was moving in the opposite direction.
A follow up in the Journal of Experimental Psychology: General in 2013 pushed further. Wise reasoning scores were associated with higher life satisfaction, less negative affect, better social relationships, less depressive rumination and greater longevity, and those associations held with socioeconomic factors, verbal ability and several personality traits controlled. In the same data, intelligence showed no association with well being at all. Two constructs that correlate at r = .12 with each other behaved completely differently against the same outcome, which is about as clean a demonstration of discriminant validity as this literature offers.
5 The Solomon Effect: Wiser About Other People's Problems
The single most useful finding in modern wisdom research is also the least flattering. Grossmann and Ethan Kross published it in Psychological Science in 2014 under the name Solomon's paradox, after the king remembered for brilliant judgment in other people's disputes and disastrous management of his own household.
Across three experiments with 693 participants, people reasoned more wisely about another person's problem than about an equivalent problem of their own. They acknowledged the limits of their knowledge more readily, gave more weight to compromise, allowed more room for future change and considered other parties' viewpoints more often, as long as the situation belonged to someone else. Facing their own version of the same conflict, the same people got measurably narrower.
Then came the manipulation. Participants instructed to reflect on their own problem from a distanced perspective, describing it as an observer would rather than through their own eyes, closed the gap. Self-distancing eliminated the asymmetry. And in the third study, the whole pattern held equally for participants aged 20 to 40 and participants aged 60 to 80, so the advantage was not something older adults had already acquired.
Sit with the implications for measurement. Wise reasoning moved substantially in a single session because of an instruction about point of view. Nothing comparable exists on the cognitive side. No one's working memory span improves because the digits belong to a friend, and no phrasing of a matrix item makes a person's fluid reasoning jump within an hour. That contrast tells you the two constructs are not just weakly correlated, they are different kinds of thing.
6 Wisdom Behaves More Like a State Than a Trait
If a single instruction can shift wise reasoning, ordinary life should shift it too. Grossmann, Tanja Gerlach and Jaap Denissen tested that directly with a daily diary study published in Social Psychological and Personality Science in 2016, tracking how the same individuals reasoned about the actual challenges their days handed them.
Wise reasoning varied within individuals across situations, and the variation was not noise. People reasoned more wisely about social situations than nonsocial ones. The situation level scores predicted outcomes such as emotional complexity and forgiveness, while trait level wisdom scores, the kind a questionnaire would produce, showed fewer of those associations. In other words, when the same person is measured repeatedly, the fluctuation carries more information than the average.
Grossmann developed the general argument in Perspectives on Psychological Science in 2017 under the title Wisdom in Context, proposing that experiential, situational and cultural factors shape wise thinking more than the field had assumed, and that an ego decentering mindset is what recovers it when personal stakes are high. In 2020 a group of wisdom researchers including Grossmann, Weststrate, Monika Ardelt, Justin Brienza, Michel Ferrari, Howard Nusbaum and John Vervaeke published a common model in Psychological Inquiry, describing wisdom as the morally grounded application of metacognition to reasoning and problem solving. Metacognition is a process word, and processes vary by occasion.
Compare this with the cognitive side, where rank order stability across long intervals is one of the better established findings in differential psychology. A person who scores near the top of their age group tends to stay near the top decades later. Wisdom instruments show nothing like that stability, and the honest reading is that a large share of what they capture belongs to the moment rather than to the person. Any correlation between a stable measure and a fluctuating one is mathematically capped before the science even begins.
7 Sternberg's Balance Theory and the Self-Report Family
Robert Sternberg approached the problem from theory rather than protocol. His balance theory, published in Review of General Psychology in 1998, treats wisdom as the application of intelligence, creativity and knowledge toward a common good, achieved by balancing three classes of interest: intrapersonal, meaning your own; interpersonal, meaning those of the people around you; and extrapersonal, meaning institutions, communities and anything larger. Balancing happens over short and long time horizons, and through three responses to circumstance: adapting to an environment, shaping it, or leaving it for another. Theory led expansions of this kind have a mixed record once measurement catches up, as Howard Gardner's eight or nine multiple intelligences illustrate: when the proposed abilities are actually measured they correlate positively rather than standing apart, which is the signature of one general factor.
What that framework buys is a way of saying that skill deployed toward purely private benefit is not wisdom, no matter how technically impressive. It makes wisdom explicitly value laden, and it explains why the same cleverness can be described as wise in one application and predatory in another. What it costs is measurability. Nobody has produced a widely adopted scoring system for balance theory comparable to the Berlin criteria, and the common good component requires the scorer to hold a position on what the common good is.
The pragmatic response from other researchers was to build questionnaires. Ardelt's Three-Dimensional Wisdom Scale, validated in Research on Aging in 2003 with 180 older adults, uses 39 items across cognitive, reflective and affective dimensions. Jeffrey Webster's Self-Assessed Wisdom Scale and Michael Levenson's Adult Self-Transcendence Inventory take similar routes with different content.
The weakness of the whole family is structural rather than incidental. Asking a person to rate their own intellectual humility puts the trait and the instrument in direct conflict, since the people best calibrated about their limits are the ones most likely to mark themselves down. Add ordinary social desirability pressure and you have a measure that can invert its own construct. The Dong meta-analysis captured the consequence numerically: wisdom measured by self-report, which the authors call phenomenological wisdom, correlated with intelligence at r = .034, a coefficient statistically indistinguishable from zero.
8 Do the Wisdom Measures Even Agree With Each Other?
This is the question that decides how much weight any of the previous sections can bear, and the answer is uncomfortable. Judith Glück and colleagues put five instruments in front of the same participants and reported the results in Frontiers in Psychology in 2013, using a sample of 47 people nominated by others as wise plus 123 controls, a design chosen so that real variation in wisdom would be present rather than assumed. Cognitive batteries are held to the opposite standard, where agreement between instruments is part of the evidence deciding whether an IQ test is accurate rather than merely plausible.
The Berlin protocols correlated with the Self-Assessed Wisdom Scale at r = .23, with the Three-Dimensional Wisdom Scale at r = .25 and with the Adult Self-Transcendence Inventory at r = .30. The questionnaires agreed better among themselves, .58 between the transcendence inventory and the three dimensional scale and .50 between the transcendence inventory and the self-assessed scale, but even there the three dimensional scale and the self-assessed scale managed only .26.
One detail from that study deserves to be better known than it is. Against the same inductive reasoning task, the Three-Dimensional Wisdom Scale correlated at r = .22 while the Self-Assessed Wisdom Scale correlated at r = negative .15. Two published questionnaires that both claim to measure wisdom relate to the same reasoning test with opposite signs. Whatever they are measuring, it is not one thing.
Rated protocols
Berlin and wise reasoning interviews. Participants think aloud about a dilemma and trained coders score the transcript. Expensive, hardest to fake, most contaminated by verbal fluency.
Self-report scales
3D-WS, SAWS, ASTI. Cheap, usable in large surveys, and structurally compromised on any trait that a well calibrated person would understate.
Situational measures
Diaries and manipulations of psychological distance. They capture within-person variation the other two families average away, but they do not produce a single tidy score.
9 The Empirical Relationship With IQ, Broken Apart
With the measurement landscape in view, the meta-analytic numbers become interpretable rather than merely small. Dong, Weststrate and Fournier synthesized thirty years of wisdom correlates, and the intelligence portion drew on 17 samples contributing 56 effect sizes. Here is the breakdown that the single pooled coefficient conceals.
Overall r = .115
Confidence interval .061 to .170. Positive, statistically reliable and small enough that it constrains almost nothing about an individual.
Performative wisdom r = .148
Wisdom scored from rated protocols. The relationship survives, modestly, when someone other than the participant does the scoring.
Phenomenological wisdom r = .034
Wisdom measured by self-report. The interval crosses zero, so the honest reading is no reliable relationship at all.
Crystallized ability r = .213
Acquired knowledge and verbal comprehension carry nearly all of the association between the two constructs.
Fluid reasoning r = .065
Novel problem solving under time pressure has close to no bearing on how wisely a person reasons about a life dilemma.
Berlin r = .224, wise reasoning r = .046
Two performance based instruments, both scored by trained raters, differ by a factor of nearly five in how much they track measured intelligence.
Translate the top line into something usable. An r of .115 corresponds to roughly 1.3 percent of shared variance. If you sorted a room by IQ and then by wisdom score, the two orderings would look close to unrelated. Predictions in that range are legitimate for describing populations and useless for describing the person in front of you, which is the same caution that applies to IQ and job performance, except that there the correlations are several times larger. Personality behaves the same way once it is measured as traits instead of sorted into types, since the trait level evidence on personality and measured ability leaves Openness to Experience as the only dimension with a real relationship, and the widest defensible gap between Myers-Briggs letter groups is about six points against a within group spread of fifteen.
10 Why the Correlation Stays Small
Four mechanisms account for most of the gap, and separating them prevents the lazy conclusion that intelligence is somehow irrelevant to good judgment.
Shared method rather than shared substance. The pattern where crystallized ability tracks wisdom at .21 while fluid reasoning manages .07 is hard to explain if the two constructs overlap conceptually, and easy to explain if part of the overlap is verbal. Both a vocabulary subtest and a think-aloud protocol reward a person who produces organized, articulate language. Some of the correlation is the measurement channel, not the thing being measured.
Opposite demands on certainty. A test item pays for arriving at the answer, quickly and confidently. Two of the five Berlin criteria pay for the opposite, for holding value conflicts open and for admitting the limits of what can be known. A person can be excellent at both, but excellence at one gives no head start at the other, and habitual fast closure may actively interfere.
Unstable measurement on one side. The diary and self-distancing results say that a large share of wise reasoning belongs to the situation. A correlation between a stable variable and a variable dominated by occasion specific variance is attenuated before any theory enters, purely by the arithmetic of reliability.
Restricted range and contested criteria. Wisdom samples often skew educated and volunteer heavy, compressing variation on both sides. And where the instruments disagree with one another as badly as section 8 documented, no correlation with a third variable can be cleaner than the instruments themselves.
None of that supports the popular inversion, the idea that intelligent people are somehow worse at judgment. There is no meta-analytic evidence for a negative relationship. What the data support is independence, which is a stranger and more useful finding than either flattering story. If you want the neighboring comparison, critical thinking versus intelligence covers a construct that sits much closer to the test.
11 Age: The Stereotype and What the Data Show
Nearly every culture packages wisdom with grey hair, so the evidence here surprises people more than any other section. Staudinger's 1999 integration found a flat relationship between age and Berlin wisdom scores across the adult range. The Dong meta-analysis put the pooled age correlation at r = .04 across 98 samples, with a confidence interval of .01 to .07. Within older adult samples specifically, the direction reverses to r = negative .11. Age is close to worthless as a predictor of wisdom scores, and among the old it may be slightly worse than worthless. Age sits awkwardly on the cognitive side too, which is why the original ratio formula built on mental age divided by chronological age gave way to deviation scores against same age peers, a change that cannot be run backwards to recover an adult's mental age.
Yet Grossmann's 2010 PNAS study did find older participants reasoning better about social conflicts. Both results are real, and reconciling them is the point of this section. The conflicts in that study belonged to other people: intergroup disputes and interpersonal problems the participant merely evaluated. When Grossmann and Kross put participants in front of their own relationship conflicts four years later, the age advantage vanished, with 60 to 80 year olds showing the same self-other asymmetry as 20 to 40 year olds.
So the defensible summary is narrow and specific. Decades of experience appear to help when you are the observer of someone else's mess. They do not appear to help when the mess is yours, which is unfortunately when wisdom matters most and when almost all real decisions occur.
There is a second reason the stereotype survives. Older adults are more often placed in advisory roles, so we encounter them while they are performing exactly the third-party reasoning where age effects show up. Selection effects do the rest, since the people we call wise elders are chosen for that reputation rather than sampled at random. For how measured cognition itself moves across the lifespan, which is a separate question with a different answer, see average IQ by age.
12 What a Cognitive Battery Can and Cannot Tell You
Everything above converges on a practical boundary, and being explicit about it is more useful than another round of definitions. A parallel gap exists for rational thinking, which correlates with measured ability and still comes apart from it in reproducible ways. The most compact instrument in that literature is the three item test of whether a first answer gets audited, where the arithmetic never goes past middle school and only 17 percent of Frederick's 3,428 respondents got all three right.
A well built battery can estimate how efficiently someone handles well-posed problems across separate domains, express that estimate as a banded profile instead of one flat figure, show which abilities are relatively stronger within a person, and predict outcomes such as training success and academic attainment with accuracy that is well documented and far from complete. Those are real capabilities, and they are the reason the instruments have lasted.
A battery cannot observe how a person deliberates when the stakes are personal, score whether their priorities serve anything beyond themselves, detect whether they seek out perspectives that contradict their own, measure whether they can hold a conclusion loosely, or watch what they actually do across the weeks that follow. Wisdom research reaches those things with transcripts, trained raters and repeated daily sampling, at a cost per participant that no automated test could carry. The limitation is architectural, not a matter of the tests being insufficiently clever yet.
The useful framing is that the two traditions ask different questions. A cognitive test asks how well this person handles problems that have answers. Wisdom research asks how this person reasons when the problem has no answer and the stakes are theirs. A high score on the first is an asset for the second, since knowledge and reasoning are inputs to any good decision, but it is an input and not a substitute. If you want the adjacent comparison with feeling and social skill rather than judgment, emotional intelligence versus IQ covers that ground.
Measuring the part that can be measured is still worth doing, as long as you know which part it is. A domain profile tells you where your reasoning, memory and speed sit relative to a norm sample, which is genuine information about one specific thing.
What is the difference between intelligence and wisdom?
Intelligence is capacity for solving problems that have determinable answers. Wisdom, as psychology operationalizes it, is a way of handling problems that do not, involving humility about one's knowledge, tolerance of uncertainty and attention to competing interests.
How strongly are the two correlated in the research?
Weakly. The pooled estimate from Dong, Weststrate and Fournier is .115, which leaves the overwhelming majority of variation in each unexplained by the other.
What is the Berlin Wisdom Paradigm?
A method developed at the Max Planck Institute in Berlin that treats wisdom as expertise about life's fundamental pragmatics and assesses it from spoken responses to dilemmas, scored by trained raters.
What are the five Berlin criteria?
Factual knowledge, procedural knowledge, lifespan contextualism, relativism of values and priorities, and recognition plus management of uncertainty. The first two are basic criteria; the last three distinguish wisdom from mere expertise.
How does a rater score a think-aloud protocol?
Transcripts are rated criterion by criterion by coders trained to a shared standard, typically with multiple raters per protocol so agreement can be checked. It is slow, human work rather than automated marking.
Which wisdom measure correlates most with IQ?
The Berlin protocols, at .224 in the meta-analytic breakdown. That is one reason some researchers suspect the format rewards verbal skill alongside the intended construct.
What does Grossmann mean by wise reasoning?
A cluster of observable habits: admitting the limits of your knowledge, expecting circumstances to change, taking other viewpoints seriously and looking for workable compromise instead of victory.
What is the Solomon effect?
The finding that the same person reasons more wisely about someone else's dilemma than about their own equivalent dilemma, named after the biblical king celebrated for judging others and criticized for his own household.
Does self-distancing improve reasoning?
In the 2014 experiments it did. Participants asked to view their own situation as an observer would reasoned about as well as they did on other people's problems, closing the gap entirely.
Is wisdom a stable trait?
Less stable than the questionnaires imply. Diary data show substantial movement within the same person from one day and one situation to the next, which is unusual for something treated as a personality characteristic.
What is Sternberg's balance theory of wisdom?
A model in which wisdom means directing intelligence, creativity and knowledge toward a common good by balancing your own interests, other people's interests and larger institutional interests over time.
Are self-report wisdom scales trustworthy?
They are convenient and internally consistent, but they inherit a design problem: the trait includes accurate self-appraisal, so the most accurate self-appraisers are penalized by their own honesty.
Do the main wisdom measures agree with each other?
Poorly. In Glück's five instrument comparison, the performance measure agreed with each questionnaire only in the .23 to .30 range, low enough that the label is doing more work than the constructs.
Why does crystallized ability track wisdom more than fluid reasoning?
Both crystallized tasks and wisdom protocols reward accumulated knowledge and fluent expression. Speeded novel problem solving shares almost nothing with a transcript about a life dilemma.
Does wisdom grow with age?
Not automatically. Across studies the age correlation is about .04, and inside older samples it turns slightly negative. Experience creates the opportunity for growth without supplying it.
Why is the "older and wiser" belief so persistent?
Partly because age effects do appear for third-party dilemmas, which is the setting where we usually encounter elders giving counsel, and partly because reputational selection puts unusually thoughtful older people in front of us.
Does wise reasoning predict well-being?
In the 2013 Grossmann study it predicted life satisfaction, relationship quality, lower rumination and longevity, holding income, verbal ability and personality constant. Measured intelligence predicted none of those outcomes in the same sample.
Can wisdom be taught or trained?
Short term shifts are demonstrable, since a distancing instruction changes reasoning within a single session. Whether repeated practice produces durable change is still open, and the honest answer is that nobody has run the long study yet.
Is emotional intelligence the same as wisdom?
No. Emotional intelligence concerns perceiving and regulating emotion, yours and other people's. Wisdom models add value conflict, uncertainty and the interests of parties outside the room.
Can an IQ test measure judgment or values?
Not by design. Scoring requires a defensible key, and questions about what a person should value have no key that a test publisher could defend without imposing their own position.
How should I read a correlation of .12?
As a statement about populations, not people. It means the two rankings barely resemble each other, and that predicting one person's wisdom from their test score would be closer to guessing than to measuring.
14 Sources Behind This Page
The claims on this page follow the published literature rather than our own assertions. These are the primary papers and reference bodies worth reading directly, with what each contributes.
Nisbett, R.E. et al. (2012). Intelligence: new findings and theoretical developments. American Psychologist, 67(2). The broad APA review of what moves measured intelligence and what does not.
Harada, C.N., Natelson Love, M.C. & Triebel, K. (2013). Normal cognitive aging. Clinics in Geriatric Medicine, 29(4), 737-752. Open access. Fluid abilities peak in the third decade while vocabulary and knowledge hold or improve into the seventies.
Salthouse, T.A. (2011). Consequences of age-related cognitive declines. Annual Review of Psychology. Open access. Laboratory declines are real yet everyday functioning largely holds, a caution against reading group curves as personal verdicts.
Plomin, R. & Deary, I.J. (2015). Genetics and intelligence differences: five special findings. Molecular Psychiatry. Open access. The standard modern review of what twin and DNA evidence does and does not show.
Gottfredson, L.S. (1997). Mainstream Science on Intelligence, the editorial signed by 52 researchers. Intelligence, 24(1), hosted by the University of Delaware. A consensus statement on what IQ tests measure.
Spearman, C. (1904). General intelligence, objectively determined and measured. American Journal of Psychology, full text at Classics in the History of Psychology. The paper where the g factor entered psychology.
Voncken, L., Albers, C.J. & Timmerman, M.E. (2019). Improving confidence intervals for normed test scores. Behavior Research Methods. Open access. Documents the mean 100, SD 15 metric and the uncertainty that norming from samples adds to any score.
Pearson Clinical Assessment Scientific Council (2023). Standardized Clinical Assessment for Practitioners: A Primer. How standard scores, percentile ranks and the standard error of measurement are meant to be read together.
Pearson (2008). WAIS-IV Score Report sample. What a real report contains: every composite paired with a percentile rank, a 95% confidence interval and a qualitative description, never a bare number.
ACIS measures six CHC domains across 20 subtests and reports each one with its own normed score and confidence interval, so you can see where you are strong and where you are not.