The TIPI measures the Big Five personality dimensions, Extraversion, Agreeableness, Conscientiousness, Emotional Stability, and Openness to Experience, with a pair of trait descriptors for each, for research settings where a full-length inventory is impractical.
Gosling, Rentfrow, and Swann designed it as a deliberate trade-off: each dimension is covered by one standard and one reverse-scored item, so internal consistency is modest by construction and the instrument's validity case rests on convergence with longer Big Five measures.
Independent validation work supports the five-factor structure with latent-variable methods, documents adequate test-retest stability, and reports generally positive results for German and Dutch adaptations.
It is free for any use without permission, but it is a research instrument for group-level analysis, not a tool for clinical assessment, counseling, or high-stakes individual decisions; note that the fifth dimension is keyed as Emotional Stability, not Neuroticism.
At a Glance
Items
10 (2 per Big Five dimension; each item pairs two trait descriptors)
Five domain scores, each the mean of two items: Extraversion, Agreeableness, Conscientiousness, Emotional Stability (higher = more stable), Openness to Experience
Validated populations
Adults; validated primarily in college and university samples
License
Free for any purpose; the authors state that no permission is needed
Original citation
Gosling, Rentfrow, & Swann (2003), Journal of Research in Personality
Introduction
The Ten-Item Personality Inventory (TIPI) is an ultra-brief self-report measure of the Big Five personality dimensions. Developed by Gosling, Rentfrow, and Swann (2003) for research settings where time is severely constrained, it covers each of the five dimensions with two items, each pairing two trait descriptors, and takes only a minute or two to complete. The instrument embodies a deliberate trade-off: with two items per dimension it cannot match the internal consistency of full-length inventories, but it retains documented convergent, discriminant, and test-retest properties that make it serviceable when a longer measure is impractical (Gosling et al., 2003).
Understanding Ultra-Brief Big Five Measurement
Comprehensive Big Five inventories can require up to 15-45 minutes, and multi-variable surveys, longitudinal panels, and online studies often cannot allocate that time to a single construct. Before instruments like the TIPI were validated, researchers in such settings faced a choice between omitting personality measurement altogether and improvising unvalidated shortcuts; the TIPI was constructed to provide a third option (Gosling et al., 2003). Each item presents a pair of related trait descriptors (for example, “extraverted, enthusiastic” for Extraversion), so that a single rating captures more of a dimension’s breadth than one adjective could, and each dimension is measured by one standard-keyed and one reverse-keyed item.
One labeling detail matters for scoring and interpretation: the TIPI keys its fifth dimension as Emotional Stability, not Neuroticism, so higher scores indicate greater emotional stability (Gosling et al., 2003; Ehrhart et al., 2009).
Theoretical Foundation
The TIPI operationalizes the Big Five (five-factor) model, which organizes personality traits into five broad dimensions identified in decades of lexical research: Extraversion, Agreeableness, Conscientiousness, Neuroticism (keyed on the TIPI as Emotional Stability), and Openness to Experience. Given only ten items, the measure targets the broad domains at the top of the trait hierarchy rather than attempting to assess narrower facets. The developers argued that low alpha coefficients are an expected arithmetic consequence of two-item scales, and that the appropriate evaluation standard for an ultra-brief measure is its convergence with longer instruments and its pattern of external correlates (Gosling et al., 2003).
⚡ Key insight: The TIPI trades internal consistency for brevity by design; its validity case rests not on alpha but on convergence with full-length Big Five measures, which is substantial at the scale level and near-complete once measurement error is modeled.
Key Features
Assessment Characteristics
10 items, two per Big Five dimension, each pairing two trait descriptors (Gosling et al., 2003)
7-point response scale from disagree strongly to agree strongly, with the stem “I see myself as:”
About 1-2 minutes to complete; the developers describe it as taking only a minute (Gosling et al., 2003)
One standard and one reverse-keyed item per dimension; the fifth dimension is keyed as Emotional Stability
Free for any purpose; the authors state that no permission is needed to use it
Rate how well ten pairs of personality traits describe you.
Scoring and Interpretation
Response Format
Each pair of traits is rated on a 7-point scale: 1 = disagree strongly, 2 = disagree moderately, 3 = disagree a little, 4 = neither agree nor disagree, 5 = agree a little, 6 = agree moderately, 7 = agree strongly (Gosling et al., 2003).
Complete TIPI Items
The TIPI is free for any purpose, so the full item set can be reproduced. The original instructions read: “Here are a number of personality traits that may or may not apply to you. Please write a number next to each statement to indicate the extent to which you agree or disagree with that statement. You should rate the extent to which the pair of traits applies to you, even if one characteristic applies more strongly than the other.” (Gosling et al., 2003)
Interpret each score on its 1.0-7.0 range. Higher scores indicate more of the labeled trait, so a higher fifth-dimension score means greater emotional stability.
Interpretation
No published cutoff scores or interpretive bands exist for the TIPI. Interpret each dimension score relative to the descriptive means and standard deviations below, or relative to your own sample.
Population Norms
Normative means and standard deviations from the development sample (Gosling et al., 2003, Appendix B):
Dimension
Sample
N
M
SD
Source
Extraversion
U.S. college students
1,813
4.44
1.45
Gosling et al. (2003)
Agreeableness
U.S. college students
1,813
5.23
1.11
Gosling et al. (2003)
Conscientiousness
U.S. college students
1,813
5.40
1.32
Gosling et al. (2003)
Emotional Stability
U.S. college students
1,813
4.83
1.42
Gosling et al. (2003)
Openness
U.S. college students
1,813
5.38
1.07
Gosling et al. (2003)
These values describe University of Texas undergraduates and assume the official scoring key above; they are descriptive reference points from the development sample, not population norms for other groups.
Research Evidence and Psychometric Properties
Reliability Evidence
Internal consistency: α = .68 (Extraversion), .40 (Agreeableness), .50 (Conscientiousness), .73 (Emotional Stability), and .45 (Openness) in the development sample (Gosling et al., 2003); Ehrhart et al. (2009) reported comparable values in an independent sample (.71, .34, .56, .65, and .52, respectively)
Low alphas are expected for two-item scales, since internal consistency depends partly on test length; the developers argue that convergent validity is the more appropriate criterion for an ultra-brief measure (Gosling et al., 2003)
Test-retest reliability over about 6 weeks (N = 180): r = .62-.77 across dimensions, mean r = .72, approaching the test-retest correlations of longer Big Five measures (Gosling et al., 2003; range corroborated by Ehrhart et al., 2009)
Convergent Validity
With the 44-item Big Five Inventory (BFI) (N = 1,813): r = .87 (Extraversion), .70 (Agreeableness), .75 (Conscientiousness), .81 (Emotional Stability), and .65 (Openness) (Gosling et al., 2003)
With the NEO-PI-R (N = 172): domain correlations of .65 (Extraversion), .59 (Agreeableness), .68 (Conscientiousness), .66 (Emotional Stability, absolute correlation with the NEO Neuroticism domain), and .56 (Openness) (Gosling et al., 2003)
With the 50-item IPIP-FFM (N = 902): scale-score correlations of .82 (Extraversion), .51 (Agreeableness), .69 (Conscientiousness), .75 (Emotional Stability), and .48 (Openness/Intellect) (Ehrhart et al., 2009)
At the latent-factor level, with measurement error modeled, TIPI-IPIP-FFM factor correlations ranged from .78 (Openness/Intellect) to 1.00 (Extraversion) (Ehrhart et al., 2009)
Factor Structure
Latent-variable confirmatory factor analysis in a large, ethnically diverse undergraduate sample (approximately 40% non-white) supported the five-factor structure with reasonable model fit (RMSEA = .08, SRMR = .05) (Ehrhart et al., 2009)
Standardized factor loadings ranged from .31 to .96, all significant, with items loading on their intended factors; the weakest loading was on one Agreeableness item (Ehrhart et al., 2009)
Discriminant pattern: cross-trait (discriminant) correlations with the BFI averaged |r| = .20 with a maximum of .36, supporting the independence of the five scores (Gosling et al., 2003)
Cross-Cultural Adaptations
German TIPI-G: construct validation with self and peer ratings against the NEO-PI-R found the ten items an efficient approximation of longer five-factor measures (Muck et al., 2007)
Dutch adaptation: generally positive results for factor structure and convergent and discriminant validity (as reviewed by Ehrhart et al., 2009)
Criterion Evidence
The TIPI shows patterns of external correlates similar to those of longer measures, with somewhat weaker magnitudes (Gosling et al., 2003). At the level of the Big Five model itself (not TIPI-specific evidence): a meta-analysis of 19 samples (N = 3,848) found that four of the five dimensions in one partner correlated with the other partner’s relationship satisfaction, namely lower neuroticism (r = −.22, the largest effect), higher agreeableness (r = .15), higher conscientiousness (r = .12), and higher extraversion (r = .06); openness was not significant (Malouff et al., 2010).
Usage Guidelines and Applications
Appropriate Applications
Large-scale surveys, panel studies, and online research where personality is a secondary or control variable
Group-level analyses in which measurement error can be handled statistically
Preliminary description before more comprehensive assessment
Research Design Considerations
Plan for measurement error: two-item scales carry substantial error, so favor larger samples, account for reliability in power planning, and consider latent-variable modeling or attenuation corrections
Report your sample’s internal consistency, even when low, and acknowledge the brevity trade-off when interpreting results, particularly null findings
Supplement with a longer Big Five measure in a subsample when key hypotheses depend on personality, and focus interpretation on effect sizes and patterns rather than precise point estimates
Use the official scoring key: reverse items 2, 4, 6, 8, and 10 and report the fifth dimension as Emotional Stability; the published norms assume this keying (Gosling et al., 2003)
Cultural Considerations
German and Dutch adaptations have shown generally positive validity results (Muck et al., 2007; Ehrhart et al., 2009); check for a validated adaptation in the target language before use
The brief length lowers translation costs, but brevity does not exempt a new translation from validation
Limitations and Cautions
Lower reliability than full-length inventories; unsuitable where precise individual measurement is required
No facet-level information; only the five broad domains are assessed
Not appropriate for clinical assessment, diagnosis, counseling, or high-stakes decisions such as hiring or placement
Limited sensitivity to change, making it a poor choice for measuring personality change in intervention studies
Validation is concentrated in adult college samples; no official age specification exists
Copyright and Usage Responsibility: Check that you have the proper rights and permissions to use this assessment tool in your research. This may include purchasing appropriate licenses, obtaining permissions from authors/copyright holders, or ensuring your usage falls within fair use guidelines.
The TIPI is free for any purpose. The authors state explicitly that anyone can use it and that no permission needs to be requested; no licensing fees apply, and the full item set may be reproduced and administered freely.
Proper Attribution: When using or referencing this scale, cite the original development:
Gosling, S. D., Rentfrow, P. J., & Swann, W. B., Jr. (2003). A very brief measure of the Big-Five personality domains. Journal of Research in Personality, 37(6), 504-528. https://doi.org/10.1016/S0092-6566(03)00046-1
Gosling, S. D., Rentfrow, P. J., & Swann, W. B., Jr. (2003). A very brief measure of the Big-Five personality domains. Journal of Research in Personality, 37(6), 504-528. https://doi.org/10.1016/S0092-6566(03)00046-1
Validation Studies:
Ehrhart, M. G., Ehrhart, K. H., Roesch, S. C., Chung-Herrera, B. G., Nadler, K., & Bradshaw, K. (2009). Testing the latent factor structure and construct validity of the Ten-Item Personality Inventory. Personality and Individual Differences, 47(8), 900-905. https://doi.org/10.1016/j.paid.2009.07.012
Muck, P. M., Hell, B., & Gosling, S. D. (2007). Construct validation of a short five-factor model instrument: A self-peer study on the German adaptation of the Ten-Item Personality Inventory (TIPI-G). European Journal of Psychological Assessment, 23(3), 166-175. https://doi.org/10.1027/1015-5759.23.3.166
Criterion Evidence (Big Five level):
Malouff, J. M., Thorsteinsson, E. B., Schutte, N. S., Bhullar, N., & Rooke, S. E. (2010). The Five-Factor Model of personality and relationship satisfaction of intimate partners: A meta-analysis. Journal of Research in Personality, 44(1), 124-127. https://doi.org/10.1016/j.jrp.2009.09.004
Frequently Asked Questions
What does the TIPI measure?
The TIPI measures the Big Five personality dimensions: Extraversion, Agreeableness, Conscientiousness, Emotional Stability, and Openness to Experience. Each dimension is assessed by two items, each pairing two trait descriptors. The fifth dimension is keyed as Emotional Stability, so higher scores indicate greater emotional stability rather than greater neuroticism.
How long does the TIPI take to complete?
About 1-2 minutes. The developers describe it as taking only a minute to complete, which is why it suits large surveys, panel studies, and online research where personality is one of many measured variables.
Is the TIPI free to use?
Yes. The authors state that anyone can use the TIPI for any purpose without asking permission, and no licensing fees apply. Proper attribution means citing Gosling, Rentfrow, and Swann (2003).
How is the TIPI scored?
Reverse-score items 2, 4, 6, 8, and 10 (reversed score = 8 minus the original score), then average the two items for each dimension: Extraversion is items 1 and 6, Agreeableness 2 and 7, Conscientiousness 3 and 8, Emotional Stability 4 and 9, Openness 5 and 10. Each score ranges from 1.0 to 7.0, with higher scores indicating more of the labeled trait.
How reliable is the TIPI?
Internal consistency in the development sample ranged from α = .40 to .73 across the five dimensions, which is expected for two-item scales because alpha depends partly on test length. Test-retest reliability over about six weeks ranged from r = .62 to .77 (mean .72), approaching the values of longer Big Five measures (Gosling et al., 2003).
How does the TIPI compare with longer Big Five measures?
Convergent correlations with the 44-item Big Five Inventory ranged from r = .65 to .87 across dimensions (Gosling et al., 2003). Against the 50-item IPIP-FFM, scale-score correlations ranged from .48 to .82, and latent-factor correlations with measurement error modeled ranged from .78 to 1.00 (Ehrhart et al., 2009). The trade-off is lower precision and no facet-level information.
Are there other versions of the TIPI?
Yes. The same development paper introduced the FIPI, a five-item variant with one item per dimension (Gosling et al., 2003). A German adaptation, the TIPI-G, was validated by Muck, Hell, and Gosling (2007), a Dutch adaptation has shown generally positive results (as reviewed by Ehrhart et al., 2009), and many later translations exist with varying validation quality.