TIPI: Ten Item Personality Inventory (Big Five)

Reviewed by: Constantin Rezlescu | Associate Professor | UCL Psychology

TL;DR

  • The TIPI measures the Big Five personality dimensions, Extraversion, Agreeableness, Conscientiousness, Emotional Stability, and Openness to Experience, with a pair of trait descriptors for each, for research settings where a full-length inventory is impractical.
  • Gosling, Rentfrow, and Swann designed it as a deliberate trade-off: each dimension is covered by one standard and one reverse-scored item, so internal consistency is modest by construction and the instrument's validity case rests on convergence with longer Big Five measures.
  • Independent validation work supports the five-factor structure with latent-variable methods, documents adequate test-retest stability, and reports generally positive results for German and Dutch adaptations.
  • It is free for any use without permission, but it is a research instrument for group-level analysis, not a tool for clinical assessment, counseling, or high-stakes individual decisions; note that the fifth dimension is keyed as Emotional Stability, not Neuroticism.

At a Glance

Items 10 (2 per Big Five dimension; each item pairs two trait descriptors)
Administration time About 1-2 minutes
Response format 7-point scale, 1 = disagree strongly, 7 = agree strongly
Scores Five domain scores, each the mean of two items: Extraversion, Agreeableness, Conscientiousness, Emotional Stability (higher = more stable), Openness to Experience
Validated populations Adults; validated primarily in college and university samples
License Free for any purpose; the authors state that no permission is needed
Original citation Gosling, Rentfrow, & Swann (2003), Journal of Research in Personality

Introduction

The Ten-Item Personality Inventory (TIPI) is an ultra-brief self-report measure of the Big Five personality dimensions. Developed by Gosling, Rentfrow, and Swann (2003) for research settings where time is severely constrained, it covers each of the five dimensions with two items, each pairing two trait descriptors, and takes only a minute or two to complete. The instrument embodies a deliberate trade-off: with two items per dimension it cannot match the internal consistency of full-length inventories, but it retains documented convergent, discriminant, and test-retest properties that make it serviceable when a longer measure is impractical (Gosling et al., 2003).

Understanding Ultra-Brief Big Five Measurement

Comprehensive Big Five inventories can require up to 15-45 minutes, and multi-variable surveys, longitudinal panels, and online studies often cannot allocate that time to a single construct. Before instruments like the TIPI were validated, researchers in such settings faced a choice between omitting personality measurement altogether and improvising unvalidated shortcuts; the TIPI was constructed to provide a third option (Gosling et al., 2003). Each item presents a pair of related trait descriptors (for example, “extraverted, enthusiastic” for Extraversion), so that a single rating captures more of a dimension’s breadth than one adjective could, and each dimension is measured by one standard-keyed and one reverse-keyed item.

One labeling detail matters for scoring and interpretation: the TIPI keys its fifth dimension as Emotional Stability, not Neuroticism, so higher scores indicate greater emotional stability (Gosling et al., 2003; Ehrhart et al., 2009).

Theoretical Foundation

The TIPI operationalizes the Big Five (five-factor) model, which organizes personality traits into five broad dimensions identified in decades of lexical research: Extraversion, Agreeableness, Conscientiousness, Neuroticism (keyed on the TIPI as Emotional Stability), and Openness to Experience. Given only ten items, the measure targets the broad domains at the top of the trait hierarchy rather than attempting to assess narrower facets. The developers argued that low alpha coefficients are an expected arithmetic consequence of two-item scales, and that the appropriate evaluation standard for an ultra-brief measure is its convergence with longer instruments and its pattern of external correlates (Gosling et al., 2003).

Key insight: The TIPI trades internal consistency for brevity by design; its validity case rests not on alpha but on convergence with full-length Big Five measures, which is substantial at the scale level and near-complete once measurement error is modeled.

Key Features

Assessment Characteristics

  • 10 items, two per Big Five dimension, each pairing two trait descriptors (Gosling et al., 2003)
  • 7-point response scale from disagree strongly to agree strongly, with the stem “I see myself as:”
  • About 1-2 minutes to complete; the developers describe it as taking only a minute (Gosling et al., 2003)
  • One standard and one reverse-keyed item per dimension; the fifth dimension is keyed as Emotional Stability
  • Free for any purpose; the authors state that no permission is needed to use it

Dimensions Assessed

  • Extraversion – sociability, assertiveness, energy
  • Agreeableness – cooperation, compassion, trust
  • Conscientiousness – organization, dependability, self-discipline
  • Emotional Stability – calm, resilient disposition versus anxiety and negative affect (higher scores = more stable)
  • Openness to Experience – intellectual curiosity, creativity, preference for novelty

Versions & Adaptations

  • Original 10-item TIPI (Gosling et al., 2003)
  • FIPI (Five-Item Personality Inventory), a five-item variant with one item per dimension, developed in the same paper (Gosling et al., 2003)
  • TIPI-G, the German adaptation, validated with self and peer ratings against the NEO-PI-R (Muck, Hell, & Gosling, 2007)
  • Dutch adaptation, with generally positive results for factor structure and convergent and discriminant validity (as reviewed by Ehrhart et al., 2009)
  • Many later translations exist; validation quality varies by language and sample

Research Applications

  • Large-scale surveys where personality is one of many measured variables
  • Longitudinal and multi-wave panel studies requiring repeated brief personality assessment
  • Online behavioral research where participant time and attention are limited
  • Cross-cultural research where translation and administration costs constrain instrument length
  • Studies using personality as a secondary or control variable rather than the primary construct

View Testable Demo

► Click here to try the TIPI: Ten Item Personality Inventory (10 items, full form)

Rate how well ten pairs of personality traits describe you.

Scoring and Interpretation

Response Format

Each pair of traits is rated on a 7-point scale: 1 = disagree strongly, 2 = disagree moderately, 3 = disagree a little, 4 = neither agree nor disagree, 5 = agree a little, 6 = agree moderately, 7 = agree strongly (Gosling et al., 2003).

Complete TIPI Items

The TIPI is free for any purpose, so the full item set can be reproduced. The original instructions read: “Here are a number of personality traits that may or may not apply to you. Please write a number next to each statement to indicate the extent to which you agree or disagree with that statement. You should rate the extent to which the pair of traits applies to you, even if one characteristic applies more strongly than the other.” (Gosling et al., 2003)

“I see myself as:”

  1. Extraverted, enthusiastic (Extraversion)
  2. Critical, quarrelsome (Agreeableness – reversed)
  3. Dependable, self-disciplined (Conscientiousness)
  4. Anxious, easily upset (Emotional Stability – reversed)
  5. Open to new experiences, complex (Openness)
  6. Reserved, quiet (Extraversion – reversed)
  7. Sympathetic, warm (Agreeableness)
  8. Disorganized, careless (Conscientiousness – reversed)
  9. Calm, emotionally stable (Emotional Stability)
  10. Conventional, uncreative (Openness – reversed)

Scoring Procedure

  1. Reverse-score items 2, 4, 6, 8, and 10 (reversed score = 8 – original score) (Gosling et al., 2003, Appendix A).
  2. Average the two items for each dimension: Extraversion = (item 1 + item 6 reversed) ÷ 2; Agreeableness = (item 7 + item 2 reversed) ÷ 2; Conscientiousness = (item 3 + item 8 reversed) ÷ 2; Emotional Stability = (item 9 + item 4 reversed) ÷ 2; Openness = (item 5 + item 10 reversed) ÷ 2.
  3. Interpret each score on its 1.0-7.0 range. Higher scores indicate more of the labeled trait, so a higher fifth-dimension score means greater emotional stability.

Interpretation

No published cutoff scores or interpretive bands exist for the TIPI. Interpret each dimension score relative to the descriptive means and standard deviations below, or relative to your own sample.

Population Norms

Normative means and standard deviations from the development sample (Gosling et al., 2003, Appendix B):

Dimension Sample N M SD Source
Extraversion U.S. college students 1,813 4.44 1.45 Gosling et al. (2003)
Agreeableness U.S. college students 1,813 5.23 1.11 Gosling et al. (2003)
Conscientiousness U.S. college students 1,813 5.40 1.32 Gosling et al. (2003)
Emotional Stability U.S. college students 1,813 4.83 1.42 Gosling et al. (2003)
Openness U.S. college students 1,813 5.38 1.07 Gosling et al. (2003)

These values describe University of Texas undergraduates and assume the official scoring key above; they are descriptive reference points from the development sample, not population norms for other groups.

Research Evidence and Psychometric Properties

Reliability Evidence

  • Internal consistency: α = .68 (Extraversion), .40 (Agreeableness), .50 (Conscientiousness), .73 (Emotional Stability), and .45 (Openness) in the development sample (Gosling et al., 2003); Ehrhart et al. (2009) reported comparable values in an independent sample (.71, .34, .56, .65, and .52, respectively)
  • Low alphas are expected for two-item scales, since internal consistency depends partly on test length; the developers argue that convergent validity is the more appropriate criterion for an ultra-brief measure (Gosling et al., 2003)
  • Test-retest reliability over about 6 weeks (N = 180): r = .62-.77 across dimensions, mean r = .72, approaching the test-retest correlations of longer Big Five measures (Gosling et al., 2003; range corroborated by Ehrhart et al., 2009)

Convergent Validity

  • With the 44-item Big Five Inventory (BFI) (N = 1,813): r = .87 (Extraversion), .70 (Agreeableness), .75 (Conscientiousness), .81 (Emotional Stability), and .65 (Openness) (Gosling et al., 2003)
  • With the NEO-PI-R (N = 172): domain correlations of .65 (Extraversion), .59 (Agreeableness), .68 (Conscientiousness), .66 (Emotional Stability, absolute correlation with the NEO Neuroticism domain), and .56 (Openness) (Gosling et al., 2003)
  • With the 50-item IPIP-FFM (N = 902): scale-score correlations of .82 (Extraversion), .51 (Agreeableness), .69 (Conscientiousness), .75 (Emotional Stability), and .48 (Openness/Intellect) (Ehrhart et al., 2009)
  • At the latent-factor level, with measurement error modeled, TIPI-IPIP-FFM factor correlations ranged from .78 (Openness/Intellect) to 1.00 (Extraversion) (Ehrhart et al., 2009)

Factor Structure

  • Latent-variable confirmatory factor analysis in a large, ethnically diverse undergraduate sample (approximately 40% non-white) supported the five-factor structure with reasonable model fit (RMSEA = .08, SRMR = .05) (Ehrhart et al., 2009)
  • Standardized factor loadings ranged from .31 to .96, all significant, with items loading on their intended factors; the weakest loading was on one Agreeableness item (Ehrhart et al., 2009)
  • Discriminant pattern: cross-trait (discriminant) correlations with the BFI averaged |r| = .20 with a maximum of .36, supporting the independence of the five scores (Gosling et al., 2003)

Cross-Cultural Adaptations

  • German TIPI-G: construct validation with self and peer ratings against the NEO-PI-R found the ten items an efficient approximation of longer five-factor measures (Muck et al., 2007)
  • Dutch adaptation: generally positive results for factor structure and convergent and discriminant validity (as reviewed by Ehrhart et al., 2009)

Criterion Evidence

The TIPI shows patterns of external correlates similar to those of longer measures, with somewhat weaker magnitudes (Gosling et al., 2003). At the level of the Big Five model itself (not TIPI-specific evidence): a meta-analysis of 19 samples (N = 3,848) found that four of the five dimensions in one partner correlated with the other partner’s relationship satisfaction, namely lower neuroticism (r = −.22, the largest effect), higher agreeableness (r = .15), higher conscientiousness (r = .12), and higher extraversion (r = .06); openness was not significant (Malouff et al., 2010).

Usage Guidelines and Applications

Appropriate Applications

  • Large-scale surveys, panel studies, and online research where personality is a secondary or control variable
  • Longitudinal designs requiring repeated brief personality measurement
  • Group-level analyses in which measurement error can be handled statistically
  • Preliminary description before more comprehensive assessment

Research Design Considerations

  • Plan for measurement error: two-item scales carry substantial error, so favor larger samples, account for reliability in power planning, and consider latent-variable modeling or attenuation corrections
  • Report your sample’s internal consistency, even when low, and acknowledge the brevity trade-off when interpreting results, particularly null findings
  • Supplement with a longer Big Five measure in a subsample when key hypotheses depend on personality, and focus interpretation on effect sizes and patterns rather than precise point estimates
  • Use the official scoring key: reverse items 2, 4, 6, 8, and 10 and report the fifth dimension as Emotional Stability; the published norms assume this keying (Gosling et al., 2003)

Cultural Considerations

  • German and Dutch adaptations have shown generally positive validity results (Muck et al., 2007; Ehrhart et al., 2009); check for a validated adaptation in the target language before use
  • The brief length lowers translation costs, but brevity does not exempt a new translation from validation

Limitations and Cautions

  • Lower reliability than full-length inventories; unsuitable where precise individual measurement is required
  • No facet-level information; only the five broad domains are assessed
  • Not appropriate for clinical assessment, diagnosis, counseling, or high-stakes decisions such as hiring or placement
  • Limited sensitivity to change, making it a poor choice for measuring personality change in intervention studies
  • Validation is concentrated in adult college samples; no official age specification exists

Import & Customize Testable Template

► Import scale to your Testable account – Add this scale. Modify instructions, edit questions, adjust presentation. Test anyone (including yourself)

► Try Testable version – View the full implementation of this scale in Testable.

► View detailed implementation guide in Testable – Step by step instructions for complete customization.

► Browse other tests and scales in Testable Library – The largest collection of ready-made psychological tests and scales.

Copyright and Usage Responsibility: Check that you have the proper rights and permissions to use this assessment tool in your research. This may include purchasing appropriate licenses, obtaining permissions from authors/copyright holders, or ensuring your usage falls within fair use guidelines.

The TIPI is free for any purpose. The authors state explicitly that anyone can use it and that no permission needs to be requested; no licensing fees apply, and the full item set may be reproduced and administered freely.

Proper Attribution: When using or referencing this scale, cite the original development:

References

Primary Development:

Validation Studies:

  • Ehrhart, M. G., Ehrhart, K. H., Roesch, S. C., Chung-Herrera, B. G., Nadler, K., & Bradshaw, K. (2009). Testing the latent factor structure and construct validity of the Ten-Item Personality Inventory. Personality and Individual Differences, 47(8), 900-905. https://doi.org/10.1016/j.paid.2009.07.012
  • Muck, P. M., Hell, B., & Gosling, S. D. (2007). Construct validation of a short five-factor model instrument: A self-peer study on the German adaptation of the Ten-Item Personality Inventory (TIPI-G). European Journal of Psychological Assessment, 23(3), 166-175. https://doi.org/10.1027/1015-5759.23.3.166

Criterion Evidence (Big Five level):

  • Malouff, J. M., Thorsteinsson, E. B., Schutte, N. S., Bhullar, N., & Rooke, S. E. (2010). The Five-Factor Model of personality and relationship satisfaction of intimate partners: A meta-analysis. Journal of Research in Personality, 44(1), 124-127. https://doi.org/10.1016/j.jrp.2009.09.004

Frequently Asked Questions

What does the TIPI measure?

The TIPI measures the Big Five personality dimensions: Extraversion, Agreeableness, Conscientiousness, Emotional Stability, and Openness to Experience. Each dimension is assessed by two items, each pairing two trait descriptors. The fifth dimension is keyed as Emotional Stability, so higher scores indicate greater emotional stability rather than greater neuroticism.

How long does the TIPI take to complete?

About 1-2 minutes. The developers describe it as taking only a minute to complete, which is why it suits large surveys, panel studies, and online research where personality is one of many measured variables.

Is the TIPI free to use?

Yes. The authors state that anyone can use the TIPI for any purpose without asking permission, and no licensing fees apply. Proper attribution means citing Gosling, Rentfrow, and Swann (2003).

How is the TIPI scored?

Reverse-score items 2, 4, 6, 8, and 10 (reversed score = 8 minus the original score), then average the two items for each dimension: Extraversion is items 1 and 6, Agreeableness 2 and 7, Conscientiousness 3 and 8, Emotional Stability 4 and 9, Openness 5 and 10. Each score ranges from 1.0 to 7.0, with higher scores indicating more of the labeled trait.

How reliable is the TIPI?

Internal consistency in the development sample ranged from α = .40 to .73 across the five dimensions, which is expected for two-item scales because alpha depends partly on test length. Test-retest reliability over about six weeks ranged from r = .62 to .77 (mean .72), approaching the values of longer Big Five measures (Gosling et al., 2003).

How does the TIPI compare with longer Big Five measures?

Convergent correlations with the 44-item Big Five Inventory ranged from r = .65 to .87 across dimensions (Gosling et al., 2003). Against the 50-item IPIP-FFM, scale-score correlations ranged from .48 to .82, and latent-factor correlations with measurement error modeled ranged from .78 to 1.00 (Ehrhart et al., 2009). The trade-off is lower precision and no facet-level information.

Are there other versions of the TIPI?

Yes. The same development paper introduced the FIPI, a five-item variant with one item per dimension (Gosling et al., 2003). A German adaptation, the TIPI-G, was validated by Muck, Hell, and Gosling (2007), a Dutch adaptation has shown generally positive results (as reviewed by Ehrhart et al., 2009), and many later translations exist with varying validation quality.
Last Updated: