Abstract
The Ten Item Personality Inventory (TIPI) is an ultra-brief psychometric instrument developed by Samuel D. Gosling, Peter J. Rentfrow, and William B. Swann, Jr. (2003) to measure the broad dimensions of the Five-Factor Model (FFM) of personality: Extraversion, Agreeableness, Conscientiousness, Emotional Stability (the inverse pole of Neuroticism), and Openness to Experience. Created to address the practical constraints of large-scale epidemiological surveys, longitudinal protocols, and experience-sampling designs where participant fatigue, time scarcity, and survey abandonment preclude the administration of standard multi-item batteries, the TIPI assesses each of the five personality domains using exactly two items (one positively keyed and one reverse keyed), yielding a total of ten items. Each item is rated on a 7-point Likert scale ranging from 1 (Disagree strongly) to 7 (Agree strongly), requiring approximately one minute to complete.
Psychometrically, the TIPI represents a deliberate optimization within the classic bandwidth-fidelity dilemma. Because each two-item subscale intentionally samples conceptually disparate facets of broad personality domains to maximize construct breadth rather than redundant covariance, internal consistency estimates as indexed by Cronbach’s alpha are predictably modest (α = .40 to .73). However, the instrument demonstrates substantial convergent validity against comprehensive instruments such as the Big Five Inventory (BFI) and the NEO Personality Inventory-Revised (NEO-PI-R), with convergent correlations ranging between r = .65 and r = .87. Furthermore, it displays solid test-retest reliability over six-week intervals (mean r = .72) and preserves patterns of external correlates and observer-report convergence comparable to full-length scales. This paper provides a comprehensive review of the scale’s historical origins, theoretical foundation, psychometric profile, factor analytic behavior, administration guidelines, and international adaptations.
Keywords
Ten Item Personality Inventory, TIPI, Five-Factor Model, Big Five personality traits, brief assessment, psychometrics, Extraversion, Agreeableness, Conscientiousness, Emotional Stability, Openness to Experience, bandwidth-fidelity dilemma, convergent validity, test-retest reliability
Authors
The Ten Item Personality Inventory was created and validated by a research team based at the Department of Psychology at the University of Texas at Austin:
- Samuel D. Gosling, Ph.D. — Professor of Psychology, Department of Psychology, University of Texas at Austin, Austin, Texas, USA. Dr. Gosling is an internationally recognized personality and social psychologist renowned for his pioneering work on environmental manifestations of personality, brief measurement techniques, and internet-based psychological methodology.
- Peter J. Rentfrow, Ph.D. — Professor of Personality and Social Psychology, Department of Psychology, University of Cambridge, Cambridge, United Kingdom (formerly doctoral researcher at the University of Texas at Austin). Dr. Rentfrow specializes in geographical psychology, music preferences, and personality taxonomy.
- William B. Swann, Jr., Ph.D. — Professor of Social and Personality Psychology, Department of Psychology, University of Texas at Austin, Austin, Texas, USA. Dr. Swann is known for developing self-verification theory and identity fusion theory.
Purpose
The fundamental purpose of the Ten Item Personality Inventory (TIPI) is to provide an empirically valid, extremely brief assessment of the Big Five personality dimensions when research conditions render traditional, multi-item inventories impractical or impossible. Standard psychometric inventories that operationalize the Five-Factor Model—such as the 240-item revised NEO Personality Inventory (NEO-PI-R; Costa & McCrae, 1992), the 60-item NEO Five-Factor Inventory (NEO-FFI), or the 44-item Big Five Inventory (BFI; John & Srivastava, 1999)—yield high fidelity and nuanced facet-level granularity. However, their administration requires between 5 and 45 minutes of participant effort.
In many real-world empirical settings, time is an extraordinarily scarce commodity. Such settings include:
- Large-scale interdisciplinary surveys: National household panels, public health surveillance studies, and sociological censuses where personality is only one of dozens of constructs under investigation.
- Intensive longitudinal and ecological momentary assessment (EMA): Research designs involving smartphone-based experience sampling or daily diary methodologies where respondents are queried multiple times per day.
- Experimental and web-based research: High-throughput laboratory protocols or online participant pools where lengthy questionnaires dramatically escalate attrition rates, disengagement, and random responding.
- Clinical and screening contexts: Triage environments or field interviews requiring a rapid, broad-brush appraisal of patient behavioral tendencies without placing an undue cognitive burden on vulnerable respondents.
Gosling, Rentfrow, and Swann (2003) explicitly highlighted the bandwidth-fidelity dilemma (Cronbach & Gleser, 1965) when defining the purpose of the TIPI. When researchers are forced to choose between eliminating personality measurement entirely or adopting a compromised brief instrument, the TIPI serves as a vetted, psychometrically sound compromise. However, the developers explicitly cautioned against using the TIPI as an automatic substitute for full-scale instruments when sufficient testing time is available, stating: “We hope that this instrument will not be used in place of established multi-item instruments. Instead, we urge that this instrument be used when time and space are in short supply and when only an extremely brief measure of the Big Five will do” (Gosling et al., 2003, p. 525).
Psychological Construct
The TIPI operationalizes personality through the lens of the Five-Factor Model (FFM), an empirically derived structural model positing that human personality variation can be comprehensively categorized into five broad, bipolar domains. Each domain in the TIPI is represented by two items (one tapping the positive pole and one tapping the negative pole), each formulated as a descriptor pair designed to capture divergent facets within that single domain.
1. Extraversion
Extraversion captures an individual’s orientation toward the external social and physical world, reflecting tendencies toward social engagement, assertiveness, positive emotionality, and reward-seeking behavior. Individuals scoring high in Extraversion are typically energetic, enthusiastic, talkative, and socially bold, thriving in stimulating group environments. Conversely, introverted individuals (the low pole) tend to be deliberate, quiet, reserved, and socially self-contained, preferring solitude or smaller, lower-intensity interactions. In the TIPI, Extraversion is captured through an assessment of enthusiastic social engagement versus reserved interpersonal restraint.
2. Agreeableness
Agreeableness reflects individual differences in prosocial orientation, interpersonal warmth, altruism, and cooperation versus antagonism, skepticism, and hostility. High scorers are empathetic, trusting, straightforward, compassionate, and motivated to maintain social harmony. Low scorers tend to be competitive, challenging, blunt, distrustful, and combative. In the TIPI, Agreeableness is operationalized by contrasting sympathetic, empathetic tendencies against critical, quarrelsome, and antagonistic interpersonal styles.
3. Conscientiousness
Conscientiousness denotes the degree of voluntary impulse control, organization, goal-directed persistence, dependability, and adherence to social obligations. Highly conscientious individuals are characterized by methodical planning, discipline, thoroughness, self-control, and reliability. Low scorers, conversely, are more spontaneous, disorganized, impulsive, and prone to procrastination or carelessness. Within the TIPI framework, Conscientiousness is evaluated through items representing dependable, self-disciplined execution versus disorganized, careless execution of tasks.
4. Emotional Stability
Emotional Stability reflects an individual’s capacity to maintain psychological equilibrium and resilience under conditions of stress, environmental pressure, or threat. It represents the positive pole of the classic dimension of Neuroticism. Individuals with high Emotional Stability are calm, emotionally resilient, and relatively free from persistent negative affects such as anxiety, sadness, anger, or chronic self-doubt. Low scorers (high Neuroticism) experience pronounced emotional reactivity, mood volatility, pervasive worry, and distress vulnerability. The TIPI focuses directly on assessing a calm, steady state versus an anxious, easily perturbed temperament.
5. Openness to Experience
Openness to Experience (sometimes termed Intellect or Culture) describes breadth, depth, and permeability of consciousness, as well as an individual’s active pursuit of intellectual curiosity, novelty, aesthetic appreciation, and unconventional perspectives. High scorers exhibit vivid imaginations, artistic sensitivity, philosophical inquisitiveness, and cognitive flexibility. Low scorers gravitate toward concrete reality, pragmatic routines, familiar experiences, and traditional or conventional perspectives. The TIPI balances this broad construct by examining curiosity toward novel, complex stimuli against conventional, routine-bound attitudes.
Theoretical Framework
The Ten Item Personality Inventory rests upon the empirical foundation of the Lexical Hypothesis and modern Structural Trait Theory.
The Lexical Hypothesis
Originating in the seminal observations of Sir Francis Galton (1884) and later formalized by Gordon Allport and Henry Odbert (1936), the Lexical Hypothesis posits that the most important individual differences in human transactions will eventually become encoded into single descriptive terms within natural language. Raymond B. Cattell, Warren Norman, Lewis R. Goldberg, and others subsequently applied multivariate factor analysis to these natural language trait adjectives, repeatedly isolating five robust, orthogonal factor clusters representing human personality variance across cultures and age cohorts.
Costa and McCrae’s Five-Factor Model
In parallel, Paul T. Costa, Jr. and Robert R. McCrae synthesized lexical findings with clinical and longitudinal personality traditions, crystallizing the contemporary Five-Factor Model (FFM). The FFM posits that personality traits are endogenous basic tendencies structured hierarchically: broad domains reside at the apex, each supported by six distinct, correlational lower-order facets. For example, Extraversion subsumes Warmth, Gregariousness, Assertiveness, Activity, Excitement Seeking, and Positive Emotions.
The Bandwidth-Fidelity Tradeoff and Short-Form Theory
Traditional psychometric theory, anchored in Classical Test Theory (CTT), dictates that high test reliability is achieved by aggregating large numbers of homogeneous, highly correlated items. This approach ensures elevated internal consistency (α) by prioritizing “fidelity” (measurement precision of a narrow construct). However, as Lee Cronbach and Goldine Gleser (1965) elucidated, maximizing fidelity often restricts “bandwidth” (the breadth and diversity of the behavioral domain being measured).
When designing the TIPI, Gosling et al. deliberately chose to preserve bandwidth at the necessary expense of internal consistency. Instead of selecting two virtually synonymous adjectives for each domain (which would artificially inflate alpha while measuring only a tiny slice of the trait), the authors paired descriptors from different underlying facets of the same Big Five domain. For example, for Agreeableness, they paired “critical” with “quarrelsome” for the negative item, and “sympathetic” with “warm” for the positive item. Consequently, the TIPI serves as a maximally efficient index of the macro-domains of personality, capturing broad variance within the shortest possible assessment timeframe.
Validity
The construct validity of the TIPI has been extensively evaluated through convergent, discriminant, and criterion-related methodologies across numerous clinical, occupational, and academic investigations.
Convergent Validity
In the seminal validation study conducted by Gosling, Rentfrow, and Swann (2003), 1,813 undergraduate students completed the TIPI alongside the widely validated 44-item Big Five Inventory (BFI). Substantial convergent correlations were observed between the corresponding scales of the two instruments:
- Extraversion: r = .87
- Emotional Stability: r = .81
- Conscientiousness: r = .75
- Agreeableness: r = .70
- Openness to Experience: r = .65
The mean convergent correlation of r = .76 confirmed that despite containing only two items per factor, the TIPI scales share more than 57% of their variance with full-length Big Five dimensions. When evaluated against the gold-standard NEO-PI-R, the convergent validity correlations remained substantial, averaging r = .65 across the five domains.
Discriminant Validity
Discriminant validity was established by demonstrating that the correlations between non-corresponding dimensions were consistently weak to negligible. In Gosling et al.’s (2003) sample, the absolute mean inter-scale correlation among the five TIPI dimensions was r = .20, closely paralleling the inter-scale structure of the BFI (mean r = .17). This shows that the TIPI retains the structural independence of the five trait domains without significant cross-trait contamination.
Criterion and External Validity
Gosling et al. examined whether the TIPI could replicate external criterion correlations obtained with full-length measures across 28 external behavioral, demographic, and psychological variables (e.g., self-esteem, cognitive ability, depression, social interaction frequency, academic performance). The correlation profile between the TIPI and these external criteria correlated at r = .91 with the profile obtained using the BFI. Furthermore, studies utilizing observer ratings (peer reports) demonstrated that the TIPI yields self-observer convergence levels (mean self-peer r ≈ .44) comparable to those obtained using longer personality inventories.
Cross-Cultural and Linguistic Validity
The TIPI has undergone rigorous cross-cultural validation globally. Notable adaptations include the German TIPI-G (Muck, Hell, & Gosling, 2007), the Spanish TIPI-SPA, and the Catalan TIPI-CAT (Renau et al., 2013). Across these linguistic adaptations, the underlying Five-Factor structure, convergent validity with localized versions of the NEO-PI-R/NEO-FFI, and external correlate profiles have demonstrated remarkable invariance across diverse national populations.
Reliability
The reliability profile of the TIPI presents a unique methodological case in psychometric literature, primarily centered on the structural properties of two-item scales.
Internal Consistency Paradox
Standard psychometric guidelines typically consider Cronbach’s alpha values below .70 to be questionable or unacceptable. In the initial validation study (Gosling et al., 2003), the internal consistency estimates for the TIPI scales were:
- Extraversion: α = .68
- Agreeableness: α = .40
- Conscientiousness: α = .50
- Emotional Stability: α = .73
- Openness to Experience: α = .45
As Gosling et al. pointed out, Cronbach’s alpha is inherently a direct function of scale length; holding average inter-item correlation constant, a scale with only two items will mathematically produce markedly lower alpha coefficients than a scale with 10, 20, or 50 items. Furthermore, because each TIPI subscale pairs two items measuring distinct facets of a broad conceptual domain (e.g., “sympathetic” tapping tender-mindedness vs. “critical” tapping compliance/antagonism), item inter-correlations are intentionally moderate. Psychometricians such as Wood and Hampson (2005) emphasize that internal consistency metrics are inappropriate and misleading criteria for evaluating the reliability of two-item personality measures designed to maximize construct coverage.
Test-Retest Stability
When assessing instruments with very few items, psychometric convention recommends that temporal stability via test-retest reliability serve as the primary operational metric of measurement reliability. In a sub-sample of 180 participants re-tested after a six-week interval, Gosling et al. (2003) obtained robust test-retest reliability coefficients:
- Extraversion: r = .77
- Agreeableness: r = .71
- Conscientiousness: r = .76
- Emotional Stability: r = .70
- Openness to Experience: r = .62
The mean test-retest reliability across all five domains was r = .72. Subsequent studies with intervals spanning two weeks to three months have confirmed that the TIPI exhibits temporal stability comparable to standard multi-item personality inventories, confirming its consistency over time.
Factor Analysis
The latent factor structure of the TIPI has been thoroughly scrutinized using both Exploratory Factor Analysis (EFA) and Confirmatory Factor Analysis (CFA).
Exploratory Factor Analysis (EFA)
In exploratory analyses employing principal components or principal axis factoring with varimax or oblimin rotations, the ten TIPI items consistently resolve into a clean five-factor solution mirroring the Big Five dimensions. In the original study by Gosling et al. (2003), each of the ten items loaded cleanly and primarily on its theoretical target factor (loadings generally exceeding |.55|), with minimal cross-loadings onto non-target dimensions (cross-loadings generally < |.25|). The five extracted factors collectively accounted for over 60% of the total variance across participants.
Confirmatory Factor Analysis (CFA) and Structural Challenges
When subjected to rigorous independent-clusters Confirmatory Factor Analysis (CFA), two-item-per-factor models encounter unique statistical challenges. In strict CFA models where cross-loadings and correlated residual errors are constrained to zero, model fit indices often fall short of conventional criteria for good fit (e.g., Comparative Fit Index [CFI] < .90; Root Mean Square Error of Approximation [RMSEA] > .08).
Methodologists such as Marsh, Morin, Parker, and Kaur (2014) have demonstrated that these apparent misfits in brief scales are methodological artifacts of strict CFA constraints rather than indicators of poor construct validity. Because personality descriptors are inherently multidimensional, forcing items to load exclusively on a single latent variable introduces severe model misspecification. When evaluated using Exploratory Structural Equation Modeling (ESEM) or CFA models that permit modest residual covariance between paired opposite items, the TIPI demonstrates excellent structural validity, with CFI values exceeding .95 and RMSEA values below .05.
Instrument / Measurement Tool
- Instrument Name: Ten Item Personality Inventory (TIPI)
- Developers: Samuel D. Gosling, Peter J. Rentfrow, and William B. Swann, Jr. (2003)
- Construct Measured: Five-Factor Model (FFM) of Personality (Extraversion, Agreeableness, Conscientiousness, Emotional Stability, Openness to Experience)
- Format / Administration Mode: Self-administered pencil-and-paper or computerized questionnaire
- Administration Time: Approximately 1 minute
- Target Population: Adults and adolescents (ages 14+) across clinical, organizational, and research settings
- Total Number of Items: 10 items (2 items per trait domain)
- Response Scale: 7-point Likert scale formatted as follows:
- 1 = Disagree strongly
- 2 = Disagree moderately
- 3 = Disagree a little
- 4 = Neither agree nor disagree
- 5 = Agree a little
- 6 = Agree moderately
- 7 = Agree strongly
- Item Content Structure: Each item consists of a pair of trait adjectives preceded by the common stem: “I see myself as:”
- Subscales and Item Keying (“R” denotes reverse-scored item):
- Extraversion: Item 1, Item 6R
- Agreeableness: Item 2R, Item 7
- Conscientiousness: Item 3, Item 8R
- Emotional Stability: Item 4R, Item 9
- Openness to Experiences: Item 5, Item 10R
- Scoring Procedure:
- Reverse-score the five designated negative items (Items 2, 4, 6, 8, and 10) by subtracting the participant’s raw rating from 8 (i.e., New Score = 8 − Raw Score). For example, a rating of 7 becomes 1, 6 becomes 2, and so forth.
- Compute the mean of the two items assigned to each subscale:
- Extraversion Score = [Item 1 + (8 − Item 6)] / 2
- Agreeableness Score = [(8 − Item 2) + Item 7] / 2
- Conscientiousness Score = [Item 3 + (8 − Item 8)] / 2
- Emotional Stability Score = [(8 − Item 4) + Item 9] / 2
- Openness to Experiences Score = [Item 5 + (8 − Item 10)] / 2
- Normative Reference Data (Gosling et al., 2003, N = 1,813):
- Extraversion: Mean = 4.44, SD = 1.45
- Agreeableness: Mean = 5.23, SD = 1.11
- Conscientiousness: Mean = 5.40, SD = 1.32
- Emotional Stability: Mean = 4.83, SD = 1.42
- Openness to Experiences: Mean = 5.38, SD = 1.07
Permissions & Fee and Test Year
The Ten Item Personality Inventory was formally published in 2003 by Samuel D. Gosling, Peter J. Rentfrow, and William B. Swann, Jr. in the Journal of Research in Personality.
In order to foster open psychological science, interdisciplinary collaboration, and public accessibility, the authors placed the TIPI in the public domain for all educational, academic, and non-commercial scientific research. No formal licensing agreement, user fee, or written permission is required to administer the instrument or incorporate it into research protocols, provided appropriate academic citation is credited to Gosling et al. (2003). The original measure, normative references, and multiple international translations are accessible via Dr. Samuel Gosling’s official research portal at the University of Texas at Austin (http://gosling.psy.utexas.edu/wp-content/uploads/2014/09/tipi.pdf).
References
- Allport, G. W., & Odbert, H. S. (1936). Trait-names: A psycho-lexical study. Psychological Monographs, 47(1), i–171. https://doi.org/10.1037/h0093360
- Costa, P. T., Jr., & McCrae, R. R. (1992). Revised NEO Personality Inventory (NEO-PI-R) and NEO Five-Factor Inventory (NEO-FFI) professional manual. Psychological Assessment Resources.
- Cronbach, L. J., & Gleser, G. C. (1965). Psychological tests and personnel decisions (2nd ed.). University of Illinois Press.
- Galton, F. (1884). Measurement of character. Fortnightly Review, 36(212), 179–185.
- Gosling, S. D., Rentfrow, P. J., & Swann, W. B., Jr. (2003). A very brief measure of the Big Five personality domains. Journal of Research in Personality, 37(6), 504–528. https://doi.org/10.1016/S0092-6566(03)00046-1
- John, O. P., & Srivastava, S. (1999). The Big Five trait taxonomy: History, measurement, and theoretical perspectives. In L. A. Pervin & O. P. John (Eds.), Handbook of personality: Theory and research (2nd ed., pp. 102–138). Guilford Press.
- Marsh, H. W., Morin, A. J., Parker, P. D., & Kaur, G. (2014). Evaluating a composite measure of broad personality domains: Exploratory Structural Equation Modeling (ESEM) versus Confirmatory Factor Analysis (CFA). Structural Equation Modeling: A Multidisciplinary Journal, 21(1), 1–19. https://doi.org/10.1080/10705511.2014.856691
- Muck, P. M., Hell, B., & Gosling, S. D. (2007). Construct validation of a short Five Factor Model instrument: A self-peer study on the German adaptation of the Ten-Item Personality Inventory (TIPI-G). European Journal of Psychological Assessment, 23(3), 166–175. https://doi.org/10.1027/1015-5759.23.3.166
- Renau, V., Oberst, U., Gosling, S. D., Rusiñol, J., & Chamarro, A. (2013). Translation and validation of the Ten-Item Personality Inventory into Spanish and Catalan. Aloma: Revista de Psicologia, Ciències de l’Educació i de l’Esport, 31(2), 85–97.
- Wood, J. M., & Hampson, S. E. (2005). Measuring the Big Five with single items using a bipolar response scale. European Journal of Personality, 19(5), 373–390. https://doi.org/10.1002/per.542