Abstract
Projective techniques are personality assessment methods that MeSH classifies under personality tests, in which a person responds to a deliberately ambiguous stimulus so that the structure they impose is read as a sign of private aspects of personality. Lawrence Frank named the family in 1939, uniting Hermann Rorschach's inkblots and Henry Murray's Thematic Apperception Test under a single projective hypothesis. Because the stimulus is unstructured, scoring is the central problem, and much of the field's history is the effort to make interpretation reliable, culminating in Exner's Comprehensive System and its successor, the Rorschach Performance Assessment System. These instruments remain among the most contested in clinical psychology: meta-analysis supports some scores while critics question the incremental validity of many. This article describes the major techniques, their scoring, and the long debate over what they measure.
Keywords: projective techniques, projective hypothesis, incremental validity
- A projective technique presents an ambiguous stimulus and treats the way a person organizes it as a projection of personality, following the projective hypothesis Lawrence Frank set out in 1939.
- The two techniques MeSH breaks out as separate descriptors are the inkblot tests and the Thematic Apperception Test; word association and figure-drawing methods belong to the same family.
- Ambiguity is the source of both the method's appeal and its difficulty: it invites rich responses but makes reliable scoring hard, which is why Exner's Comprehensive System and the later R-PAS exist.
- The instruments are heavily contested; meta-analysis validates some scores, but the standing question is incremental validity, whether a projective score adds anything beyond cheaper measures.
What Projective Techniques Are
A projective technique is an assessment method that presents a person with an ambiguous stimulus and interprets the response as a disclosure of personality. The defining feature is the ambiguity of the material. Where a self-report questionnaire asks a direct question and records a chosen answer, a projective task supplies an inkblot, a vague picture, or an unfinished sentence and leaves the person to decide what it is, what is happening, or how it ends. The premise is that with little structure in the stimulus, whatever structure appears in the answer must come from the respondent.
Lawrence Frank gave the family its name and its rationale in 1939, arguing that a projective method induces the person to reveal their private world of meaning by imposing organization on an unstructured field (Frank, 1939). The idea drew on the clinical instruments already in use: Hermann Rorschach had published his inkblot method in 1921, and Christiana Morgan and Henry Murray introduced the Thematic Apperception Test in 1935 as a way to elicit the fantasies and needs that Murray's personology took to organize behavior (Morgan & Murray, 1935). Frank's term unified these under one hypothesis: the projective hypothesis, that responses to ambiguous stimuli express enduring dispositions the person may not report directly.
Figure 1
Three projective stimulus formats and the free responses they invite
The interactive figure below makes the projective hypothesis concrete. As the stimulus is made more ambiguous, the number of equally defensible readings grows, so more of what the response contains must be supplied by the viewer rather than by the blot.
Modern authors increasingly prefer the label performance-based to projective, on the grounds that the tests measure what a person actually does with a task rather than an assumed act of psychological projection, and that the older name imports a contested theory into the instrument's very description (Meyer & Kurtz, 2006). The distinction matters because the projective hypothesis is a theory about mechanism, whereas the tests can be evaluated as behavior samples without committing to it.
Types of Projective Techniques
MeSH places projective techniques beneath personality tests and divides them into two narrower descriptor classes, distinguished by the kind of ambiguous material the person responds to.
| Subtype | MeSH descriptor | What the person responds to | Representative instrument |
|---|---|---|---|
| Ink Blot Tests | D007282 | Symmetric ambiguous inkblots, reporting what each might be | Rorschach test |
| Thematic Apperception Test | D013803 | Ambiguous interpersonal scenes, telling a story about each | Thematic Apperception Test (TAT) |
The division is one of stimulus format, not of theory: both classes rest on the same projective hypothesis and differ only in whether the ambiguous field is a perceptual pattern to be identified or a social scene to be narrated. The two categories are not exhaustive of practice — word association tasks and figure-drawing methods are equally projective — but they are the two the thesaurus breaks out as separate headings. As with any MeSH placement, the classification is an indexing scheme built for literature retrieval, not a theoretical taxonomy of personality assessment: it records how the published work is catalogued, not a claim that these two subtypes exhaust the projective family.
The Major Techniques
The Rorschach is the prototype. A person is shown ten symmetric inkblots in a fixed order and asked, for each, what it might be; the examiner records the responses verbatim and then, in an inquiry phase, asks what about the blot prompted each percept. What is scored is not the content alone but the determinants of the response: whether the person used the whole blot or a detail, whether form, color, shading, or perceived movement drove the percept, and how conventional or idiosyncratic the reading is. These formal features, rather than the manifest imagery, carry the interpretive weight.
The Thematic Apperception Test works by narrative rather than perception. The person is shown a series of ambiguous pictures, most depicting people in situations open to many readings, and is asked to tell a story for each: what led up to the scene, what is happening, what the characters feel, and how it turns out. The stories are read for recurring needs, conflicts, and views of relationships, following the personology Murray built the test to serve (Morgan & Murray, 1935). Contemporary TAT research scores the narratives with structured systems and asks how far the pictures themselves, as opposed to the person, shape the ratings (Siefert et al., 2016).
Beyond these two, word association tests present single words and record what the person says back, and figure-drawing tasks ask the person to draw a person, a house, or a tree and read the drawing for clues to self-concept. All share the projective logic: an underspecified task, a free response, and an inference from how the response is organized to what the person is like.
Scoring and the Comprehensive System
Ambiguity creates the scoring problem. A stimulus open to many readings yields responses open to many codings, and for decades the Rorschach was administered and scored in several incompatible ways, so that a score meant little without naming the school that produced it. John Exner resolved this in 1974 by combining the empirically defensible elements of the competing systems into a single standardized method, the Comprehensive System, with fixed administration, explicit coding rules, and normative reference data. Standardization is what lets two examiners assign the same codes to the same response, and studies of the Comprehensive System reported that many of its variables could be scored with acceptable interrater agreement (Meyer et al., 2002).
Reliability, the consistency of scoring, is a precondition for validity but not a guarantee of it: raters can agree precisely on a code that predicts nothing. The demonstration below separates the two ideas, showing how imposing scoring structure raises interrater agreement without, by itself, saying anything about whether the agreed-upon score is meaningful.
Exner's system became the field standard, but its norms were later criticized for making healthy respondents look disturbed, and a successor was built to address this. The Rorschach Performance Assessment System, introduced in 2011, retained the variables with the best empirical support, added an international normative sample, and adjusted scores for the number of responses a person gives. Interrater-reliability studies of the newer system, in both nonpatient United States and European samples, report agreement comparable to or better than the Comprehensive System it replaced (Kivisalu et al., 2016); (Pignolo et al., 2017).
The Validity Debate
Whether projective techniques measure what they claim to has been argued for half a century, and the argument turns on construct validity, the degree to which test scores behave as the underlying theory says the construct should (Cronbach & Meehl, 1955). A widely cited review concluded that the scientific status of the major projective techniques was weak, that many scores lacked demonstrated validity, and that some were used well beyond their evidence (Lilienfeld, Wood, & Garb, 2000). A companion analysis traced how the Rorschach controversy arose from a mismatch between the instrument's clinical popularity and its thin empirical base (Garb, Wood, Lilienfeld, & Nezworski, 2005). Part of that mismatch has a documented cognitive root: Loren and Jean Chapman showed that clinicians reliably report seeing associations between projective-test signs and symptoms that the data do not contain, an illusory correlation driven by the prior expectation that the sign and the symptom go together rather than by any real covariation (Chapman & Chapman, 1967). This is why confident clinical experience is not evidence that a sign is valid: the same expectancy that makes a sign feel diagnostic also manufactures the correlation that seems to confirm it. The professional body responded with an official statement defending the responsible use of the Rorschach while acknowledging that not all of its scores were equally supported (Society for Personality Assessment, 2005).
The decisive move was to stop arguing about the instrument as a whole and to test its variables one at a time. A large systematic review and meta-analysis of the Comprehensive System found that some Rorschach variables had substantial validity against relevant external criteria while others had little, so that the honest verdict was neither wholesale endorsement nor wholesale rejection but a variable-by-variable ledger (Mihura, Meyer, Dumitrascu, & Bombel, 2013). Critics reanalyzed the same evidence and argued that publication bias and coding choices inflated the supportive findings (Wood, Garb, Nezworski, Lilienfeld, & Duke, 2015), and the original authors replied defending their standards and conclusions (Mihura, Meyer, Bombel, & Dumitrascu, 2015). The exchange narrowed the dispute from whether the Rorschach works to which specific scores work and how well.
Underneath the meta-analytic argument sits a sharper test: incremental validity, whether a projective score improves prediction beyond what cheaper, simpler measures already provide (Sechrest, 1963). A test can be reliable and even correlate with a criterion yet add nothing once history, interview, and a self-report inventory are in hand. The demonstration below partitions predictive variance to show why incremental validity, not raw validity, is the demanding standard a costly projective test must meet.
Worked Example
Consider a clinician deciding whether adding a Rorschach to an assessment battery is worth the hour it costs. Suppose an existing battery of history and a self-report inventory already predicts a clinical outcome with a multiple correlation of R = 0.50, so it accounts for R squared, or 0.25, of the variance in the outcome. Suppose the Rorschach score correlates r = 0.30 with the same outcome on its own, accounting for 0.09 of the variance in isolation.
The question incremental validity asks is not what the Rorschach explains alone but what it adds. If the Rorschach were entirely redundant with the existing battery, it would add nothing despite its 0.09; if it were entirely independent, it would add close to its full 0.09. Suppose it overlaps the battery moderately, sharing about half of its predictive variance with what is already measured. Then the unique variance it contributes is roughly 0.09 times (1 minus 0.5), which is about 0.045.
Adding that increment moves the explained variance from 0.25 to about 0.295, and the battery's multiple correlation from 0.50 to the square root of 0.295, or about 0.54. The gain is real but modest: a 4.5 percentage-point rise in variance explained, bought at the cost of a lengthy administration and expert scoring. Whether that trade is worthwhile depends on the decision at stake. The lesson is the one the incremental-validity standard always teaches: a test's value is measured at the margin, against everything already known, not by its correlation with the outcome in isolation.
Discussion
The enduring tension in projective assessment is between richness and rigor. The same ambiguity that lets an inkblot or a picture elicit material a questionnaire would never reach is what makes the responses hard to score reliably and harder still to validate. For much of the twentieth century the instruments were used far more confidently than their evidence warranted, and the critical literature of the 1990s and 2000s was a necessary correction to that overreach (Lilienfeld et al., 2000).
The correction did not settle into simple rejection. By forcing the debate down to the level of individual scores, meta-analysis produced a differentiated picture in which some variables earn their place and others do not (Mihura et al., 2013). The reframing of these instruments as performance-based tests, evaluated by ordinary psychometric standards rather than defended or attacked as a bloc, is the mature form of the field (Meyer & Kurtz, 2006). Read that way, a projective technique is neither a royal road to the unconscious nor a discredited relic, but a behavior sample whose specific scores must each carry their own evidence, and whose worth in any assessment is set by what they add beyond the cheaper measures already at hand (Bornstein, 2017).
Current Directions
Contemporary work runs along two tracks. The first is the continued psychometric rehabilitation of the Rorschach through the Rorschach Performance Assessment System. Recent studies compare the newer system directly against Exner's Comprehensive System, asking which better separates clinical from nonclinical respondents, with results that support the reformed system's discriminant validity (Pianowski, de Villemor-Amaral, & Meyer, 2023). Reliability studies in new national samples continue to test whether its scoring holds up outside its development sample (Pignolo et al., 2017).
The second track applies the same structured, evidence-first stance to the Thematic Apperception Test. Rather than reading stories impressionistically, current TAT research scores narratives with explicit rating systems and asks a pointed methodological question: how much of a rating reflects the person and how much is pulled by the particular card, since some pictures push almost everyone toward the same themes (Siefert et al., 2016). Across both tracks the governing idea is the one that reframed the whole field: treat these as performance-based tests and hold them to evidence-based assessment standards, score by score (Bornstein, 2017).
Key Researchers
Robert F. Bornstein (Adelphi University). Contemporary theorist of process-focused, evidence-based assessment who reframes projective methods as performance-based tests judged by what a person does with the stimulus; see his faculty profile.
John E. Exner (1928-2006). Built the Rorschach Comprehensive System in 1974, integrating competing scoring schools into one standardized method that dominated practice for a generation; see Wikipedia.
Lawrence K. Frank (1890-1968). Coined the term projective methods in 1939 and articulated the projective hypothesis that unified the family; see Wikipedia.
Luciano Giromini (University of Turin). Contemporary Rorschach researcher whose interrater-reliability studies test the scoring reforms of the Rorschach Performance Assessment System; ORCID 0000-0002-9540-4803.
Gregory J. Meyer (University of Toledo). Co-developer of the Rorschach Performance Assessment System and lead analyst of the interrater reliability and validity of Rorschach scores; ORCID 0000-0002-8869-3838.
Joni L. Mihura (University of Toledo). Lead author of the systematic meta-analytic review of individual Rorschach variables that grounded the modern, score-by-score case for the instrument; ORCID 0000-0003-0627-9869.
Henry A. Murray (1893-1988). Co-created the Thematic Apperception Test in 1935 and developed the need-based personology it was built to assess; see Wikipedia.
Hermann Rorschach (1884-1922). Devised the inkblot method in 1921, the founding projective technique, treating the perception of ambiguous blots as a window on personality; see Wikipedia.
James M. Wood (University of Texas at El Paso). Leading critic of projective techniques whose reanalyses press the case that many Rorschach indices lack demonstrated incremental validity; see Wikipedia.
Glossary
- Ambiguous stimulus
- A test material with little inherent structure, such as an inkblot or a vague picture, whose openness is what lets a response reveal the respondent rather than the material.
- Apperception
- The act of interpreting new perceptions in light of past experience; the Thematic Apperception Test is named for the way a person's stories reshape an ambiguous picture through prior needs and conflicts.
- Comprehensive System
- John Exner's 1974 standardization of Rorschach administration, coding, and norms, which unified previously incompatible scoring schools into one method.
- Construct validity
- The degree to which test scores behave as the theory of the underlying construct predicts; the standard against which projective scores are judged.
- Determinant
- In Rorschach scoring, the feature of the blot that drove a percept, such as form, color, shading, or perceived movement; the formal properties that carry interpretive weight.
- Illusory correlation
- The perception of a relationship between two variables, such as a test sign and a symptom, that is absent or weaker in the data than the observer believes, driven by prior expectation; a documented reason clinical confidence in projective signs outran their evidence.
- Incremental validity
- The extent to which a test improves prediction beyond what cheaper or existing measures already provide; the demanding standard a costly projective test must meet.
- Ink blot tests
- The MeSH subtype of projective techniques using symmetric inkblots, of which the Rorschach is the prototype.
- Interrater reliability
- The consistency with which independent scorers assign the same codes to the same responses; a precondition for validity but not a substitute for it.
- Performance-based assessment
- A relabeling of projective testing that describes the tests as samples of what a person does with a task, avoiding the theoretical commitment carried by the word projective.
- Personology
- Henry Murray's theory of personality organized around needs and press, which the Thematic Apperception Test was designed to assess.
- Projective hypothesis
- The premise that responses to an ambiguous stimulus express enduring dispositions the person may not report directly, because the structure of the answer must originate in the respondent.
- Rorschach Performance Assessment System
- The 2011 successor to the Comprehensive System, retaining the best-supported variables, adding international norms, and adjusting for the number of responses a person gives.
- Thematic Apperception Test
- A projective technique in which a person tells stories about ambiguous interpersonal pictures, scored for recurring needs, conflicts, and views of relationships.
- Word association test
- A projective method that presents single words and records the responses given back, reading the associations for signs of conflict or preoccupation.
Frequently Asked Questions
What is a projective technique?
It is a personality assessment method that presents an ambiguous stimulus, such as an inkblot or a vague picture, and interprets how the person organizes it as a disclosure of personality. The premise is that with little structure in the material, the structure of the response must come from the respondent.
What is the projective hypothesis?
It is the assumption, named by Lawrence Frank in 1939, that responses to ambiguous stimuli express enduring dispositions a person may not report directly. It is a theory about mechanism, which is why many researchers now prefer to describe the tests without committing to it.
What are the main projective tests?
The two that MeSH breaks out as separate descriptors are the inkblot tests, of which the Rorschach is the prototype, and the Thematic Apperception Test. Word association tasks and figure-drawing methods belong to the same family.
How is the Rorschach scored?
A person reports what each of ten symmetric inkblots might be, and the examiner codes the formal features of the response, such as whether the whole blot or a detail was used and whether form, color, shading, or movement drove the percept. Exner's Comprehensive System standardized this coding, and the later Rorschach Performance Assessment System refined it.
Are projective tests valid?
It depends on the specific score. Meta-analysis of Rorschach variables found that some have substantial validity while others have little, so the honest answer is a variable-by-variable ledger rather than a single verdict for the instrument as a whole.
What is incremental validity and why does it matter?
Incremental validity is whether a test improves prediction beyond what cheaper, existing measures already provide. It matters because a projective test can correlate with an outcome yet add nothing once history, interview, and a self-report inventory are in hand, and its value is measured at that margin.
Why do some psychologists call them performance-based tests?
Because the label describes what the tests actually do, sample a person's behavior on a task, without importing the contested theory of psychological projection. The change of name is an attempt to evaluate the instruments by ordinary psychometric standards.
Are projective techniques still used?
Yes, though under closer scrutiny than in the past. Current work rehabilitates the Rorschach through the Rorschach Performance Assessment System and applies structured scoring to the Thematic Apperception Test, holding both to evidence-based assessment standards.
References
Bornstein, R. F. (2017). Evidence-based psychological assessment. Journal of Personality Assessment, 99(4), 435-445. https://doi.org/10.1080/00223891.2016.1236343
Chapman, L. J., & Chapman, J. P. (1967). Genesis of popular but erroneous psychodiagnostic observations. Journal of Abnormal Psychology, 72(3), 193-204. https://doi.org/10.1037/h0024670
Cronbach, L. J., & Meehl, P. E. (1955). Construct validity in psychological tests. Psychological Bulletin, 52(4), 281-302. https://doi.org/10.1037/h0040957
Frank, L. K. (1939). Projective methods for the study of personality. The Journal of Psychology, 8(2), 389-413. https://doi.org/10.1080/00223980.1939.9917671
Garb, H. N., Wood, J. M., Lilienfeld, S. O., & Nezworski, M. T. (2005). Roots of the Rorschach controversy. Clinical Psychology Review, 25(1), 97-118. https://doi.org/10.1016/j.cpr.2004.09.002
Kivisalu, T. M., Lewey, J. H., Shaffer, T. W., & Canfield, M. L. (2016). An investigation of interrater reliability for the Rorschach Performance Assessment System (R-PAS) in a nonpatient U.S. sample. Journal of Personality Assessment, 98(4), 382-390. https://doi.org/10.1080/00223891.2015.1118380
Lilienfeld, S. O., Wood, J. M., & Garb, H. N. (2000). The scientific status of projective techniques. Psychological Science in the Public Interest, 1(2), 27-66. https://doi.org/10.1111/1529-1006.002
Meyer, G. J., Hilsenroth, M. J., Baxter, D., Exner, J. E., Fowler, J. C., Piers, C. C., & Resnick, J. (2002). An examination of interrater reliability for scoring the Rorschach Comprehensive System in eight data sets. Journal of Personality Assessment, 78(2), 219-274. https://doi.org/10.1207/S15327752JPA7802_03
Meyer, G. J., & Kurtz, J. E. (2006). Advancing personality assessment terminology: Time to retire objective and projective as personality test descriptors. Journal of Personality Assessment, 87(3), 223-225. https://doi.org/10.1207/s15327752jpa8703_01
Mihura, J. L., Meyer, G. J., Dumitrascu, N., & Bombel, G. (2013). The validity of individual Rorschach variables: Systematic reviews and meta-analyses of the Comprehensive System. Psychological Bulletin, 139(3), 548-605. https://doi.org/10.1037/a0029406
Mihura, J. L., Meyer, G. J., Bombel, G., & Dumitrascu, N. (2015). Standards, accuracy, and questions of bias in Rorschach meta-analyses: Reply to Wood, Garb, Nezworski, Lilienfeld, and Duke (2015). Psychological Bulletin, 141(1), 250-260. https://doi.org/10.1037/a0038445
Morgan, C. D., & Murray, H. A. (1935). A method for investigating fantasies: The Thematic Apperception Test. Archives of Neurology and Psychiatry, 34(2), 289-306. https://doi.org/10.1001/archneurpsyc.1935.02250200049005
Pianowski, G., de Villemor-Amaral, A. E., & Meyer, G. J. (2023). Comparing the validity of the Rorschach Performance Assessment System and Exner's Comprehensive System to differentiate patients and nonpatients. Assessment, 30(8), 2417-2432. https://doi.org/10.1177/10731911221146516
Pignolo, C., Giromini, L., Ando', A., Ghirardello, D., Di Girolamo, M., Ales, F., & Zennaro, A. (2017). An interrater reliability study of Rorschach Performance Assessment System (R-PAS) raw and complexity-adjusted scores. Journal of Personality Assessment, 99(6), 619-625. https://doi.org/10.1080/00223891.2017.1296844
Sechrest, L. (1963). Incremental validity: A recommendation. Educational and Psychological Measurement, 23(1), 153-158. https://doi.org/10.1177/001316446302300113
Siefert, C. J., Stein, M. B., Slavin-Mulford, J., Sinclair, S. J., Haggerty, G., & Blais, M. A. (2016). Estimating the effects of Thematic Apperception Test card content on SCORS-G ratings: Replication with a nonclinical sample. Journal of Personality Assessment, 98(6), 598-607. https://doi.org/10.1080/00223891.2016.1167696
Society for Personality Assessment. (2005). The status of the Rorschach in clinical and forensic practice: An official statement by the Board of Trustees of the Society for Personality Assessment. Journal of Personality Assessment, 85(2), 219-237. https://doi.org/10.1207/s15327752jpa8502_16
Wood, J. M., Garb, H. N., Nezworski, M. T., Lilienfeld, S. O., & Duke, M. C. (2015). A second look at the validity of widely used Rorschach indices: Comment on Mihura, Meyer, Dumitrascu, and Bombel (2013). Psychological Bulletin, 141(1), 236-249. https://doi.org/10.1037/a0036005