Abstract
Thematic Apperception Test is a type of projective techniques: a picture-story method in which a respondent invents a dramatic narrative for each of a series of ambiguous scenes, and the examiner content-analyzes those stories for the motives, conflicts, and relationships the respondent reads into them. Henry Murray and Christiana Morgan introduced the cards and the apperceptive rationale in 1935; the decades since produced not one scoring system but many, and the test's scientific standing turns entirely on which one is used. This article treats the TAT as a case study in narrative content coding: how the empirically structured systems earn a reliability that impressionistic reading cannot, why a motive measured from stories diverges from the same motive on a questionnaire, and why the picture-story exercise resists the classical internal-consistency reliability that fits ordinary tests.
Keywords: Thematic Apperception Test, narrative scoring, apperception, implicit motives
The Thematic Apperception Test is, with the Rorschach, one of the two founding projective instruments, but it inverts the Rorschach's premise. Where the inkblot test scores the perceptual process and largely discards the content, the TAT scores the content and largely discards the perception: what matters is the story a respondent builds, not how the picture that prompted it was seen. A respondent is shown a card depicting one or more people in a deliberately unspecified situation and asked to tell a story with a past, a present, and an outcome — what led up to the scene, what is happening, what the characters are feeling and thinking, and how it turns out (Murray, 1943). Henry Murray and Christiana Morgan introduced the method and its rationale in 1935 (Morgan & Murray, 1935). The interpretive wager, apperception, is that in supplying the unstated the respondent projects their own motives and relational templates onto the ambiguous figures. What makes the test scientifically interesting is that this wager can be cashed out in incompatible ways, and only some of them survive measurement scrutiny.
- The TAT is a projective picture-story test: a respondent narrates ambiguous scenes and the examiner codes the stories for projected motives, conflicts, and relationships.
- Its interpretive basis is apperception, the reading of personal meaning into an ambiguous stimulus, so the scored datum is the story's content rather than any perceptual feature.
- There is no single TAT: the same cards support many scoring systems, and reliability and validity are properties of the coding system used, not of the cards themselves.
- The empirically structured systems — need-for-achievement coding, the SCORS-G object-relations scale, and defense-mechanism coding — achieve the inter-coder reliability that global impressionistic interpretation cannot.
- A motive scored from TAT stories (an implicit motive) diverges from the same motive on a self-report questionnaire (a self-attributed motive) and predicts a different class of behavior, so the low correlation between them is a finding, not a failure.
What the Thematic Apperception Test Is
The TAT is a performance-based personality test built from pictures. The standard materials are a set of mostly achromatic cards, each showing one or more people in a scene whose situation, relationships, and outcome are left unspecified, together with one entirely blank card. Murray's original set contained thirty-one cards, from which an examiner selects around twenty, some keyed to the respondent's age and sex, and administers them across two sittings (Murray, 1943). For each card the respondent is asked to make up a story: what has led to the depicted moment, what is happening now, what the characters are thinking and feeling, and how the situation will end.
The scored record is the set of stories, not the respondent's account of what the picture literally shows. Because the scenes are underspecified, no story is entailed by a card; the respondent must import characters' intentions, histories, and outcomes that the picture does not contain, and it is this imported material that the test treats as diagnostic. This is the projective hypothesis in its narrative form: the structure a respondent supplies to an ambiguous social scene is taken to reflect the motives and relational templates through which they habitually construe such scenes (Murray, 1938).
That commitment to narrated content is what distinguishes the TAT from its perceptual sibling and places it in the study of motivation and social cognition rather than perception. Murray named the underlying process apperception: perception shaped by the perceiver's own needs, prior experience, and expectations, so that the same picture becomes a different scene for different people. Where the Rorschach codes how a respondent organizes an ambiguous perceptual field, the TAT codes what a respondent narrates into an ambiguous social one. The interpretive frame Murray supplied was personology, his taxonomy of psychogenic needs — achievement, affiliation, power, and the rest — and of the environmental press that elicits them; a TAT story was to be read for the need-and-press interplay it dramatized.
Scoring Systems for Story Content
The central methodological fact about the TAT is that Murray's cards outlived his scoring scheme. The apperceptive rationale specified what to look for in principle but not a reliable procedure for counting it, and over the following decades several distinct, mutually incompatible systems grew up around the same pictures (Cramer, 1999). The consequence is that the phrase the TAT names a stimulus set, while a TAT score is defined only relative to a system: two examiners can administer the identical cards and produce records that share no common metric, because one counted achievement imagery and the other rated object relations. Any statement about the test's reliability or validity is therefore incomplete until the scoring system is named.
The systems that earned scientific standing share a common form: each specifies explicit content categories, defined closely enough that independent coders applying them to the same story largely agree. The need-for-achievement system codes a story first for the presence of achievement imagery — competition with a standard of excellence, a unique accomplishment, or long-term involvement in attaining a goal — and then for a set of subsidiary categories such as a stated need, instrumental activity, anticipation of success or failure, obstacles, and the overall thema, summing them into a motive score (McClelland et al., 1953). The Social Cognition and Object Relations Scale rates each story on anchored dimensions of interpersonal functioning — the complexity of representations of people, the affective tone of relationships, the capacity for emotional investment, and the understanding of social causality — turning narrative into a rated profile rather than a count (Westen, 1991). Cramer's Defense Mechanism Manual scores denial, projection, and identification from story content along a developmental sequence (Cramer, 1999).
Note. Schematic of the one-story-many-scores logic; the three systems are illustrative of the structured coding traditions, not an exhaustive list.
What all three share, and what free clinical interpretation of the pictures lacks, is that the score is a count or an anchored rating rather than a holistic impression. This is the same move that Exner's standardization made for the Rorschach, but made piecemeal by different investigators for different constructs, which is why the TAT has no single comprehensive system and why its psychometric standing must be assessed one system at a time (Keiser & Prather, 1990).
Table 1. Three structured TAT scoring systems and what each codes.
| System | Construct scored | Unit of measurement |
|---|---|---|
| Need for achievement | Implicit achievement motive | Presence count of defined imagery categories summed across cards |
| SCORS-G | Object relations and social cognition | Anchored ratings on multiple relational dimensions per story |
| Defense Mechanism Manual | Denial, projection, identification | Defense usage scored along a developmental sequence |
Implicit Versus Self-Attributed Motives
The most theoretically important result to come out of TAT motive research is that a motive scored from a person's spontaneous stories and the same motive as the person rates it on a questionnaire are weakly correlated at best, and often essentially unrelated (Spangler, 1992). To a critic this looks like a validity failure: if a TAT achievement score and a self-report achievement score disagree, surely at least one of them is wrong. McClelland and colleagues reframed the disagreement as evidence in the paper that gave the distinction its name and its theoretical basis (McClelland et al., 1989). The two instruments tap two different motivational systems — implicit motives, aroused by intrinsic task incentives and expressed in spontaneous behavior, which operant methods like the TAT capture; and self-attributed motives, the values a person consciously endorses, which self-report captures (McClelland et al., 1989).
Spangler's two meta-analyses put numbers to the distinction. TAT-based and questionnaire-based achievement measures each predict outcomes, but different outcomes: the story-based measure predicts spontaneous, long-run behavioral trends such as entrepreneurial activity and performance over time, while the questionnaire measure predicts responses to explicit social demands and immediate, choice-like tasks (Spangler, 1992). Winter and colleagues generalized the pattern into a two-channel model of personality in which motives, best measured by the TAT, and traits, best measured by self-report, are distinct systems that interact rather than redundant measures of one thing — traits channel the expression of motives rather than substituting for them (Winter et al., 1998).
The methodological upshot is a lesson that recurs across assessment. The classic complaint that TAT motive scores fail to correlate with questionnaire scores mistakes a substantive finding for a psychometric defect. A low convergent correlation is exactly what a two-systems account predicts, so the appropriate validity test is predictive — does each measure forecast the class of behavior it is theorized to govern — not convergent against a differently-defined construct. A measure is not invalidated by disagreeing with a different measure of a different thing.
The Consistency Problem
The TAT has long appeared unreliable by one standard index and defensible by others, and sorting out which index applies is the crux of its psychometric reputation. Internal-consistency reliability — Cronbach's alpha computed across cards — is typically low, and for a straightforward reason: the cards are deliberately heterogeneous. Each pulls for different themes, so a respondent's achievement imagery on one card need not track their imagery on another, and an index that assumes every card measures the same homogeneous construct will read that heterogeneity as inconsistency (Gruber & Kreuzpointner, 2013).
Gruber and Kreuzpointner argued that alpha is simply the wrong model for a picture-story exercise. Alpha presumes parallel items tapping one latent dimension; TAT cards are neither parallel nor unidimensional by design, so a category-based reliability that respects the instrument's multidimensional structure yields a truer and substantially higher estimate (Gruber & Kreuzpointner, 2013). Inter-coder reliability, meanwhile, is high for the structured systems, precisely because their categories are explicitly defined and can be applied the same way by different raters (McClelland et al., 1953).
This mirrors the Rorschach lesson at a different joint. The reliability the TAT genuinely earns is coding agreement under a structured system; the reliability it appears to fail, internal consistency, is being demanded under a measurement model that does not fit the instrument. As with any test, reliability is necessary but not sufficient for validity, and the kind of reliability that is even appropriate to ask for depends on what the score is supposed to be (Cramer, 1999).
Validity and the Scoring System
Because the TAT is a family of scoring systems rather than a single instrument, validity claims must be system-specific, and the general critique of projective techniques made exactly this demand. That critique argued that many TAT uses rested on interpretive traditions with thin predictive support, and that global, impressionistic reading of the pictures should be held to the same evidentiary standard as any other test before it informs a clinical or forensic decision (Lilienfeld et al., 2000).
The structured systems answer the critique on its own predictive ground rather than retreating from it. The need-for-achievement system has meta-analytic evidence that it forecasts the operant outcomes it is theorized to govern, which is the criterion validity a motive measure should be judged by (Spangler, 1992). Object-relations coding discriminates clinical from nonclinical groups and tracks the interpersonal constructs it names, with construct-validity evidence continuing to accumulate for the SCORS-G in clinical samples (Stein et al., 2012). Cramer's defense coding shows the developmental progression and the criterion associations its theory predicts (Cramer, 1999).
The settlement parallels the Rorschach's variable-by-variable resolution. Validity is not a property of the TAT but of a particular scoring system applied to particular cards for a particular inference. Free interpretation of the pictures is weakly supported; specific, structured, empirically anchored systems are supported for the constructs they were built to measure. The correct unit of analysis, once again, is the score and not the stimulus (Lilienfeld et al., 2000).
The Thematic Apperception Test in Motion
The three demonstrations below make the test's logic manipulable. The first builds a need-for-achievement score from a story, showing how presence-coded content categories sum into a motive score and how the same card yields different scores for different narratives. The second lays out the implicit-versus-self-attributed dissociation, showing why two measures of the same-named motive can be uncorrelated yet both valid. The third runs the internal-consistency problem, showing why heterogeneous cards depress classical reliability and how aggregation changes the estimate.
Summed score 6 of 6, from the achievement-saturated story.
Each defined category a story contains adds one point. The striving story carries all six and scores 6; a story told to the identical card that only describes the scene carries none and scores 0. Toggle the categories by hand and the count tracks the narrated content, not the picture — the concrete meaning of scoring the story rather than the stimulus.
The scoring demonstration assembles a single story's achievement-motive score. Toggling the content categories a story contains — achievement imagery, a stated need, instrumental activity, anticipation, an obstacle, and an achievement thema — builds the summed score the way a trained coder would, and switching between two narratives told to the same card shows that the picture fixes nothing: an achievement-saturated story and a purely descriptive one produce entirely different scores from identical stimuli. This is the concrete meaning of scoring the narrated content rather than the card.
At a convergent correlation of 0.10 the two measures overlap by only 1.0%, yet the implicit measure still explains 12.2% of the long-run outcome against the self-attributed measure’s 0.3%.
Slide the convergence as low as it realistically goes and each measure keeps forecasting its own class of behaviour — the story-based motive the spontaneous long-run trend, the questionnaire motive the deliberate immediate choice. A near-zero correlation between them coexists with two genuinely valid measures, which is why the dissociation is a finding rather than a defect.
The motives demonstration sets the true correlation between an implicit (story-based) motive and a self-attributed (questionnaire-based) motive and shows the two predicting different criteria. Sliding the convergence between the measures shows that even at a realistically low correlation each still forecasts its own class of outcome — the implicit measure the spontaneous long-run behavior, the self-attributed measure the deliberate immediate choice — so a near-zero convergent correlation coexists with two genuinely valid measures. It makes visible why the dissociation is a finding rather than a defect.
6 cards at per-card reliability 0.10 give an aggregate reliability of 0.40; reaching 0.80 would take 36 comparable cards.
At the default r = 0.10, six cards land at R = 0.40, and pushing the aggregate to a conventional 0.80 would demand 36 comparable cards — an impractically long exercise from heterogeneous stimuli. That mismatch between the classical model and the picture-story method is exactly what motivated category-based reliability estimates for the instrument.
The reliability demonstration runs the aggregation arithmetic behind the consistency problem. Setting the number of cards and the modest reliability of a single card computes the Spearman-Brown reliability of the aggregate, showing how a short exercise of heterogeneous cards lands at a low classical coefficient and how many cards a given per-card reliability would need to reach a conventional target. It makes concrete why internal consistency is an unforgiving and ill-fitting standard for the picture-story exercise, and why category-based estimates were proposed instead.
Worked Example
Begin with the achievement-motive score the first demonstration builds. A story is coded for the presence of defined categories, each contributing one point: achievement imagery (the story is fundamentally about competing with a standard of excellence), a stated need for achievement, instrumental activity toward the goal, positive anticipation of success, an obstacle overcome, and an overall achievement thema. A story containing all six scores 1 + 1 + 1 + 1 + 1 + 1 = 6, the maximum on this simplified scheme. A second story told to the same card that merely describes the scene — a person sitting at a desk, with no goal, need, or striving — contains none of the categories and scores 0. Identical stimulus, scores of 6 and 0: the score lives in the narrative, not the picture.
Now the implicit-versus-self-attributed dissociation. Suppose the story-based motive predicts a spontaneous long-run outcome at r = 0.35 and the questionnaire motive predicts it at r = 0.05, while for a deliberate immediate choice the pattern reverses. The variance each explains in the long-run outcome is r-squared: 0.35-squared = 0.1225, about 12%, for the implicit measure, against 0.05-squared = 0.0025, about 0.25%, for the self-attributed one — roughly a forty-nine-fold difference in predictive variance despite both carrying the name achievement. Treating the two as interchangeable, or averaging them into one composite, would discard exactly the systematic difference that makes each useful for its own criterion.
Finally, the reliability of aggregation. The Spearman-Brown formula gives the reliability of a test of k comparable parts, each of reliability r, as R = kr / (1 + (k − 1)r). Take a single card's contribution as a modest r = 0.10. A six-card exercise gives R = (6 × 0.10) / (1 + 5 × 0.10) = 0.60 / 1.50 = 0.40. To reach a conventional R ≈ 0.80 at that per-card reliability would take k = (0.80 × (1 − 0.10)) / (0.10 × (1 − 0.80)) = 0.72 / 0.02 = 36 comparable cards. The classical model thus demands an impractically long exercise from heterogeneous cards, which is precisely the mismatch that motivated category-based reliability estimates for the instrument (Gruber & Kreuzpointner, 2013). (The counts here are chosen to show the arithmetic, not to report any study's exact figures.)
Discussion
The Thematic Apperception Test earns its place in cognitive psychology less as a clinical instrument than as a sustained lesson in what a projective measure has to do to be believed, and the lesson differs from the Rorschach's at an instructive point. The Rorschach's problem was that one standardized system had to be sorted variable by variable; the TAT's problem is prior to that — there is no single system to sort, only a shared set of cards beneath a plurality of incompatible coding traditions. Everything that can be said about the test's reliability or validity is therefore indexed to a system, and the systems that succeeded did so by the same means: explicit, countable content categories in place of holistic impression (McClelland et al., 1953).
The deepest contribution to general psychology came from a result that first looked like a failure. The near-independence of story-measured and self-reported motives could have been read as proof that the TAT measures nothing stable; instead it became the empirical anchor for the distinction between implicit and self-attributed motivation, one of the more durable findings in the psychology of motivation (Spangler, 1992). The methodological moral is that a low correlation between two instruments is uninformative until one knows whether they were built to measure the same construct, and the TAT's history is the standing illustration.
The modern settlement is neither the wholesale dismissal the strongest critics urged nor the free interpretive practice the test's tradition once licensed. Structured systems with predictive evidence are retained and used for the constructs they were validated against; global impressionistic reading of the pictures is set aside; and the difference between the two is a matter of which coding manual is in the examiner's hands. A test that spent decades as an emblem of subjective clinical judgment turns out, when scored the right way, to be an unusually clear demonstration of why the unit of psychometric analysis is the score rather than the apparatus (Lilienfeld et al., 2000).
Current Directions
The most active contemporary line of TAT research is the development and validation of the Social Cognition and Object Relations Scale-Global (SCORS-G), a refinement of Westen's original scale into a set of anchored rating dimensions with a detailed manual for clinicians and researchers (Stein & Slavin-Mulford, 2017). The program has pursued exactly the psychometric groundwork the earlier projective literature was faulted for skipping: construct-validity studies in clinical samples establishing that the rated dimensions behave as the theory of object relations predicts (Stein et al., 2012), and careful attention to a confound specific to the picture-story method — that the pull of individual cards, not only the respondent, shapes the ratings. Work estimating how much TAT card content itself drives SCORS-G scores, replicated in nonclinical samples, is the direct analogue of the Rorschach's response-complexity adjustment: an attempt to separate the signal of the person from the artifact of the stimulus (Siefert et al., 2016).
The motive-measurement tradition has modernized on a parallel track, migrating from the clinical TAT to the picture-story exercise as a research instrument and building the standardized infrastructure that operant measurement long lacked: expert-coded story databases, updated picture norms, and shared coding resources that let implicit-motive scoring be applied and audited consistently across laboratories (Sch\u00f6nbrodt et al., 2021).
That maturation has been accompanied by candid internal criticism rather than only advocacy. Recent critical reviews of the SCORS-G and TAT in clinical practice have catalogued the psychometric limitations that remain — variability in card sets and administration, gaps in normative data, and the ethical implications of drawing high-stakes inferences from a still-imperfect measure — and have called for tighter standardization before the instrument's clinical use outruns its evidence (Sinclair et al., 2023). The common thread across the current work is the one the whole history points to: progress comes from treating a specific scoring system as the object of study, validating and where necessary constraining it, rather than defending or dismissing the cards as a whole.
Common Misconceptions
- The TAT reveals hidden meaning from what a person sees in the pictures.
- The pictures are only a prompt. What is scored is the story a respondent constructs — its motives, conflicts, and relationships — under an explicit coding system, not a symbolic decoding of the image itself (Murray, 1943).
- There is one standard way to score the TAT.
- There is not. The same cards support several incompatible systems — need for achievement, object-relations rating, defense coding, and others — and a TAT score has no meaning until the system is specified (Cramer, 1999).
- The TAT is invalid because its scores do not match self-report questionnaires.
- Story-based and questionnaire measures tap distinct motivational systems and predict different behaviors, so their low correlation is a substantive finding rather than evidence that either is wrong (Spangler, 1992).
- Low internal consistency means the TAT is simply unreliable.
- Internal consistency assumes homogeneous, parallel items, which heterogeneous picture cards are not; inter-coder agreement for structured systems is high, and category-based reliability estimates fit the instrument better (Gruber & Kreuzpointner, 2013).
Glossary
- Apperception.
- Perception shaped by the perceiver's own needs, prior experience, and expectations; the process by which a respondent reads personal meaning into an ambiguous picture, giving the test its name.
- Defense Mechanism Manual.
- Cramer's system for scoring denial, projection, and identification from TAT stories along a developmental sequence, one of the empirically structured coding schemes for the test.
- Environmental press.
- In Murray's personology, the pressure a situation or object exerts on a person to act, the external counterpart to an internal need whose interplay a TAT story was meant to dramatize.
- Implicit motive.
- A motive aroused by intrinsic task incentives and expressed in spontaneous behavior, measured by operant methods such as the TAT; distinct from the motive a person consciously endorses.
- Inter-coder reliability.
- The degree to which two raters independently applying the same coding system to the same story assign the same scores; the reliability the structured TAT systems most clearly secure.
- Need for achievement.
- The implicit motive to compete with a standard of excellence, scored from TAT stories through defined imagery categories in the system developed by McClelland and Atkinson.
- Object relations.
- The mental representations of self and others and of relationships between them; the interpersonal domain that the Social Cognition and Object Relations Scale rates from TAT narratives.
- Operant measure.
- A measure that scores spontaneously generated behavior, such as a freely told story, rather than a response to a fixed set of options; the TAT is an operant measure of motives, contrasted with a respondent measure like a questionnaire.
- Personology.
- Murray's theory of personality built on psychogenic needs and environmental press, the interpretive frame within which the TAT was originally to be read.
- Picture-story exercise.
- A general term for a TAT-type task in which a respondent writes or tells stories to ambiguous pictures, emphasizing that the instrument is a method rather than a single fixed test.
- Projective hypothesis.
- The premise that responses to an unstructured stimulus are shaped by the respondent's own dispositions rather than by the stimulus, so the response reveals the person.
- SCORS-G.
- The Social Cognition and Object Relations Scale-Global Rating Method, an anchored multidimensional system for rating object relations and social cognition from TAT stories and the most active contemporary TAT coding program.
- Self-attributed motive.
- A motive a person consciously endorses and reports on a questionnaire; predicts deliberate immediate choices and is largely independent of the corresponding implicit motive.
- Thema.
- In Murray's scheme, the overall dramatic pattern of a story — the interplay of a need and the press it meets — and, in achievement scoring, one of the categories contributing to the motive score.
Key Researchers
Leopold Bellak (1916-2000). Austrian-American psychiatrist whose Bellak scoring system and the derivative Children's Apperception Test turned the TAT into a structured ego-function assessment and extended apperceptive storytelling to child populations. Wikipedia - Wikidata
Phebe Cramer (1935-2021). Psychologist at Williams College whose Defense Mechanism Manual scored denial, projection, and identification from TAT stories, providing one of the few developmentally validated coding schemes for the instrument. Wikipedia - Wikidata
David C. McClelland (1917-1998). Harvard psychologist who recast TAT stories as a content-analytic measure of implicit motives, building the need-for-achievement scoring system that made apperceptive fantasy a research tool for the science of motivation. Wikipedia - Wikidata
Christiana D. Morgan (1897-1967). Lay psychoanalyst and artist at the Harvard Psychological Clinic who co-created the TAT with Henry Murray and prepared many of the original picture stimuli. Wikipedia - Wikidata
Henry A. Murray (1893-1988). Director of the Harvard Psychological Clinic who co-authored the TAT and embedded it in personology, his theory of psychogenic needs and environmental press, giving the test its interpretive frame. Wikipedia - Wikidata
Michelle B. Stein. Clinical psychologist at Massachusetts General Hospital and Harvard Medical School who co-developed the Social Cognition and Object Relations Scale-Global (SCORS-G), the leading contemporary system for rating TAT narratives, and authored its clinician's guide. Google Scholar
Drew Westen. Psychologist at Emory University whose Social Cognition and Object Relations Scale reframed TAT interpretation around measurable dimensions of object relations and social cognition, seeding the empirical narrative-scoring tradition. Wikipedia - Wikidata
Frequently Asked Questions
What is the Thematic Apperception Test?
It is a projective personality test in which a respondent is shown a series of ambiguous picture cards and asked to tell a dramatic story for each, and the examiner content-analyzes those stories for the motives, conflicts, and relationships the respondent projects into them (Murray, 1943).
How is the TAT scored?
There is no single method. The same cards support several structured systems, each coding a different construct: need for achievement by imagery categories, object relations by anchored ratings, and defense mechanisms along a developmental sequence, so a score is defined only relative to the system used (Cramer, 1999).
What is apperception?
Apperception is perception shaped by the perceiver's own needs, memories, and expectations. In the TAT it is the process by which a respondent reads personal meaning into an underspecified scene, which is why the story is treated as revealing the storyteller (Murray, 1938).
Why do TAT motive scores not match questionnaire scores?
Because they measure different things. A story-based score captures an implicit motive expressed in spontaneous behavior, whereas a questionnaire captures a self-attributed motive a person consciously endorses; the two predict different classes of behavior, so a low correlation between them is expected (Spangler, 1992).
Is the TAT reliable?
Inter-coder reliability is high for the structured systems because their categories are explicit. Internal-consistency reliability looks low, but that reflects the deliberate heterogeneity of the cards rather than true unreliability, and category-based estimates fit the instrument better (Gruber & Kreuzpointner, 2013).
Is the TAT valid?
Validity depends on the scoring system. Structured systems such as need-for-achievement coding and the SCORS-G have predictive and construct-validity evidence, whereas global impressionistic interpretation of the pictures is weakly supported (Lilienfeld et al., 2000).
How is the TAT different from the Rorschach?
The Rorschach scores the perceptual process by which a respondent organizes an inkblot and largely ignores the content, whereas the TAT scores the narrated content of stories told to pictures and largely ignores the perception. One measures how a person sees, the other what a person tells (Murray, 1943).
What is the SCORS-G?
The Social Cognition and Object Relations Scale-Global is an anchored, multidimensional system for rating object relations and social cognition from TAT stories; it is the most active contemporary TAT research program and comes with a detailed manual for clinical and research use (Stein & Slavin-Mulford, 2017).
References
Cramer, P. (1999). Future directions for the Thematic Apperception Test. Journal of Personality Assessment, 72(1), 74-92. https://doi.org/10.1207/s15327752jpa7201_5
Gruber, N., & Kreuzpointner, L. (2013). Measuring the reliability of picture story exercises like the TAT. PLoS ONE, 8(11), e79450. https://doi.org/10.1371/journal.pone.0079450
Keiser, R. E., & Prather, E. N. (1990). What is the TAT? A review of ten years of research. Journal of Personality Assessment, 55(3-4), 800-803. https://doi.org/10.1080/00223891.1990.9674114
Lilienfeld, S. O., Wood, J. M., & Garb, H. N. (2000). The scientific status of projective techniques. Psychological Science in the Public Interest, 1(2), 27-66. https://doi.org/10.1111/1529-1006.002
McClelland, D. C., Atkinson, J. W., Clark, R. A., & Lowell, E. L. (1953). The achievement motive. Appleton-Century-Crofts.
McClelland, D. C., Koestner, R., & Weinberger, J. (1989). How do self-attributed and implicit motives differ? Psychological Review, 96(4), 690-702. https://doi.org/10.1037/0033-295X.96.4.690
Morgan, C. D., & Murray, H. A. (1935). A method for investigating fantasies: The Thematic Apperception Test. Archives of Neurology and Psychiatry, 34(2), 289-306. https://doi.org/10.1001/archneurpsyc.1935.02250200049005
Murray, H. A. (1938). Explorations in personality. Oxford University Press.
Murray, H. A. (1943). Thematic Apperception Test manual. Harvard University Press.
Schönbrodt, F. D., Hagemeyer, B., Brandstätter, V., Czikmantori, T., Gröpel, P., Hennecke, M., Israel, L. S. F., Janson, K. T., Kemper, N., Köllner, M. G., Kopp, P. M., Mojzisch, A., Müller-Hotop, R., Prüfer, J., Quirin, M., Scheidemann, B., Schiestel, L., Schulz-Hardt, S., Sust, L. N. N., Zygar-Hoffmann, C., & Schultheiss, O. C. (2021). Measuring implicit motives with the Picture Story Exercise (PSE): Databases of expert-coded German stories, pictures, and updated picture norms. Journal of Personality Assessment, 103(3), 392-405. https://doi.org/10.1080/00223891.2020.1726936
Siefert, C. J., Stein, M. B., Slavin-Mulford, J., Sinclair, S. J., Haggerty, G., & Blais, M. A. (2016). Estimating the effects of Thematic Apperception Test card content on SCORS-G ratings: Replication with a nonclinical sample. Journal of Personality Assessment, 98(6), 598-607. https://doi.org/10.1080/00223891.2016.1167696
Sinclair, S. J., Carpenter, E. K., Cowie, S., AhnAllen, C. G., & Haggerty, G. (2023). A critical review of the Social Cognition and Object Relations Scale-Global and Thematic Apperception Test in clinical practice and research: Psychometric limitations and ethical implications. Psychological Assessment, 35(9), 778-790. https://doi.org/10.1037/pas0001263
Spangler, W. D. (1992). Validity of questionnaire and TAT measures of need for achievement: Two meta-analyses. Psychological Bulletin, 112(1), 140-154. https://doi.org/10.1037/0033-2909.112.1.140
Stein, M. B., Slavin-Mulford, J., Sinclair, S. J., Siefert, C. J., & Blais, M. A. (2012). Exploring the construct validity of the Social Cognition and Object Relations Scale in a clinical sample. Journal of Personality Assessment, 94(5), 533-540. https://doi.org/10.1080/00223891.2012.668594
Stein, M. B., & Slavin-Mulford, J. (2017). The Social Cognition and Object Relations Scale-Global Rating Method (SCORS-G): A comprehensive guide for clinicians and researchers. Routledge.
Westen, D. (1991). Social cognition and object relations. Psychological Bulletin, 109(3), 429-455. https://doi.org/10.1037/0033-2909.109.3.429
Winter, D. G., John, O. P., Stewart, A. J., Klohnen, E. C., & Duncan, L. E. (1998). Traits and motives: Toward an integration of two traditions in personality research. Psychological Review, 105(2), 230-250. https://doi.org/10.1037/0033-295X.105.2.230