The 13-item Lennox–Wolfe Revised Self-Monitoring Scale describes two reported tendencies: sensitivity to other people’s expressive behavior and perceived ability to adjust one’s self-presentation. It can help frame reflection on differences across situations, but it does not explain why a particular behavior changed. Japanese and Hebrew adaptation findings support the intended two-part account in those studies; a Canadian student study reported a three-factor solution, and a later favorable result concerned a further revision. Read findings in light of the instrument version, sample, and analysis.
What does Lennox–Wolfe mean by self-monitoring?
The 13-item Lennox–Wolfe Revised Self-Monitoring Scale describes two reported tendencies: sensitivity to other people’s expressive behavior and perceived ability to modify one’s self-presentation. In “Revision of the Self-Monitoring Scale,” Lennox and Wolfe use self-monitoring for this bounded combination of attending to social cues and adjusting presentation. A response on the scale is a person’s report about these capacities; it does not record what changed in a particular conversation or establish why it changed.
That scope matters because “revised self-monitoring” can refer to different instruments, and evidence about one version cannot automatically describe another. This article focuses on the 1984 Lennox–Wolfe measure, first explaining why its authors chose a narrower definition, then examining how later analyses and adaptations treated its structure. The distinction keeps the scale’s intended subject clear while allowing the evidence about that subject to be assessed on its own terms.
The phrase should therefore be read as the name of the authors’ measured construct, not as a general synonym for being socially adaptable. Its two-part scope concerns reported attention to expressive behavior and reported capacity to alter presentation; it does not encompass every reason someone may act differently from one setting to another. The scale’s label tells a reader what its authors intended the questions to represent. It does not, on its own, identify the cause, success, or meaning of any specific shift. Those boundaries make it possible to ask a more precise question of the research: do studies of this version reproduce the structure its authors intended?
Why did the 1984 revision narrow the construct?
Lennox and Wolfe’s stated reason for narrowing self-monitoring was a mismatch between the theory and the earlier scale’s structure. In the abstract to “Revision of the Self-Monitoring Scale,” they report that the original Snyder scale had a stable factor structure, but that structure did not correspond to the theory’s proposed five components. In other words, the observed organization of responses did not line up with the broader conceptual breakdown the theory offered. For the authors, retaining every proposed component under one measure would not make the instrument a faithful test of that account.
They also describe a problem with earlier attempts to write items that seemed to represent the theory directly: those face-valid item sets were strongly related to social anxiety. Lennox and Wolfe treated this association as a reason to reconsider what the scale should cover. Their response was to define the construct more narrowly around two areas they intended to assess, rather than continue to represent the wider theory through items whose relationship to another concept complicated interpretation. This is the authors’ rationale as summarized in their abstract; it does not establish that the earlier theory was false, or tell us what any individual’s answers mean.
The distinction is between a broad idea and an instrument’s operational choice: the broad account proposed several components, while this revision selected a more limited scope for its measure. Narrowing can make a scale’s intended content easier to identify and later studies easier to compare against that stated target. It also means the scale cannot stand in for every part of the earlier account. A study finding a different factor pattern therefore bears on how well this particular item set organizes responses; by itself, it does not settle the truth of every broader claim about social behavior.
This history also changes how to read the word “revision.” The change was not simply a new edition intended to preserve the entire earlier theory in a shorter format. As the abstract presents it, the authors were making a judgment about which content could be represented by a more focused instrument after the earlier measurement approach failed to map neatly onto the proposed components and its theory-led items showed a strong association with social anxiety. Their narrower definition is consequently part of the scale’s meaning, not merely a technical detail about item construction. A later study can support or question the organization of this selected content without answering whether the broader theory’s other ideas matter in social life.
The article can therefore keep two questions separate: what the 1984 authors chose to measure, and whether later data reproduce that choice. Keeping them apart prevents a structural result from being mistaken for a verdict on the whole theory.
How do the two reported dimensions differ?
In “Revision of the Self-Monitoring Scale,” Lennox and Wolfe distinguish sensitivity to other people’s expressive behavior from perceived ability to modify one’s own presentation. The first concerns information a person says they attend to in others; the second concerns a change they say they can make in themselves when circumstances call for it. One is oriented toward receiving and interpreting social cues, the other toward adjusting one’s outward expression. They can matter in the same interaction, but they are different actions and should not be collapsed into one description.
That difference is easiest to see by separating the direction of attention. Sensitivity points outward: what changes in another person’s expression, tone, or manner is noticed and considered. Modification points inward: what aspect of one’s own manner might be altered for this audience or setting. Neither dimension logically guarantees the other. Someone could report attending closely to others’ expressive behavior without reporting much inclination or perceived capacity to change their own presentation. Conversely, someone might report being able to alter how they come across without saying that close attention to other people’s cues is what prompts the shift. These are conceptual possibilities, not descriptions of people selected by a score.
Illustrative, not evidence: imagine a person at a project meeting who notices that a colleague’s replies have become brief, then chooses to ask a more direct question. Noticing the change and deciding what it might mean belong to the cue-oriented side of the distinction; changing one’s own approach belongs to the presentation-oriented side. The example cannot tell us whether the colleague was impatient, distracted, or simply short on time. It only shows why attending to someone else and altering one’s own expression are not interchangeable, even when they occur in sequence.
The distinction also explains Lennox and Wolfe’s recommendation about scores. Their abstract reports two subscales as well as a total and advises that all three be treated separately. A total can summarize the instrument’s combined content, but it cannot replace the information conveyed by keeping the two intended areas distinct. If a reader sees only the total, the outward-facing and self-presentation components are less visible; if the subscales are considered, they retain the separation built into the authors’ account. This is a statement about how the authors advised treating their scale’s scores, not a rule for inferring what a particular person is like from a high or low result.
Keeping the labels apart does not mean the tendencies never coexist; it means coexistence should not be assumed from the name of a combined construct. In ordinary interaction, cue attention, interpretation, and self-adjustment may occur close together, but they remain distinguishable questions: what did the person report attending to, and what did they report being able to change?
Separating the dimensions also prevents a common conceptual shortcut: treating adaptation as evidence that a person first detected a social signal. A person may change style because a setting has a formal convention, because they planned to speak differently, or because someone explicitly asked for a change. Likewise, attending to another person’s expression need not produce any outward adjustment. Lennox and Wolfe’s two-part account keeps the reported cue-related tendency apart from the reported capacity to vary presentation, so a description can preserve which part is actually being discussed instead of supplying an assumed cause for the other.
The authors’ conceptual separation is a starting model for describing the instrument, not proof that response data will always divide cleanly into precisely these two parts. Later factor studies address that empirical question. For the present distinction, the useful takeaway is narrower: noticing what another person expresses and changing one’s own presentation are related possibilities in social life, but one does not stand in for the other.
What does sensitivity to expressive behavior describe?
In “Revision of the Self-Monitoring Scale,” Lennox and Wolfe describe the first dimension in terms of sensitivity to other people’s expressive behavior. Its subject is the information a person reports noticing in social expression: a shift in tone, facial movement, pace, or engagement may draw attention. The measure’s label concerns that reported attentiveness. It does not certify that the person noticed every relevant cue, assigned it the right meaning, or knows another person’s internal state.
Those are separate steps. First, something observable may change; second, an observer may notice it; third, the observer may interpret what it means. Even a clear observation does not by itself settle the interpretation, because similar behavior can have more than one explanation. The person who becomes quieter could be reacting to the conversation, concentrating on a detail, or dealing with something unrelated to it. Establishing their actual state would require information beyond the observer’s report of sensitivity. The scale’s dimension speaks to the observer’s reported orientation, not independent confirmation of the whole chain.
Illustrative, not evidence: suppose a listener notices that a friend pauses before answering a question. The pause is the noticed cue; “they may be uncomfortable” is one interpretation. The friend might instead be choosing words carefully or recalling something. Nothing in the act of noticing decides among those explanations. This small distinction matters because people can accurately report that they attend to expression while remaining uncertain about what a particular expression means.
This boundary is useful whenever an observer is tempted to move from “I noticed something” to “I know what it meant.” The former is a report about attention; the latter is a conclusion about another person. A careful description can retain the specific cue and mark the interpretation as tentative, leaving open what further conversation or context might clarify. That is a way to use the dimension’s vocabulary without turning it into a claim of special access to other people’s thoughts.
So the first dimension is best understood as a tendency relevant to social perception: reported attention or sensitivity to expressive cues. It is not a guarantee of perceptiveness, a measure of emotional intelligence, or proof that someone can detect deception. The bounded meaning leaves room for ordinary uncertainty: cues can be ambiguous, interpretations can differ, and another person’s state is not directly available from expression alone.
In practical language, then, “sensitive to expressive behavior” describes where attention is said to go, not whether the resulting judgment is correct. A reader can keep the observation concrete—what changed in the other person’s expression—and hold any explanation lightly until more context is available. That distinction preserves the scale’s intended subject without treating a reported tendency as a reliable reading of any single encounter.
What does ability to modify self-presentation describe?
In “Revision of the Self-Monitoring Scale,” Lennox and Wolfe call the second dimension ability to modify self-presentation. The phrase describes a person’s perceived capacity to alter how they present themselves when a situation calls for a change. It concerns what someone believes they can do: vary tone, formality, expressiveness, or another feature of their outward manner. The label does not, by itself, say what prompted a change or whether a change occurred in any particular encounter. It names a reported capacity within this scale’s intended construct, not a complete account of social behavior.
That distinction separates capacity from motive. A person may change their manner because they want to fit a setting, because a task requires clarity, because another person requested it, or for a reason they have not considered. The subscale name does not choose among those explanations. It also does not settle voluntariness: a person might experience a shift as deliberate, habitual, constrained by a role, or simply easier than behaving another way. The instrument’s description of adjustment capacity leaves these paths open rather than treating one motive as part of the definition.
Illustration, invented: someone may know how to use a more formal style when speaking with a new client than when chatting with a familiar colleague. That contrast illustrates perceived flexibility in presentation. It does not reveal whether the person changed to show respect, follow workplace convention, feel safer, or meet an explicit expectation. Nor does the ability label tell us whether the shift worked: the client might still misunderstand, or the colleague might prefer more formality. Describing the change and explaining it are separate tasks.
The word “ability” can sound like a rating of social competence, but Lennox and Wolfe’s source establishes the name and intended scope of a subscale, not a universal ranking of social skill. Being able to alter one’s presentation is not automatically beneficial in every setting. A context may reward consistency, and a person may reasonably choose not to adjust even when they could. Conversely, a change that seems effective to its author may be read differently by others. The scale’s term does not establish whether adaptation was wise, authentic, successful, or preferable to remaining consistent.
A careful real-life description can therefore stay close to the distinction: “I think I can make my manner more formal when needed” reports perceived capacity. “I did it because I wanted approval,” “I chose freely,” or “it helped” adds claims about motive, control, or outcome that require other evidence. This wording preserves what the dimension is meant to describe while leaving those further questions genuinely open. It also prevents a capacity label from becoming a verdict about how a person ought to act.
Which ‘revised’ scale are you reading?
“Revised Self-Monitoring Scale” is not a precise instrument identifier on its own. The lineage includes Snyder’s original 25-item scale, the distinct 18-item form introduced in “On the nature of self-monitoring: matters of assessment, matters of validity,” and Lennox and Wolfe’s 13-item measure in “Revision of the Self-Monitoring Scale.” The latter two are both described as revised, yet their item sets and measurement histories differ. A paper’s use of the shared phrase therefore does not establish which questions were administered or which findings apply.
The item count helps distinguish them, but should be read with the authors and citation rather than used alone. A reliable identification records the scale’s full or author-linked name, publication year, number of items, and the paper cited for its construction or revision. For this article’s target, the identifying combination is Lennox–Wolfe, 1984, 13 items, and “Revision of the Self-Monitoring Scale.” Snyder and Gangestad’s distinct form is associated with 1986 and 18 items; Snyder’s original is the 25-item version. If an article supplies only “revised self-monitoring,” check its methods and references before carrying over a result.
The Spanish adaptation illustrates why that check matters. “Validity and Reliability of the Spanish Version of the Revised Self-Monitoring Scale” identifies its object as Snyder and Gangestad’s 1986 18-item measure. Its title could be mistaken for a study of Lennox and Wolfe’s 13-item scale if the version were inferred from the phrase alone. The Spanish paper’s results belong to the version it actually adapted; the title’s similarity does not make those results evidence about a different item set. This example is about identifying the instrument, not comparing the adaptation’s findings or quality.
Related scales can share a construct history while differing in item content and proposed structure. That relationship is not enough to assume equal scores, interchangeable subscales, or transferable validation evidence. Each claim must stay attached to the instrument and sample studied. If the method section names one version but a secondary summary uses only the generic label, prefer the method and trace its cited source; if item count or provenance is missing, treat the identity as unresolved rather than filling it in from the word “revised.”
A compact citation check for your own notes is: who developed this form, in what year, with how many items, and which source documents that version? Keep those details beside any finding you quote. This simple record prevents a result for the 18-item Snyder–Gangestad form, for example, from silently becoming a claim about Lennox–Wolfe’s 13-item measure. It does not declare one form superior; it makes clear which instrument a statement describes and what evidence would be needed before comparing versions.
This matters when reading a comparison across papers as well as when identifying one study. If two authors report different factor patterns, first ask whether they used the same version before treating the disagreement as a replication failure. If they used different versions, the difference may concern item selection as well as participants or analysis. That possibility is a reason to check the methods, not a conclusion about what caused the findings to diverge. Clear version labels make the comparison question answerable; a generic scale name does not.
Sources: Validity and Reliability of the Spanish Version of the Revised Self-Monitoring Scale; On the nature of self-monitoring: matters of assessment, matters of validity; An Analysis of the Dimensionality and Reliability of the Lennox and Wolfe Revised Self-Monitoring Scale; Revision of the Self-Monitoring Scale
What did the Japanese adaptation reproduce?
In “A study of revised Self-Monitoring Scale,” the Japanese-language adaptation reported a factor analysis that yielded the two intended factors: sensitivity to others’ expressive behavior and ability to modify self-presentation. The abstract also describes the scale’s internal consistency as acceptable. The result therefore supports a specific point about this adaptation: responses in the studied Japanese-language version organized around the two areas Lennox and Wolfe had set out to measure, and the reported consistency was considered adequate by the study’s authors.
That is a meaningful replication result because factor organization concerns how items relate to one another in the data, not just whether the scale has a recognizable name or translated wording. Here, the reported analysis recovered the intended two-part arrangement. It gives readers evidence that the proposed distinction was not confined to the original English-language account. But the accessible abstract supplies limited detail about the sample, translation procedures, analytic choices, and the numerical consistency estimates. Without those details, we cannot assess how robustly the result was established or compare its strength directly with another adaptation.
Internal consistency and factor structure answer related but different questions. A two-factor result concerns whether the item responses clustered in the intended two areas; an acceptable consistency result concerns whether items within a reported scale showed a sufficiently coherent pattern by the study’s criterion. Neither result alone establishes that the scale captures every aspect of self-monitoring, that the factors are independent, or that respondents behave as their answers imply. The abstract’s summary reports both kinds of evidence, but it does not expose the coefficients or thresholds needed to judge their size against a common benchmark. Keeping those claims separate avoids treating one favorable property as proof of all the others.
The same abstract reports correlations between the adapted scale and other personality measures. Those associations add correlational context: the Japanese study did not examine the factor arrangement in isolation, but also considered how its scores related to other measured personality constructs. The accessible record does not provide enough detail here to reconstruct the full pattern or its magnitude, so the safe conclusion stays at that level. A correlation indicates that scores varied together in this sample; it does not show that one tendency caused another, nor does it tell a reader what a particular person’s responses mean in an interaction.
The finding should therefore travel with its study label. It supports the two intended factors in this Japanese adaptation and reports acceptable internal consistency and correlations with other personality measures. It does not establish that every Japanese version, every Japanese speaker, or every translated item functions equivalently to the original. Country alone cannot explain why a structure appeared, and this one report cannot settle questions about cultural differences. The useful addition is narrower: the intended organization appeared in one adaptation, while the available abstract leaves important design details unavailable for closer appraisal.
Accordingly, the Japanese study supports a local structural finding, not a universal definition of the construct.
Sources: A study of revised Self-Monitoring Scale

What did the Hebrew study add?
“The psychometric properties of the revised self-monitoring scale (RSMS) and the concern for appropriateness scale (CAS) in Hebrew” examined a Hebrew translation across two Israeli samples totaling 1,294 participants, according to the Hebrew University repository record. The study included the Revised Self-Monitoring Scale (RSMS), the Concern for Appropriateness Scale (CAS), and additional measures. Its two-sample design and combined sample make the evidence base broader in scale than a single small adaptation report, although the repository abstract remains the accessible basis for the details summarized here.
For the RSMS, the repository reports that the general total and subscale organization was replicated, with item 12 as an exception. This supports a qualified version of the intended account: the overall pattern was broadly present in those data, but the result was not a perfect item-by-item reproduction. That exception matters because factor claims concern the relations among particular items as well as a broad resemblance to the proposed dimensions. The evidence does not license treating every item as interchangeable or assuming that the same structure would necessarily appear in another translation or population.
A two-sample evaluation can contribute in a way a single pooled statement cannot: the study’s reported total of 1,294 participants was distributed across two samples, giving the researchers an opportunity to examine the proposed account in more than one group within the same evaluation. The repository summary does not give enough detail here to describe the groups’ composition or claim that the result independently generalized to populations beyond them. The cautious advantage is therefore about the study design and breadth of observation as reported, not proof of broad representativeness.
The item-12 qualification also prevents a broad factor label from hiding the level at which evidence differs. A generally reproduced subscale arrangement can coexist with one item that does not behave in the expected way. For readers, that distinction argues against translating “replicated” into “every question works identically.” It also does not show why item 12 was exceptional: the repository record, as summarized in the plan, does not establish whether wording, translation, sample composition, or another feature accounts for it. Any explanation of the exception would go beyond the accessible finding.
The paper also evaluated CAS, but its fit was less satisfactory. CAS is a separate scale, so its result must not be folded into the RSMS finding as if the two instruments were subscales of one measure. The study’s reported support for separate RSMS and CAS analyses answers a measurement question about how each scale behaved in this evaluation; the weaker CAS fit qualifies conclusions about CAS itself. It neither erases the RSMS replication nor supplies extra validation for it. Keeping the instruments distinct is essential to representing what the Hebrew study actually tested.
The contribution is thus two-part: a larger, two-sample Hebrew evaluation generally reproduced the RSMS total and subscale organization while identifying an item-level exception, and it reported a less satisfactory fit for the related but distinct CAS. Those findings make the evidence more informative than a simple statement that “the scale replicated,” because they show where the broad pattern held and where qualification was needed. At the same time, a repository abstract does not provide all the analytic and sample detail needed for an independent close evaluation. The result remains specific to this translation and study; it does not establish universal measurement invariance or resolve how the pattern behaves across languages and settings.
For cross-study reading, this means asking which version, language, sample, and item-level result each paper actually reports before treating similar scale names as interchangeable. The Hebrew report adds a qualified instance to that record, rather than a final ruling on the measure.
Why did the Canadian analysis complicate the two-factor account?
In “The Lennox and Wolfe Revised Self-Monitoring Scale: latent structure and gender invariance,” the confirmatory analysis tested whether responses followed the proposed two-factor arrangement and compared it with alternative models. The sample comprised 836 Canadian university psychology students: 561 women and 275 men. In this dataset, the hypothesized two-factor account was not supported by the confirmatory comparisons. The analyses instead supported a correlated three-factor solution. This matters because confirmatory analysis asks whether an explicitly specified account fits the observed pattern of item responses; a model that is plausible as a conceptual description is not automatically the only arrangement those responses can support.
The three-factor account retained a sensitivity factor and a modified ability factor, while separating a third difficulty-modifying factor. That additional factor consisted of two negatively worded items that otherwise belonged to the ability content. In plain terms, two questions framed in the negative shared enough response pattern to be represented together in this sample, rather than simply joining the rest of the ability items. Thus, the result complicates a clean division in which every item belongs neatly to one of two substantive dimensions. It adds an item-wording-related grouping to the model the researchers found supported; it does not establish why respondents answered those items similarly.
The word “correlated” adds another useful qualification to the Canadian model. The three factors were modeled as related dimensions, rather than as wholly independent compartments. So the supported alternative did not replace one tidy two-part account with three unrelated kinds of person. It separated the item responses into three distinguishable components while allowing those components to covary. For interpretation, that preserves the possibility that sensitivity and presentation adjustment overlap in people’s responses even when the items do not reduce to only two factors. It also shows why factor counts should not be translated directly into a list of fixed personality types: the analysis concerns the structure among responses, not a taxonomy of people.
That distinction is important for interpretation. If a factor includes items because of their wording as well as their subject matter, a factor label may combine content and form. A reader could otherwise assume that every cluster maps directly onto a separate psychological tendency. The Canadian result shows why that assumption needs testing: the two negatively worded ability items formed a distinct component in the reported solution, even though their content was linked to ability to modify presentation. The source supports describing this observed composition. It does not license a further story about confusion, inattention, acquiescence, or any other response process unless evidence directly tests that explanation.
The researchers also examined whether the measure operated similarly for women and men in their sample. They reported broad or partial support for invariance across gender for specified parameters, while some individual item parameters differed. Invariance is not a single all-or-nothing property: evidence that selected parameters are comparable can coexist with differences in others. The result therefore permits a bounded statement about the parameters the analysis supported, alongside recognition that the item-level picture was not uniform. It should not be compressed into a claim that every item functioned identically across the two groups, nor into the opposite claim that no comparison was possible.
The scope of this evidence remains the study’s 836 psychology students at one Canadian university and the models evaluated there. Its three-factor result challenges the idea that the two-factor structure must appear unchanged in every sample; it does not show that sensitivity and adjustment are useless descriptions, that the scale has no descriptive value, or that a three-factor score should replace the published approach for readers. Those are different claims from identifying which model best represented this dataset. The practical contribution is a reason to keep the measurement question open: the intended distinction can be useful as a description while its exact item structure remains dependent on evidence from the population and version being studied.
For the exact question of what the scale describes, this study therefore changes the certainty of the structural claim more than the everyday vocabulary. Sensitivity and adjustment remain the intended content labels, but the Canadian responses warn against assuming that the item architecture always mirrors those two labels without remainder. The evidence points to a model to investigate, not a new reader-facing category.
Sources: Validity and Reliability of the Spanish Version of the Revised Self-Monitoring Scale
What did O’Cass evaluate—and what did it find?
“A psychometric evaluation of a revised version of the Lennox and Wolfe revised self-monitoring scale” evaluated a further revision of the measure. That wording matters: the paper is not simply another test of the unchanged 13-item form discussed as this article’s target. The accessible abstract reports confirmatory factor analysis using data from 450 respondents and says the results showed a broadly similar structure, with improved fit and reliability. These findings are relevant to the scale’s development, but the object evaluated was a revised version. A favorable result for that measure cannot by itself confirm that the original 13 items have the same structure or measurement properties.
The abstract’s statement about fit concerns how the evaluated model corresponded to the data analyzed in that study. Its report of improved reliability concerns the evaluated measurement under that analysis. Both are useful psychometric findings, but neither is a certificate that every interpretation made from a score is warranted, that the model will fit equally well in other groups, or that the measure captures all relevant behavior. The abstract does not provide the coefficients or effect sizes needed here to quantify the improvement, so the comparison should remain at the level it reports: fit and reliability improved for the revision under evaluation.
“Broadly similar” is also a limited structural description, not a claim that the revised items or every parameter matched the earlier form. Without the full report, the abstract does not let us determine exactly which parts of the structure were retained, changed, or compared, so that phrase should not be expanded into a detailed replication claim. A similar arrangement can be informative about the revision’s organization while leaving open how closely its items correspond to the 13-item version. The result belongs to the revised instrument as studied, and the summary does not establish item-by-item equivalence between forms.
The abstract also reports differences or associations involving consumer confidence, subjective knowledge, and concern for image across self-monitoring levels. This places the revised measure in relation to selected consumer variables. The result does not show that self-monitoring caused those differences, establish the direction or size of every relationship from the accessible summary, or predict what an individual reader will do as a consumer. Group-level associations in a study and a particular person’s choices are distinct kinds of evidence. Since only the abstract was accessible, details about the measures, analyses beyond the stated CFA, and how those comparisons were made cannot responsibly be filled in.
This distinction also sets a sensible boundary for comparisons across the papers. A direct claim that one form has better fit than another would require knowing which models and data were compared, along with the relevant estimates. The accessible abstract reports improvement for the further revision, but does not expose those details. It supports reporting the authors’ broad psychometric conclusion, not independently calculating the size of the gain or attributing it to a particular item change. That narrower reading keeps a favorable result in view while preserving the difference between summary evidence and a full technical appraisal.
The sample count is available, but the full article was not accessed. For that reason, this account does not add recruitment details, coefficients, effect sizes, or a more specific description of the revision’s item set than the abstract provides. These are not minor omissions to be guessed from the title: they are details needed to compare the revised instrument precisely with the 13-item target and judge the reported gains quantitatively. Naming the source and its access limit keeps the favorable summary useful without making it sound more complete than the evidence reviewed here.
Read alongside the Canadian analysis, O’Cass’s result answers a different measurement question. The Canadian paper compared structural models for the Lennox–Wolfe scale in a specific student sample; O’Cass reports on a further revised measure in 450 respondents. Their findings need not be ranked as a positive and a negative verdict on one identical instrument. They concern different evidence objects and available levels of detail. O’Cass provides abstract-level evidence that a further revision showed a broadly similar structure and improved fit and reliability, with reported consumer-variable relationships. It does not automatically settle the structure of the unchanged 13-item form, or erase the sample-specific result from the Canadian analysis.
Sources: On the nature of self-monitoring: matters of assessment, matters of validity
How should the mixed structural findings be read together?
Taken together, the findings support a measured conclusion: sensitivity to expressive cues and perceived ability to adjust presentation are useful descriptions of what the Lennox–Wolfe scale was designed to ask about, and that two-part account appears in reported Japanese and Hebrew adaptation results. It is not established as the one latent structure that must emerge unchanged in every study. The Canadian analysis supports a correlated three-factor model in its sample, while the favorable O’Cass result concerns a further revision. These results qualify how confidently we can describe the item structure; they do not require discarding the two intended ideas as ordinary-language descriptions. The strongest competing reading is that a three-factor model makes the two-part account too simple as a description of responses; the adaptation results keep that concern from becoming a reason to discard the intended distinction outright.
The adaptation evidence is supportive in different ways. “A study of revised Self-Monitoring Scale” reports the two intended factors in its Japanese adaptation. “The psychometric properties of the revised self-monitoring scale (RSMS) and the concern for appropriateness scale (CAS) in Hebrew” reports general reproduction of the RSMS organization across two Israeli samples, alongside an exception for item 12. Read together, they show that the intended organization was recoverable in more than one reported adaptation, though not as a claim of flawless item behavior in every case. The details available for those reports also differ, which limits how finely their results can be compared.
“The Lennox and Wolfe Revised Self-Monitoring Scale: latent structure and gender invariance” provides a meaningful complication: confirmatory comparisons in 836 Canadian university students supported a correlated three-factor solution rather than the hypothesized two-factor account. That is evidence against treating the two-factor arrangement as automatic across samples. It is not a direct refutation of the adaptation findings, because a model result from one sample does not tell us which unmeasured feature caused it to differ. Nor does it decide whether the two named content areas remain useful for describing the scale’s intended subject.
“A psychometric evaluation of a revised version of the Lennox and Wolfe revised self-monitoring scale” belongs in a separate place in this comparison. It reports a broadly similar structure and improved fit and reliability for a further revision, based on 450 respondents. Because that paper evaluated a revised version, its result cannot be counted as a direct confirmation of the unchanged 13-item form. The available abstract does not provide enough detail to align its item set and model decisions precisely with the other analyses. Its favorable finding is relevant to the instrument’s development, but it answers a related measurement question rather than resolving the structure of the exact form discussed here.
Part of the apparent disagreement may also come from asking different analytic questions. A study that tests whether a specified two-factor model fits its responses is making a direct comparison between that model and alternatives. A summary that says an adaptation yielded two factors reports a compatible pattern, but without all the same candidate models and fit information it cannot be assumed to have performed an equivalent contest. Likewise, a broad replication statement and an item-level exception can coexist: the overall grouping may resemble the intended account while one response does not join its expected subscale. Those descriptions operate at different levels of detail, so they should not be flattened into simple votes for or against two factors.
Translation is another reason to be precise without speculating. A translated form carries the same named construct and may retain its intended grouping, but the reports summarized here do not provide a matched design that isolates translation from sample composition, administration, or analysis. The Hebrew item-12 exception identifies where that report qualified its result; it does not tell us whether wording caused the exception. Similarly, the Canadian two-item factor identifies a pattern in the responses, not the process that produced it. These are clues about where further comparison would be informative, not evidence for a particular cultural or response-style explanation.
This distinction separates two conclusions that are easy to merge. One is about the construct the authors chose to represent: sensitivity and adjustment remain the scale’s intended paired themes. The other is about how a particular collection of item responses is best organized statistically: the studies do not yield one uniformly repeated answer on that point. The first conclusion comes from the instrument’s stated scope and is echoed by some adaptation findings; the second must remain conditional on version, sample, and analysis. A reader can therefore use the two themes as questions for reflection without treating factor labels as a final map of every person or setting. That is the defensible middle ground: retain the named construct’s practical distinction, while leaving the universal structure claim open to better matched evidence.
The studies are therefore not exact replications whose different answers can be averaged into a single verdict. They involve translations and different populations; their samples are not matched, and the Canadian report uses confirmatory model comparisons while the accessible adaptation summaries provide differing levels of analytic detail. The further O’Cass revision adds possible version differences as well. No matched comparison in the accessible evidence holds the language, population, item wording, version, and model specification constant while changing only one factor. The reason for the differing structural results is consequently unresolved. Assigning it to culture, translation, item wording, or sample composition would exceed what these studies establish.
A practical interpretation can still be careful and useful. Imagine, purely as an invented illustration, that someone says they speak differently in a formal meeting than in a familiar conversation. The observable report is a difference in speaking across those two settings. Their account of noticing others’ expression or adjusting their own presentation is a further self-report about what they attended to and did. Possible explanations remain open: the person may be responding to role expectations, a different topic, familiarity with the people present, or simply having more opportunity to speak in one setting. The contrast alone does not identify which explanation is right.
A modest comparison principle follows: describe what changed before naming the tendency. Then compare occasions that are similar except for one relevant condition, such as familiarity, while noting whether the speaking difference recurs. If it appears repeatedly under comparable circumstances, an interpretation about that condition becomes more plausible than it was after one contrast. That is a reason to take the interpretation seriously, not proof of what caused the behavior; several conditions may still vary together, and the person’s account may not capture every feature of the interaction. The useful result is a more specific description of when a presentation shift seems to occur, rather than a verdict about the person.
Sources: The psychometric properties of the revised self-monitoring scale (RSMS) and the concern for appropriateness scale (CAS) in Hebrew; A psychometric evaluation of a revised version of the Lennox and Wolfe revised self-monitoring scale; Validity and Reliability of the Spanish Version of the Revised Self-Monitoring Scale; On the nature of self-monitoring: matters of assessment, matters of validity
What should you ask next?
After noticing a shift, ask: “What felt different here, and what options did I think I had?” Write down the concrete situation and your answer in a sentence. Treat it as a prompt for reflection on that episode, not a diagnostic or validated procedure; one remembered moment cannot establish a general pattern. If you want to consider how several tendencies may connect, the private Context Profile at [/assessment](/assessment) offers a non-validated reflection tool. For more learning about personality patterns, explore the [personality profile library](/topics). If you are thinking about work, use the question to compare the conditions and demands of a role with experiences you have actually had; a profile cannot choose or predict the right job for you. You decide whether the comparison gives you a useful next question. You might return to the note after another naturally occurring example and see whether the same question still seems useful; there is no need to force a pattern from unlike situations. You can keep the note private and ordinary: what was happening, what you noticed, and what response seemed available at the time. The point is to make your next reflection more specific, not to score yourself.
Questions readers ask
What does the Lennox–Wolfe Revised Self-Monitoring Scale describe?
Its 13 items concern reported sensitivity to other people’s expressive behavior and perceived ability to modify one’s own self-presentation. These are descriptions of reported tendencies, not a record of what happened in a particular interaction.
Does the scale explain why someone changes their presentation?
No. A reported tendency can prompt reflection, but it does not establish the reason for a particular change. Consider the specific situation and what options seemed available before drawing a broader conclusion.
Is the scale’s two-factor structure settled?
The studies summarized here do not establish one structure across every population and version. Japanese and Hebrew adaptation studies reported support for the intended dimensions, while an analysis of 836 Canadian university students supported a correlated three-factor solution. A later study evaluated a further revision.
Sources and notes
- Revision of the Self-Monitoring Scale
Lennox and Wolfe describe a 13-item scale restricted to sensitivity to others’ expressive behavior and ability to modify self-presentation; they report two subscales and a total, and advise treating all three scores separately.
- The Lennox and Wolfe Revised Self-Monitoring Scale: latent structure and gender invariance
Confirmatory factor analyses of responses from 836 Canadian university students did not support the hypothesized two-factor account and supported a correlated three-factor solution, including a factor made from two negatively worded items; gender invariance was mostly, not completely, supported at item-parameter level.
- A study of revised Self-Monitoring Scale
A Japanese-language adaptation of the Lennox–Wolfe instrument was factor-analyzed; the abstract reports the two intended factors and acceptable internal consistency, alongside correlations with other personality measures.
- The psychometric properties of the revised self-monitoring scale (RSMS) and the concern for appropriateness scale (CAS) in Hebrew
A Hebrew translation was examined in two Israeli samples totaling 1,294 participants; the two-subscale structure generally replicated, except for RSMS item 12, and confirmatory analyses supported separate RSMS and CAS scales.
- A psychometric evaluation of a revised version of the Lennox and Wolfe revised self-monitoring scale
O’Cass used confirmatory factor analysis on data from 450 respondents for a further revision and reported similar structure, improved reliability and fit, and differences in consumer confidence, subjective knowledge and concern for image across self-monitoring levels.
- Validity and Reliability of the Spanish Version of the Revised Self-Monitoring Scale
The Spanish adaptation studied Snyder and Gangestad’s 1986 Revised Self-Monitoring Scale, a separate 18-item instrument; its reported structure and findings must not be attributed to Lennox–Wolfe’s 13-item scale.
- On the nature of self-monitoring: matters of assessment, matters of validity
Snyder and Gangestad’s 1986 paper discusses self-monitoring assessment and validity and presents a new 18-item scale, establishing a separate instrument lineage from the Lennox–Wolfe measure.
- An Analysis of the Dimensionality and Reliability of the Lennox and Wolfe Revised Self-Monitoring Scale
The 1990 article is an additional published dimensionality and reliability investigation of Lennox–Wolfe’s measure; its reported abstract notes minimal overlap with Snyder’s 25-item original and 18-item revised form, reinforcing that these measures should not be conflated.
Apply it to your own pattern
Compare your tendencies with the work setting
From this guide: Reflect on when you adjust your communication and what role conditions make that easier or harder.
The Lennox–Wolfe scale describes two reported tendencies, while work also brings expectations about pace, communication, authority, and recovery. If you want to consider how this pattern sits beside other everyday tendencies, the private Context Profile offers continuums and a descriptive work-style lens for reflection. Use it to form questions about role conditions and your own examples; it does not select a career or predict fit.
