The Wong and Law Emotional Intelligence Scale (WLEIS) is a 16-item self-report measure of perceived emotional tendencies across four factors. At work, its factor scores can help you choose a behavior to examine in a recurring interaction; they do not verify what happened or predict your individual job performance.
What does a WLEIS response actually represent?
After a difficult feedback exchange, a worker may want a result that settles whether they handled it well. The Wong and Law Emotional Intelligence Scale (WLEIS) cannot replay that conversation. In “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample,” the instrument is described as a 16-item self-report scale, with four items for each of four factors. The respondent reports how they see their emotional tendencies; the answers are not an observer’s record of what happened in a particular meeting.
Imagine, as an illustration rather than a reported case, that a reviewer says a handoff was unclear and the worker leaves unsure whether they listened or became defensive. A WLEIS result can give them a structured prompt for reflection on emotional habits. It cannot determine whether the reviewer's example was accurate, whether the worker interrupted, or what either person intended. Those questions call for evidence from the exchange itself: the feedback, the work, and, where appropriate, a specific account of what each person observed. The scale also does not supply the missing context: what standard the reviewer used, which part of the handoff caused concern, or whether the worker had a chance to clarify it. Those details belong to the work conversation, not to the questionnaire response. The scale also does not supply the missing context: what standard the reviewer used, which part of the handoff caused concern, or whether the worker had a chance to clarify it. Those details belong to the work conversation, not to the questionnaire response.
The scale’s authors grounded it in an ability model, and some discussions classify WLEIS in ability terms. That theoretical lineage does not change what a respondent does when completing this questionnaire: they report about themselves. For the worker deciding what to do with the result, the useful distinction is between a perception to examine and a documented behavior to verify. The four factors can focus that reflection; they do not turn the score into a verdict on the exchange.
What does self-emotion appraisal help a worker notice?
Self-emotion appraisal is the WLEIS factor about recognizing and understanding one's own emotions. In the feedback illustration, the worker might notice tension before deciding what the reviewer's comment means. They could name defensiveness, embarrassment, or irritation as the experience they are having. This is an inward observation: it gives the worker a clearer account of their present state, not yet an explanation of its cause.
That distinction matters because a feeling and the interpretation attached to it are different kinds of information. The worker may genuinely feel embarrassed; the feeling is not imaginary simply because its source is uncertain. But “I feel exposed” does not by itself show that the reviewer intended to humiliate them, that the criticism was unfair, or that the work was poor. The emotion marks something worth attending to. The account of why it arose remains a question that can be checked against the comment, the task, and what was said in the conversation.
The original WLEIS study, “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample,” treats self-emotion appraisal as one of four distinct factors and describes the scale as self-report. A rating on this factor therefore tells the worker how they appraise their own capacity to recognize and understand emotion. It does not independently test whether a particular label fits a particular moment. In practice, the score can point toward a useful line of inquiry: when feedback lands, does the worker notice an emotional response clearly enough to describe it, or do they mainly register a broad sense that something feels wrong? That question is a prompt for reflection, not a conclusion supplied by the scale.
Even a worker who is attentive to emotion may use an incomplete label. “Angry” might cover disappointment, worry about consequences, or a sense of being misunderstood; the exact wording of feedback and the setting can shape what stands out. A factor score cannot establish which of those explanations is right. That does not erase the value of noticing the feeling. It means the worker should keep the observation close to what they can report—“I felt tense when the handoff was called unclear”—and leave the cause open until the surrounding evidence gives them reason to be more specific. A worker who jumps from a bodily signal to a confident account of another persons motive can use that jump as a cue to slow the interpretation down. The score cannot resolve the gap; a description of the moment and the available conversation can help keep the two claims distinct.
For this example, self-emotion appraisal contributes a small but important piece of information: the worker can separate their immediate experience from the judgment they are forming about the other person or the work. They might record the feeling and the exact comment that preceded it, then examine those separately. The article is not claiming that WLEIS tested this response or that naming an emotion improves the conversation. It shows how the factor’s focus on awareness can make a reflection more precise: first identify what is being felt; then treat the explanation as something to investigate.
What can a worker infer from appraisal of other people's emotions?
In the WLEIS, others' emotion appraisal concerns a respondent's perceived ability to recognize and understand emotions in other people. “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample” identifies it as one of the scale's four self-reported factors. That makes it a prompt about how a person sees this capacity; it is not a test in which a reviewer’s feelings are independently observed and scored. A worker may use the factor to ask what they notice in an exchange, but the score cannot certify that their reading of a particular colleague is accurate.
Return to the feedback exchange: a reviewer says the handoff was unclear. The worker asks which section needs attention, and the reviewer gives a shorter answer than before, with a flatter tone. The change is a cue the worker can notice. It might mean the reviewer is frustrated, pressed for time, concentrating on the task, or simply speaking briefly. Even if frustration seems plausible, the cue does not tell the worker what caused it or whether the reviewer’s concern is about the work, the process, or something else. A shift in tone is information about the interaction as experienced by the worker, not direct access to the other person's private state.
That gap matters because people often act on an interpretation as though it were an observed fact. If the worker concludes, “They think I am careless,” they may apologize for a motive that was never expressed, defend themselves against an accusation that was never made, or stop asking about the handoff. None of those moves answers the practical question raised by the original feedback: what, specifically, should change? Appraisal can help the worker register that the exchange feels different. It cannot supply the reviewer’s explanation or convert a possible emotional cue into proof of intent.
A more grounded next move keeps the request about the work open. After listening, the worker might say, “Could you point me to the part of the handoff that needs the most attention?” Or, if the moment feels tense, “Would it be all right if I check one detail about the feedback?” The question asks for an example or permission to clarify; it does not require the worker to announce a diagnosis of the reviewer’s mood. The answer can then be compared with the actual handoff and the standard being applied. If the reviewer names a missing decision or unclear owner, that gives the worker something concrete to inspect. If the answer remains vague, the uncertainty remains; a guessed emotion does not fill it.
This is a practical illustration of how a worker might use the construct, not evidence that these exact words improve feedback conversations. Timing and power matter. A direct question can sound challenging when someone is rushed, when the exchange is already heated, or when the worker has little room to disagree safely. Listening without interruption may be the first useful step. The worker can ask whether the reviewer has time to clarify, take a pause, or return to the issue later. Asking permission does not make every conversation comfortable, but it gives the other person a chance to engage rather than treating the worker’s interpretation as settled.
The useful distinction is between noticing a cue and knowing what it means. The factor can focus reflection on whether a worker tends to attend to others’ expressed emotion. In the moment, however, a changed tone should invite curiosity about the task and the conversation, not certainty about motive. The reviewer’s concrete example, the work itself, and any explanation they choose to give are better grounds for deciding what needs attention.
When can use of emotion help a worker act on feedback?
The WLEIS use-of-emotion factor concerns a respondent’s perceived ability to draw on emotion in pursuing goals. The Colombian manager study, “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample,” describes this as one of four self-report dimensions. In practical terms, the factor invites a worker to consider how they believe emotion can direct effort or keep a goal in view. It does not establish that a feeling has correctly identified the problem, that the chosen goal is attainable, or that greater intensity will produce a better decision.
Consider the same feedback about an unclear handoff. The worker feels disappointed and urgently wants to show that the work was done properly. That urgency may focus attention: which deadline was at risk, what information did the next person need, and where did the handoff leave a gap? Disappointment may also point to a concern about fairness or recognition. Those reactions can help identify what the worker wants to examine. They do not settle whether the handoff met the agreed standard, whether the review was fair, or whether the reviewer misunderstood the work. Emotion can direct attention toward a question; evidence from the task must help answer it.
A useful sequence is to name the concern in plain terms, then inspect what can be checked. Was a date missed? Was an owner unclear? Did the receiving colleague lack a decision or file? If the task record shows a real omission, the worker can revise the handoff and say what changed. If the record seems complete but the reviewer points to a specific ambiguity, the worker can ask what wording or detail would have made the transfer usable. If the disagreement is instead about how the feedback was delivered or how responsibilities were assigned, that process issue may need its own conversation. The feeling helps distinguish what deserves attention; it cannot decide which of these explanations is right.
This distinction also helps with motivation. A worker may draw on concern about a deadline to begin a revision, but effort only helps if the work can be changed and the worker has enough time, authority, and information to do it. Where the goal is outside their control, trying to intensify motivation may leave the actual obstacle untouched. The next step could be to request a missing decision, renegotiate a deadline, or make the dependency visible. Those actions address the conditions around the task. A self-rating about using emotion does not show that workload, resources, or organizational authority are adequate.
Nor should the worker treat a strong feeling as proof that the first interpretation is true. Feeling indignant does not by itself establish unfair treatment; feeling anxious does not demonstrate that the work will fail. A feeling can still matter as a signal that the worker sees a risk, unmet expectation, or value at stake. The distinction is between taking that signal seriously enough to investigate and treating it as a verdict. Looking at the feedback alongside the task record can reveal whether the emotional concern corresponds to a concrete work issue, a process concern, or both.
In the illustration, emotion might give the worker energy to make the handoff clearer or persistence to ask for the example they need. The sequence—identify the concern, inspect the work, then choose whether to revise, clarify, or raise a process issue—is an editorial application of the factor, not an effect tested by WLEIS. It shows what a worker could do with a reflection prompt without claiming that a score creates motivation or improves judgment. A next step still depends on a feasible goal and the conditions for pursuing it.
Does regulation mean staying calm?
Regulation in the Wong and Law Emotional Intelligence Scale (WLEIS) is the respondent’s perceived ability to manage their emotional response. It points toward a question about choice under pressure: when criticism stings, can the worker notice the reaction and decide what to do with it? That is different from appearing calm. A person can hide irritation while remaining consumed by it, and a person can show concern or firmness while still choosing their words and next step. In “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample,” WLEIS is treated as a self-report measure. A regulation rating therefore describes how the respondent sees this capacity; it cannot establish what they did in one meeting.
Suppose a reviewer says a handoff was unclear, and the worker feels an immediate urge to interrupt: the documents were sent on time, and the criticism seems unfair. A brief pause could create room to decide whether to ask which part failed, explain what was sent, or request time to check the record. The pause is useful because it preserves options, not because silence or composure is always the right outcome. If the worker answers at once, the conversation may shift to defending intent before either person has named the handoff problem. If they wait long enough to identify what they need to know, they can still disagree while keeping attention on the claim.
Suppressing frustration for the sake of smoothness has a different purpose. A worker might say “That’s fine” while believing the delivery was belittling, then leave without clarifying either the work issue or the impact of the exchange. Nothing in that performance of calm shows that the reaction has been managed. The worker could instead say, “I want to understand the concern. Could we look at the part that was unclear?” If criticism becomes personal or disrespectful, they may name that boundary: “I can discuss the handoff, but I need us to keep the feedback focused on the work.” These are reasoned examples of response selection, not scripts tested by WLEIS.
Regulation also includes choosing not to answer immediately. If the worker is too activated to listen or the standard is unclear, “I need a little time to review the handoff; can we return to this later today?” protects accuracy and leaves the issue open. The other person’s role and the meeting conditions matter: a junior employee may have little power to redirect a senior reviewer, and a rushed or public exchange can narrow the available choices. A score cannot tell whether a boundary is safe to state in a particular workplace. Where a response could carry a cost, documenting the concern or finding a more appropriate setting may be more realistic than insisting on a direct confrontation.
That is why regulation should not become a demand that a worker absorb poor treatment. The factor can focus reflection on perceived influence over one’s response; it does not say that every visible emotion is a failure or that agreement is evidence of skill. In the handoff discussion, the worker might remain frustrated, challenge the reviewer’s account, or ask for a pause. The relevant distinction is whether they can choose a response that serves the issue they need to address, given the power and conditions in front of them.
How do the four factors interact in one difficult conversation?
The four factors are more useful together because a difficult exchange does not arrive sorted into separate emotional tasks. Consider this invented illustration, not a reported case: in a project review, a colleague says, “Your handoff left us guessing about the next step.” The worker feels a flush of irritation and thinks, “I sent everything yesterday.” A single factor cannot explain whether the exchange is about missing information, a strained relationship, unclear ownership, or a delivery problem. The worker has to move from their own reaction toward a response while the meaning of the comment is still uncertain.
First, self-emotion appraisal gives the worker a chance to notice the irritation rather than treat it as the whole story. They might recognize that the criticism touches a concern about being seen as unreliable. That observation is about their own experience. It does not show that the colleague intended to insult them. If the worker skips this distinction, they may answer the presumed insult and never discover what was missing from the handoff.
Next, others’ emotion appraisal can help the worker attend to the colleague’s words and manner. Perhaps the colleague sounds pressed and points to a specific gap; perhaps the remark is sharp but gives no example. Either cue can inform the worker’s tentative read, but neither proves motive. A concrete reference to a missing decision would change the likely explanation more than a guessed mood would. If the standard for a complete handoff was never agreed, even accurate attention to both people’s feelings cannot resolve that process problem. The worker may need to ask what information the recipient expected.
Their irritation can also reveal what matters to them: accuracy, recognition for work already completed, or a clear division of responsibility. This is the use-of-emotion idea in practical form. The feeling helps select a question worth asking; it cannot decide whether the handoff was adequate. Looking at the message sent yesterday, the worker may find that the file was included but the next owner was not named. Or the record may show that the recipient had the needed details and the disagreement concerns a different expectation. Those possibilities lead to different conversations, so checking the task matters more than defending an assumption.
Then regulation shapes how the worker responds. They might say, “I sent the file yesterday, but I may have left the next owner unclear. Which step were you expecting me to name?” This answer holds both the fact they remember and a genuine opening to learn. If the colleague’s delivery crossed a line, the worker could instead pause the work discussion long enough to say that they want to address the handoff while keeping the exchange respectful. Either response remains possible; the choice depends on what the colleague said, the work evidence, and the worker’s room to speak.
This sequence—notice one’s feeling, form a provisional reading of the other person, identify what the feeling brings into focus, and choose a response—is an editorial way to connect the factors. It is not a causal pathway that WLEIS has validated. Time pressure, seniority, safety, or an ambiguous handoff standard could change what the worker can reasonably infer or say. Reading a single factor as the explanation would hide those dependencies; reading the sequence as a prompt keeps the score subordinate to the actual exchange.

Why can the total score and four factors tell different stories?
A four-factor profile and a single total answer different questions. Separate factor scores preserve distinctions among the scale’s domains, which can help a worker choose a narrower subject for reflection. A total score compresses those distinctions into one summary. Whether that compression is defensible depends on evidence that the factors also reflect a broader common dimension in the particular version being used. Factor labels alone do not settle the issue.
In “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample,” both a correlated four-factor model and a higher-order model showed favorable fit. In the first, the four dimensions are related but remain distinguishable. In the second, their relationships can also be represented through a broader factor above them. For that Bogotá manager sample, the findings supported examining the domains separately and offered support for a general score as well. The paper also notes that findings elsewhere have not consistently supported a general factor, so its result is evidence about this sample rather than a universal scoring rule.
The Peruvian study, “Escala de Inteligencia Emocional de Wong-Law (WLEIS-S): propiedades psicométricas y datos normativos en población adulta peruana,” favored a bifactor model. This structure allows a broad common factor to account for responses while retaining factor-specific variance. In practical terms, it asks whether the overall pattern and the four narrower patterns each contribute information. That model is not simply a competing name for the Colombian higher-order result: the structures allocate shared and domain-specific variation differently. A score report’s decision to show both kinds of result therefore needs support from the structure tested for that version.
Fit statistics summarize how well a proposed structure reproduces relationships among answers in a dataset. They do not reveal which behavior produced a worker’s score, whether a colleague experiences that behavior, or what happened in a particular disagreement. Stronger fit for one structure is a reason to examine how the version’s scores are interpreted; it is not a description of the person behind them. For an individual using a profile to reflect, four scores can preserve useful distinctions. Combining them may be reasonable where the tested model and scoring guidance support a general factor, but the total necessarily gives up some detail. A total can make a concise overview easier to communicate, yet it may conceal a profile in which one domain differs from the others. That difference can matter if the aim is to decide what to notice in a future interaction. Conversely, treating four scores as wholly independent would also overstate what the factor labels mean when the domains are related. The practical choice is between a supported summary and supported detail, not between one true score and four unrelated traits.
The Colombian and Peruvian results do not crown one country’s structure as correct. The studies examined different versions and samples, and both used non-probability recruitment. Their contrast leaves several possibilities open: sample composition, wording or adaptation, and analytic choices may affect which structure fits. A worker comparing two reports should therefore ask which version produced each score and what factor model its evidence supports. If the answer is unclear, the four-domain profile is a more transparent prompt for reflection than assuming an overall number has a settled meaning. Neither presentation identifies a cause of workplace conduct; that requires examples of what the worker actually did and how the situation unfolded. A report that supplies only a total leaves the reader unable to see whether its interpretation depends on a general factor or simply adds together domains for convenience. A report that supplies subscales should likewise explain why those distinctions are reported and whether a broad score is also supported. The score architecture shapes the question a worker can sensibly ask next: an overall result invites broad reflection, while a differentiated profile may help select a particular area to examine. Neither should be treated as a record of what happened at work. To combine scores responsibly, a user needs more than an attractive overall summary: the scoring rule should follow evidence for that version, and the report should make clear what information the aggregation removes.
Sources: Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample; Escala de Inteligencia Emocional de Wong-Law (WLEIS-S): propiedades psicométricas y datos normativos en población adulta peruana
When can a worker trust a translated or adapted version?
A translated WLEIS can be useful for reflection when researchers have examined that version in a relevant group. That evidence answers a local question: do the translated items and scores behave plausibly among the people studied? It does not automatically answer a broader one: can scores be compared fairly across languages, occupations, or demographic groups? Before drawing either conclusion, the worker needs to know which version was administered and what claims its validation actually supports.
“Validity and Reliability of the Korean version of the Wong and Law Emotional Intelligence Scale for Nurses” reports a translation and validation study with nurses recruited from two South Korean hospitals. The study provides evidence about that Korean version in that occupational setting. Its authors also identify the absence of a detailed measurement-invariance analysis as a limitation. A nurse using the version for personal reflection may still find the result useful; someone comparing it with an English-language report cannot infer equivalence from the local validation alone. The sample does not represent every Korean speaker or every worker.
Comparability requires a more specific test. Measurement invariance examines whether a measure operates similarly across groups named in an analysis; depending on the level established, researchers may be able to compare relationships among scores or their levels. The Peruvian WLEIS-S study reports invariance evidence by sex and age group within its study, as well as percentile norms for adults in metropolitan Lima. Those results support interpretation against those stated reference groups and norms. They do not show that Korean nurses, English-speaking managers, and Lima adults share a common scale or can be ranked on the same yardstick. This distinction matters when a report presents a percentile as if it were a portable position. A percentile describes standing relative to a reference group under the scoring rules used; it is not an inherent amount of emotional skill. Even a carefully constructed local norm answers only the comparison it was built to answer.
A worker looking at a translated report can make the decision concrete with four questions: Which language version did I answer? Was it validated with people resembling the group I want to compare myself with? Does the report identify the population behind its norms, if it gives a percentile or rank? Was this version shortened or otherwise adapted from the one studied? A favorable result on one version does not transfer automatically to a shortened form, and a percentile has little meaning without its reference group. These questions determine whether a result is best used for private reflection or whether a particular comparison is supported.
The wording of an adaptation matters because translation is not just replacing each word. A phrase about managing one’s emotions may carry different everyday associations across languages or occupations; reviewers need to check whether respondents understand the intended question. Shortening can also remove content or change the balance among domains. Those are reasons to look for documentation of the actual form administered, not reasons to assume every translation is defective. A report that names the version and its studied population gives the reader a way to judge whether the evidence is close to their purpose.
For personal use, a locally studied translation can offer a structured prompt even when cross-language equivalence has not been established. For comparison, the bar is higher: the evidence must cover the specific groups, version, and score interpretation involved. The Korean nurses study and the Peruvian adult norms each answer bounded questions about their own contexts; neither supplies a universal reference for workers. If a report cannot name its language adaptation or comparison group, treat its factor descriptions as prompts to investigate behavior rather than as a basis for comparing yourself with people elsewhere. Where the report gives no norm group, a percentile or rank should not be read as a meaningful standing among workers.
Sources: Validity and Reliability of the Korean version of the Wong and Law Emotional Intelligence Scale for Nurses; Escala de Inteligencia Emocional de Wong-Law (WLEIS-S): propiedades psicométricas y datos normativos en población adulta peruana
How does self-report differ from an ability test?
For a worker asking whether a WLEIS result reflects their habits or their ability to handle an emotion problem, the response format is the first clue. The Wong and Law scale asks respondents to endorse statements about how they generally perceive their emotional tendencies. In “Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample,” WLEIS is described as a 16-item self-report instrument. A worker answers about themselves; they do not have to solve an emotion task during the assessment. The resulting pattern therefore begins with self-perception: what the respondent believes they usually notice, understand, regulate, or do.
An ability test makes a different demand. The respondent performs tasks designed to sample emotional problem solving, and a scoring method determines how responses are evaluated. The score concerns performance on those sampled tasks under those instructions. The distinction is about what the person does to produce a result, not whether emotions matter at work. A self-report can be a useful entry to reflection on perceived tendencies; a task measure can address performance on its particular problems. Neither response format, by itself, observes how someone handles every meeting, deadline, disagreement, or repair. A task score is therefore evidence about performance within a defined assessment, not a direct recording of conduct in an unobserved conversation. To carry it into development, the worker would still need to ask whether the task resembles a challenge they meet and what evidence from their actual work supports that connection.
The label attached to a scale cannot settle that question. The Colombian validation paper describes a theoretical debate about WLEIS: its roots are associated with ability-based emotional intelligence, while its actual items ask people to rate themselves. The paper discusses that ancestry and debate, but the respondent still supplies self-descriptions rather than answers to scored emotion problems. For interpretation, the act required by the item matters more than a model’s family tree. Calling WLEIS an ability measure without explaining its response format risks making a worker think the score records demonstrated task performance when it records endorsed statements.
The same check applies when comparing named measures. “The relation between emotional intelligence and job performance: A meta-analysis,” available here through its publisher abstract, distinguishes ability, self/peer-report, and mixed measurement streams. That broad comparison supports treating response demands as meaningfully different. The abstract alone does not justify a detailed claim about how particular instruments score answers, which one is superior, or what WLEIS predicts. Those questions require evidence about the specific measure and version, not a borrowed conclusion from a broad category.
A practical choice follows from the question. If the worker wants a starting point for examining perceived patterns—perhaps whether they tend to notice frustration before replying—a self-report such as WLEIS can organize that reflection. The next step is to compare the perception with an actual example, rather than treat endorsement as proof of what happened. If the question is how someone performs on emotion problems, look for a documented ability measure that specifies its tasks and scoring approach. Its result answers that narrower performance question; it does not establish how the person behaves across the full range of workplace conditions.
Ability performance remains bounded by the tasks selected, their language and context, and the scoring rules applied. A person may do well on a particular assessment yet face time pressure, power differences, unclear expectations, or a relationship history that the test did not reproduce. Conversely, self-report can reveal a respondent’s own account while missing behavior they do not notice or recall. These are consequences of the methods’ different demands, not grounds for collapsing them into one universal EQ result. When the decision is about employment, neither format alone provides a sufficient basis to rank or judge a worker; each leaves out evidence about the real work and circumstances.
For individual development, start with the uncertainty you actually want to reduce. “What do I think I usually do when feedback stings?” points toward self-report plus a concrete episode to examine. “How do I perform on these emotion-recognition or reasoning tasks under this scoring method?” points toward an ability assessment with accessible documentation. The first is a question about perceived habits; the second is a question about sampled task performance. Keeping those questions apart helps a WLEIS profile do its useful job: suggest where a worker might look more closely, without quietly changing what the instrument asked them to do.
Sources: Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample; The relation between emotional intelligence and job performance: A meta-analysis
What do workplace outcome studies add—and what can they not predict?
Workplace research adds two different kinds of evidence: studies asking whether EI measures are associated with work outcomes, and studies asking whether training is followed by measured change. Both can inform a worker’s expectations, but neither directly answers whether a particular WLEIS factor predicts what one person will do next. The distinction matters because an association across employees and an average change after an intervention arise from different designs and support different conclusions.
“A Meta-Analysis of the Relationships Between Emotional Intelligence and Employee Outcomes” reports a positive pooled association between self-report EI and job performance: its table includes 31 samples and 10,438 people, with a corrected correlation of rho = .33. Across EI streams, the review reports an overall job-performance estimate of rho = .30 from 68 samples and 23,269 people. Those are broad syntheses across measures and employee samples. The self-report subgroup is not a WLEIS-only estimate, and neither pooled figure says that a specific worker’s WLEIS factor will forecast their performance.
The review’s breadth is useful precisely because it gives a group-level view across multiple studies, while also limiting how narrowly the result can be applied. Its included research varies in instruments, performance criteria, and study designs; some outcomes rely on self-report. The evidence covers English-language quantitative studies published from 1990 through early 2020, and the review notes that contextual moderators were unavailable. A pooled relationship can summarize a recurring pattern across that evidence base without telling us what caused the association or whether it holds in a given role, team, or individual case.
A worker can therefore read the finding as context: emotional-intelligence measures have shown positive relationships with employee outcomes in aggregate. It is not a personal forecast. A favorable WLEIS response does not guarantee strong performance, and a lower response does not establish that someone will struggle. The meta-analysis did not test whether changing a WLEIS score changes work outcomes. Other features of the role, the opportunity to use a skill, the quality of a process, or the way performance is rated may also shape the relationship observed across studies. The data support an association at a broad level; choosing among these explanations for an individual would require evidence the pooled estimate does not provide.
The workplace training evidence asks a separate question. “Training emotional competencies at the workplace: a systematic review and metaanalysis” synthesized 50 pre-post samples involving 3,233 participants and a controlled subset of 27 studies. In the controlled comparison, the review estimated an average standardized mean difference of .47. That estimate concerns measured change across diverse workplace emotional-competency interventions, not a WLEIS-guided program or a specific practice prescribed by one factor score. It gives a reason to consider development plausible across the interventions studied, without converting that average into a promise about one employee.
The distribution around the average is central to that interpretation. The review reports high heterogeneity in pre-post results and a prediction interval for the controlled effect that crosses zero. Its risk-of-bias assessment classified 18 studies as critical and 28 as serious; only six were randomized trials. Interventions, target competencies, and outcome measures varied, most measures were self-report, and follow-up was limited. Those features leave considerable uncertainty about what would happen in a new setting, which worker might benefit, and whether a measured change would carry into everyday work behavior.
The controlled and pre-post results also answer different counterfactuals. A pre-post comparison asks whether participants’ measured scores changed between the start and end of a program; without a comparison group, changes may reflect other events or measurement conditions. A controlled comparison asks how outcomes differ between intervention and comparison conditions, but its estimate still pools the interventions and studies that met the review’s criteria. The prediction interval crossing zero is a reminder that the average cannot be assumed to describe every setting represented by a future study. Since only a small number of included studies were randomized, the overall estimate should not be read as a clean causal effect that can be assigned to any one workshop or emotional skill. The review’s results are relevant to whether workplace development merits investigation; they do not identify the best curriculum for a WLEIS profile.
The two reviews can both be true without answering the same question. A positive average association across employees does not show that an intervention causes better job performance. An average measured improvement after varied training does not establish that WLEIS identifies who will improve, or that its four factors prescribe effective exercises. One body of evidence describes relationships among measures and outcomes; the other evaluates change after different interventions. Combining their estimates into a single claim about WLEIS would erase the difference in design and overstate what either review tested.
For an individual worker, the sensible use is modest and concrete. Treat a score as a prompt for selecting one behavior to observe, then compare it with specific examples over time: what was said after feedback, whether a pause changed the reply, or whether a repair conversation clarified the impact. This is a practical way to learn from one’s own work, not an intervention effect established by these reviews. If the examples change, the worker has something more relevant to discuss than a score alone; if they do not, the score did not guarantee improvement. The research supports taking emotional skills and development seriously while leaving individual outcomes open to observation.
A behavior check can be made specific enough to learn from without pretending to be a research trial. Before a difficult exchange, the worker might name the action they want to notice, such as asking for one example before defending a decision. Afterward, they can record what the other person said, what they did, and whether the exchange clarified the issue. Repeating that observation across situations may reveal whether the behavior appears under different pressures, or whether a particular process or relationship changes what is possible. One successful conversation cannot establish durable change, and a missed attempt does not diagnose an inability. The value is narrower: concrete episodes can test whether a self-description is recognizable in the worker’s current practice. That makes the result useful as a starting question while keeping the outcome tied to what happened, rather than to an average drawn from different people, measures, and programs.
Sources: Training emotional competencies at the workplace: a systematic review and metaanalysis; A Meta-Analysis of the Relationships Between Emotional Intelligence and Employee Outcomes
What is one useful next conversation?
Ask a trusted colleague or reviewer for one concrete example of the behavior you want to understand. A useful prompt is: “When have you seen me notice someone’s concern before responding?” or “Can you recall a time I asked for clarification after difficult feedback?” Keep the question about something observable. You are seeking a description of an exchange, not agreement with a factor score or a judgment about your emotional intelligence.
Listen for what the person saw or heard, then compare that account with your own memory. If they describe several moments, ask which single example best shows the behavior you are trying to understand, and what happened just before and after it. Their perspective comes from particular conversations and circumstances; treat it as one example to examine, not an unquestionable rating. If it points to a moment worth revisiting, choose one small behavior to notice next time, such as asking a follow-up question before explaining your intent. Make the cue specific enough to recognize in the exchange: what you asked, paused to hear, or said in reply. If discussing work with another person feels inappropriate, use a private note about a recent exchange and what you might try in a similar one.
For a separate private reflection tool, the Emotional Skills Profile offers 32 questions and an educational report that keeps recent-behavior summaries distinct from judgments on authored scenarios. It is a development prompt, not a WLEIS form or a substitute for one. [Explore the Emotional Skills Profile](/assessment), or [view the report information](/report).
Questions readers ask
Is the Wong and Law Emotional Intelligence Scale an ability test?
WLEIS asks people to rate their perceived emotional tendencies. Although its theoretical roots are associated with ability-based emotional intelligence, its self-report format does not test performance on emotion problems. The Colombian manager validation paper describes this distinction and the debate around classification.
Sources
- Psychometric Properties of the Wong and Law Emotional Intelligence Scale in a Colombian Manager Sample
Defines WLEIS as a 16-item self-report scale with four four-item factors; tests four-related-factor and higher-order models in Colombian managers and discusses debate over interpreting WLEIS as ability versus trait-like self-perception.
- Validity and Reliability of the Korean version of the Wong and Law Emotional Intelligence Scale for Nurses
Reports translation and validation of a Korean WLEIS version with 210 nurses from two South Korean hospitals; abstract reports content validity index .90, adequate construct validity, and overall Cronbach alpha .91. Full text also describes limitations of convenience sampling and lack of detailed measurement-invariance analysis.
- Escala de Inteligencia Emocional de Wong-Law (WLEIS-S): propiedades psicométricas y datos normativos en población adulta peruana
Studies 782 convenience-sampled adults aged 18–59 in metropolitan Lima using the Spanish WLEIS-S. Confirmatory analyses favor a bifactor model and report general and factor-specific reliability, sex and age-group invariance evidence, and sample-specific percentile norms. The paper also records favorable-response tendencies on positively worded items.
- Training emotional competencies at the workplace: a systematic review and metaanalysis
Systematic review and meta-analysis includes 50 pre-post samples (N=3,233) and 27 controlled comparisons of workplace emotional-competency training. It reports a moderate controlled effect (SMD .47) but substantial heterogeneity and a prediction interval crossing zero; 18 included studies had critical and 28 serious risk of bias, and only six were randomized trials. Most interventions and measures are not WLEIS-specific.
- A Meta-Analysis of the Relationships Between Emotional Intelligence and Employee Outcomes
Synthesizes organizational-outcome associations across ability, self-report, and mixed EI streams. For job performance it reports 31 self-report EI samples with N=10,438 and corrected correlation rho=.33; the overall estimate is rho=.30 across 68 samples and N=23,269. The review includes English-language studies published through early 2020 and notes missing contextual moderators; estimates pool instruments and are not WLEIS-specific individual forecasts.
- The relation between emotional intelligence and job performance: A meta-analysis
The publisher record and abstract summarize a meta-analysis comparing ability, self/peer-report, and mixed EI measurement streams and their relation to job performance. Use it only for its reported cross-stream comparison and broad pooled association; it does not establish that a WLEIS factor predicts a particular worker's performance.
Apply it to the real situation
Turn an emotional-skill result into one behavior to notice
From this guide: WLEIS can suggest a domain to reflect on; a concrete example from your own work can help you examine how that pattern shows up.
If you want a private starting point for reflecting on emotional habits, the Emotional Skills Profile offers 32 questions about recent behavior and emotional situations, followed by practical reflection guidance. It is a separate educational tool, not WLEIS or a normed ability test. Use the report to choose one observable behavior to examine in feedback, disagreement, or repair.
