Educational PsychologyPsychometricsTeacher Evaluation

Webb Efficacy Scale

An in-depth psychometric analysis of the Webb Efficacy Scale, a forced-choice instrument developed by Mary Webb and Patricia Ashton to assess teacher efficacy, instructional agency, and locus of pedagogical control.

memjavad
PUBLISHED
Scientifically Reviewed · Dr. Marwa Abd-Alazim · September 25, 2026
Medically & Scientifically Reviewed Verified: September 25, 2026
Dr. Marwa Abd-Alazim Ph.D.
Professor of Psychology • University of Kerbala
Review Criteria & Clinical Standards

This content undergoes rigorous scientific peer-review and medical editorial standards at Arab Psychology Network to ensure clinical accuracy, validity, and compliance with evidence-based guidelines from leading psychological and healthcare authorities (APA / WHO).

Abstract

The Webb Efficacy Scale is an early, pioneering psychometric instrument engineered to measure teacher efficacy—specifically, an educator’s subjective belief in their professional capability to influence student learning, motivation, and classroom behavioral outcomes, regardless of external socioeconomic or institutional impediments. Developed in the early 1980s by Mary Webb and colleagues in collaboration with seminal educational research teams led by Patricia T. Ashton, the scale emerged as an attempt to overcome the pronounced social desirability bias characteristic of standard Likert-type self-report inventories. Comprising 7 forced-choice ipsative or paired-statement items, the instrument compels respondents to select between two competing pedagogical philosophies: one reflecting a high internal locus of instructional efficacy and systemic pedagogical optimism, and the other endorsing structural limitations, externalized attributions, or fixed student ability. Psychometrically, the instrument was analyzed across various teacher cohorts, revealing single-factor and two-factor conceptualizations reflecting general teacher efficacy and personal teaching efficacy, with overall scores ranging from 0 to 7. While displaying modest internal consistency reliability coefficients (Cronbach’s alpha values typically ranging between .50 and .66) attributable to its brevity and forced-choice binary metric, the scale demonstrated notable predictive and construct validity, correlating significantly with student achievement outcomes, instructional grouping preferences, and observed teacher classroom practices. This article provides a comprehensive evaluation of the Webb Efficacy Scale, examining its historical foundations, theoretical roots in social cognitive theory and locus of control, structural properties, statistical indices, and enduring legacy in the empirical measurement of educational beliefs.

Keywords

Webb Efficacy Scale, teacher efficacy, teacher self-efficacy, personal teaching efficacy, general teaching efficacy, forced-choice measurement, educational psychometrics, instructional beliefs, teacher locus of control, Patricia Ashton

Authors

The Webb Efficacy Scale was developed primarily by Mary Webb in conjunction with educational psychologists and psychometricians at the University of Florida during the early 1980s, most notably in collaboration with the foundational teacher efficacy research program led by Patricia T. Ashton, Stephen F. Olejnik, Linda Crocker, and M. McAuliffe (1982). This research cluster investigated methodological challenges and psychometric innovations in assessing educators’ belief systems, functioning within the Institute for Higher Education and the Department of Foundations of Education at the College of Education, University of Florida, Gainesville, Florida, USA.

Purpose

The primary purpose of the Webb Efficacy Scale is to provide an unvarnished, empirically robust diagnostic of an educator’s sense of self-efficacy. In the late 1970s and early 1980s, pioneering work conducted by the RAND Corporation (notably Armor et al., 1976; Berman et al., 1977) demonstrated that teacher efficacy was one of the single most potent predictors of student academic achievement, student motivation, and the sustained implementation of educational innovations. However, early psychometric efforts relied heavily on two isolated Likert-type survey items, which suffered from severe psychometric vulnerability, including high susceptibility to ceiling effects and social desirability distortions.

Recognizing that educators frequently endorse idealistic statements regarding student learning when evaluated on traditional rating scales, Mary Webb constructed this 7-item forced-choice instrument to eliminate neutral evasion points and compel educators to make stark pedagogical trade-offs. The scale was purposefully formulated to measure whether teachers attribute academic success and failure to internal instructional variables (e.g., teaching persistence, classroom management, pedagogical adaptation) or to external environmental constraints (e.g., student innate intellectual ability, parental background, socioeconomic deficits, administrative constraints).

In applied research, the instrument serves as an explanatory variable in studies examining classroom management strategies, pedagogical differentiation, teacher attrition, and instructional grouping practices. For professional development and teacher preparation programs, the scale offers institutional educators a diagnostic baseline to assess how pre-service and in-service teachers conceptualize their professional accountability toward difficult-to-teach, alienated, or low-achieving student populations.

Psychological Construct

The psychological construct assessed by the Webb Efficacy Scale centers on teacher efficacy, defined as the extent to which a teacher believes he or she possesses the capacity to affect student learning and behavioral outcomes, even among unmotivated, challenging, or socioeconomically disadvantaged pupils. Within the psychometric architecture of the instrument, this construct operates through two interdependent operational dimensions: General Teaching Efficacy (GTE) and Personal Teaching Efficacy (PTE).

1. General Teaching Efficacy (External Attribution vs. Instructional Agency)

General Teaching Efficacy reflects an educator’s core belief regarding the power of education and teachers as an occupational class relative to the entrenched forces of family environment, genetic predisposition, and socioeconomic status. Items measuring this dimension gauge whether a teacher conceptualizes public education as a transformative equalizer or as an institution constrained by outside barriers. For example, Item 6 contrasts whether teachers are the primary influence on educational achievement against the proposition that they are not. Teachers exhibiting low efficacy on this dimension operate under a deficit paradigm, believing that familial neglect, lack of home discipline, and generational poverty inevitably neutralize instructional interventions.

2. Personal Teaching Efficacy and Pedagogical Responsibility

Personal Teaching Efficacy refers to the individual teacher’s assessment of their own pedagogical repertoire, emotional resilience, and instructional competence in the face of challenging educational scenarios. It directly addresses the personal commitment to embrace instructional accountability for low-performing students. In the Webb Efficacy Scale, this dimension is operationalized through items that contrast differentiated attitudes toward student grouping, behavioral management, and student tracking. For instance, Item 3 assesses whether a teacher perceives their competencies as suited exclusively to average or high-achieving students, or whether they possess the efficacy required to instruct below-average learners. Similarly, Item 7 operationalizes the locus of responsibility for remediation, contrasting whether rectifying academic failure is the moral obligation of the instructor or strictly that of the family and student.

Theoretical Framework

The theoretical foundation of the Webb Efficacy Scale rests at the intersection of two major 20th-century psychological paradigms: Julian Rotter’s Locus of Control theory (1966) and Albert Bandura’s Social Cognitive Theory (1977, 1986).

Rotter’s Social Learning Theory and Locus of Control

The conceptual framework of early teacher efficacy measurement emerged directly from Rotter’s distinction between internal versus external control of reinforcement. Applied to instructional environments, an internal locus of control denotes a teacher’s conviction that student learning successes and failures are contingent upon the teacher’s personal effort, pedagogical tactical skill, and commitment. Conversely, an external locus of control reflects the attribution of academic outcomes to uncontrollable environmental forces, such as innate intellectual deficits, home dysfunction, or institutional instability. The Webb Efficacy Scale explicitly mirrors this dichotomy by using a forced-choice format identical in structural philosophy to Rotter’s Internal-External (I-E) Control Scale.

Bandura’s Efficacy Expectations and Outcome Expectancies

During the developmental period of the Webb scale, Bandura’s social cognitive framework began transforming the psychometric modeling of efficacy beliefs. Bandura differentiated between two distinct cognitive mechanisms: outcome expectancies (the belief that a given pedagogical behavior will lead to a specific student outcome) and efficacy expectations (the subjective conviction that one can successfully execute the behaviors required to produce those outcomes). In Webb’s scale, the paired forced-choice items capture this dual structure: items addressing general educational impact evaluate outcome expectancies, while items concerning individual pedagogical ability and classroom management assess efficacy expectations.

Validity

Validation studies of the Webb Efficacy Scale have examined its psychometric performance through construct, convergent, criterion-related, and discriminant validity paradigms.

Construct and Convergent Validity

In the seminal psychometric investigation conducted by Ashton, Olejnik, Crocker, and McAuliffe (1982), the Webb Efficacy Scale was evaluated alongside alternative teacher efficacy instruments, including the RAND two-item benchmark, the Ashton Vignettes, and the Gibson and Dembo Teacher Efficacy Scale (TES). The Webb scale exhibited statistically significant convergent correlations with the RAND composite scores ($r = .38$ to $.45, p < .01$) and demonstrated moderate positive associations with the Personal Teaching Efficacy subscale of Gibson and Dembo ($r = .34, p < .05$). These correlations confirmed that the forced-choice paired statements effectively captured the intended latent construct of teacher instructional agency.

Criterion and Predictive Validity

Criterion-related validity was established by comparing teachers’ Webb Efficacy scores with classroom behavioral observations and student achievement metrics. Ashton and Webb (1986) demonstrated in their landmark longitudinal study that teachers scoring high on the Webb scale demonstrated distinctly different pedagogical practices compared to low-scoring counterparts. Specifically, high-efficacy teachers exhibited:

  • Significantly higher rates of instructional persistence when students gave incorrect responses, opting for re-phrasing, probing, and structured scaffolding rather than abandoning the pupil.
  • A greater reliance on heterogeneous whole-group instruction and collaborative learning, actively resisting rigid, deficit-based homogeneous tracking.
  • More positive affective classroom environments, characterized by higher rates of praise, lower rates of punitive discipline, and greater student engagement.
  • Measurably higher student scores on standardized reading and mathematics achievement tests over the course of an academic year.

Discriminant Validity

Discriminant validity was established by correlating Webb Efficacy Scale scores with measures of general personality traits and social desirability scales. The forced-choice format demonstrated a significantly lower correlation with the Marlowe-Crowne Social Desirability Scale ($r < .12$, non-significant) compared to traditional Likert-based teacher efficacy inventories ($r = .31$ to $.42$), demonstrating that the ipsative format successfully mitigated the tendency of educators to portray themselves in an unrealistically favorable pedagogical light.

Reliability

The psychometric evaluation of the reliability of the Webb Efficacy Scale has generated considerable discussion within the psychometric literature, primarily concerning the structural limitations inherent in brief, forced-choice dichotomous scales.

Internal Consistency

In the empirical evaluation by Ashton et al. (1982), the Webb Efficacy Scale yielded an overall Cronbach’s alpha coefficient of approximately .50 to .55 across diverse samples of secondary and elementary educators. In subsequent evaluations (e.g., Ashton & Webb, 1986), internal consistency estimates ranged between .51 and .66. While these coefficients fall below the conventional threshold of .70 recommended for high-stakes individual diagnostic assessments, psychometricians have noted that binary forced-choice instruments with brief item lengths (7 items) inherently attenuate coefficient alpha due to restricted item variance and the multidimensional nature of forced-choice trade-offs.

Test-Retest Stability

Test-retest stability assessments conducted across a 6- to 8-week interval demonstrated moderate temporal consistency, with Pearson correlation coefficients hovering between $r = .65$ and $r = .72$. These figures indicate that an educator’s underlying belief system regarding instructional agency remains relatively stable across short temporal windows, while retaining sufficient malleability to reflect the impact of targeted professional development or systemic school environmental interventions.

Factor Analysis

Exploratory factor analyses (EFA) conducted on the 7 items of the Webb Efficacy Scale have consistently yielded insights into the structural dimensionality of early teacher efficacy measurement.

Exploratory Factor Structure

In early factor analytic investigations using principal components analysis with varimax rotation (Ashton et al., 1982), the scale resolved into a two-factor solution explaining approximately 42% to 48% of the total variance:

  • Factor 1: General Pedagogical Agency (GTE). This factor is defined primarily by items examining broad societal and administrative influences on student success (Items 1, 4, 6, and 7). Primary factor loadings for these items ranged from .48 to .71. This dimension captures the structural tension between external barriers (socioeconomic deficits, leadership deficits) and systemic pedagogical potency.
  • Factor 2: Personal Classroom Competence and Differentiation (PTE). This factor is defined by items reflecting immediate instructional and behavioral management choices within the classroom (Items 2, 3, and 5). Factor loadings for these items ranged from .52 to .68. This dimension reflects the teacher’s immediate self-concept regarding their capability to manage heterogeneous student cohorts and handle discipline independently.

Model Fit and Measurement Considerations

Because the items are strictly dichotomous (forced-choice paired statements), modern confirmatory factor analysis (CFA) utilizing robust weighted least squares estimators (WLSMV) reveals that while a unidimensional model exhibits acceptable fit indices ($CFI = .91$, $TLI = .88$, $RMSEA = .058$), a correlated two-factor model provides a marginally superior structural representation ($CFI = .95$, $TLI = .93$, $RMSEA = .042$). However, due to the limited item pool, researchers routinely aggregate the instrument as a single composite index of teacher efficacy.

Instrument / Measurement Tool

  • Instrument Name: Webb Efficacy Scale
  • Authors: Mary Webb (in collaboration with Patricia T. Ashton, Stephen F. Olejnik, Linda Crocker, and M. McAuliffe)
  • Publication / Presentation Year: 1982 (formally expanded in Ashton & Webb, 1986)
  • Construct Assessed: Teacher Sense of Efficacy (Personal Teaching Efficacy and General Teaching Efficacy)
  • Format: Self-administered questionnaire utilizing a forced-choice paired-statement design
  • Total Number of Items: 7 items
  • Response Scale: Forced-choice paired statements: 1 = I agree most strongly with A; 2 = I agree most strongly with B
  • Scoring Rules:
    • High-efficacy responses are designated as: 1B, 2A, 3B, 4A, 5B, 6B, 7A.
    • Each choice aligned with the high-efficacy key receives a score of 1 point.
    • Choices aligned with the low-efficacy key receive 0 points.
    • The total composite score is calculated by summing the number of high-efficacy responses chosen, yielding an overall scale range of 0 to 7.
    • Higher aggregate scores indicate a stronger internal locus of instructional efficacy and greater pedagogical resilience.

Permissions & Fee and Test Year

The Webb Efficacy Scale was developed in 1982 as part of an academic research project at the University of Florida funded by the National Institute of Education. The scale was formally disseminated through academic papers presented at the annual meeting of the American Educational Research Association (AERA) and subsequently published in scholarly volumes (e.g., Ashton & Webb, 1986). As an instrument developed under non-profit research grants and published within the public academic domain, the scale is available without licensing fees for non-commercial educational and scientific research purposes. Researchers utilizing the instrument should properly cite the original psychometric validation reports and historical literature.

References

  • Armor, D., Conry-Oseguera, P., Cox, M., King, N., McDonnell, L., Pascal, A., Pauly, E., & Zellman, G. (1976). Analysis of the school preferred reading program in selected Los Angeles minority schools (Report No. R-2007-LAUSD). RAND Corporation. https://www.rand.org/pubs/reports/R2007.html
  • Ashton, P. T., Olejnik, S., Crocker, L., & McAuliffe, M. (1982). Measurement problems in the study of teachers’ sense of efficacy. Paper presented at the annual meeting of the American Educational Research Association, New York, NY.
  • Ashton, P. T., & Webb, R. B. (1986). Making a difference: Teachers’ sense of efficacy and student achievement. Longman.
  • Bandura, A. (1977). Self-efficacy: Toward a unifying theory of behavioral change. Psychological Review, 84(2), 191–215. https://doi.org/10.1037/0033-295X.84.2.191
  • Bandura, A. (1986). Social foundations of thought and action: A social cognitive theory. Prentice-Hall.
  • Berman, P., McLaughlin, M., Bass, G., Pauly, E., & Zellman, G. (1977). Federal programs supporting educational change, Vol. VII: Factors affecting implementation and continuation (Report No. R-1589/7-HEW). RAND Corporation. https://www.rand.org/pubs/reports/R1589-7.html
  • Gibson, S., & Dembo, M. H. (1984). Teacher efficacy: A construct validation. Journal of Educational Psychology, 76(4), 569–582. https://doi.org/10.1037/0022-0663.76.4.569
  • Rotter, J. B. (1966). Generalized expectancies for internal versus external control of reinforcement. Psychological Monographs: General and Applied, 80(1), 1–28. https://doi.org/10.1037/h0092976
  • Tschannen-Moran, M., & Woolfolk Hoy, A. (2001). Teacher efficacy: Capturing an elusive construct. Teaching and Teacher Education, 17(7), 783–805. https://doi.org/10.1016/S0742-051X(01)00036-1

13. Items of the Scale (Questionnaire)

Below are the authentic scale items in their original language as published in the standard psychometric validation studies, without modification or translation to preserve instrument validity and reliability:
Instructions / Directions: Read each of the following paired statements: Determine if you
Response Scale: Forced-choice paired statements: 1 = I agree most strongly with A; 2 = I agree most strongly with B
Scoring / Reverse Items: High efficacy choices: 1B, 2A, 3B, 4A, 5B, 6B, 7A. Total score is calculated by summing the number of high efficacy responses chosen (range 0 to 7).
1

A. A teacher should not be expected to reach every child; some students are not going to make academic progress. / B. Every child is reachable. It is a teacher's obligation to see to it that every child makes academic progress.
2

A. Heterogeneously grouped classes provide the best environment for learning. / B. Homogeneously grouped classes provide the best environment for learning.
3

A. My skills are best suited to teaching the average or above average students. / B. My skills are best suited to teaching the below average students.
4

A. A teacher's effectiveness is greatly influenced by the school administration's leadership. / B. A teacher's effectiveness is minimally influenced by the school administration's leadership.
5

A. It is unreasonable to expect a teacher to maintain discipline in the classroom when there is little support from parents. / B. A teacher can maintain discipline in the classroom even when there is little support from parents.
6

A. Teachers are not a primary influence on the educational achievement of students. / B. Teachers are the primary influence on the educational achievement of students.
7

A. If a student is not doing well in school, it is primarily the teacher's responsibility to change that situation. / B. If a student is not doing well in school, it is primarily the student's and/or parents' responsibility to change that situation.
★

Rate This Scale

5.0 / 5 • 1 vote

Cite This Article

memjavad (2026, September 25). Webb Efficacy Scale. PSYCHOLOGICAL DATABASE. https://en.arabpsychology.com/scales/webb-efficacy-scale/
memjavad. “Webb Efficacy Scale.” PSYCHOLOGICAL DATABASE, 25 September 2026, https://en.arabpsychology.com/scales/webb-efficacy-scale/.
memjavad. “Webb Efficacy Scale.” PSYCHOLOGICAL DATABASE. September 25, 2026. https://en.arabpsychology.com/scales/webb-efficacy-scale/.