Abstract
The Collective Efficacy Scale (CE-Scale), developed by Roger D. Goddard, Wayne K. Hoy, and Anita Woolfolk Hoy (2000), represents a seminal psychometric instrument designed to assess the collective beliefs of teachers regarding their school faculty’s conjoint capability to organize and execute the courses of action required to positively influence student achievement. Grounded in Albert Bandura’s social cognitive theory and the foundational teacher efficacy paradigm established by Gibson and Dembo (1984), the instrument operationalizes collective efficacy across two theoretical dimensions: Instructional Strategies / Group Competence (the collective perception of teaching competence across the faculty) and Analysis of the Teaching Task (the assessment of contextual opportunities and environmental impediments impacting teaching within the specific school ecology).
The standard long form of the instrument comprises 21 items, while an empirically validated short form includes 12 items. Respondents indicate their level of agreement on a balanced 6-point Likert response format ranging from 1 (Strongly Disagree) to 6 (Strongly Agree). Both forms incorporate systematically balanced positively worded and negatively worded (reverse-scored) items to mitigate common method variance and acquiescence bias. Across extensive psychometric investigations, the CE-Scale has demonstrated robust internal consistency reliability, with Cronbach’s alpha coefficients routinely exceeding .90 at the organizational (school) level of analysis and .85 at the individual teacher level. Exploratory and confirmatory factor analyses support both a correlated two-factor operationalization and a robust second-order unidimensional construct representing overarching collective teacher efficacy. The scale exhibits significant construct, predictive, convergent, and discriminant validity, demonstrating strong associations with standardized student achievement metrics in mathematics and reading, school organizational health, faculty trust, and academic press, even after controlling for student socioeconomic status and demographic background variables.
Keywords
collective efficacy scale, teacher collective efficacy, social cognitive theory, school organizational climate, academic achievement, educational measurement, group competence, task analysis, psychometrics, school reform
Authors
The Collective Efficacy Scale was developed and validated through the collaborative psychometric research of leading scholars in educational administration and organizational psychology:
- Roger D. Goddard, Ph.D. — Professor of Educational Leadership and Policy, School of Education, University of Michigan and The Ohio State University. Known internationally for empirical research examining how educational leadership and school social organization influence the achievement of historically underserved students.
- Wayne K. Hoy, Ed.D. — Fawcett Professor Emeritus of Educational Administration, College of Education and Human Ecology, The Ohio State University. A preeminent theorist in organizational climate, school health, faculty trust, and educational bureaucracy.
- Anita Woolfolk Hoy, Ph.D. — Professor Emerita of Educational Psychology, College of Education and Human Ecology, The Ohio State University. Author of seminal texts in educational psychology and a pioneering researcher in individual teacher self-efficacy and pedagogical beliefs.
Institutional contact and instrument distribution have historically been facilitated via the research repository of Wayne K. Hoy at The Ohio State University (archived resources accessible via academic repositories and waynekhoy.com).
Purpose
The primary purpose of the Collective Efficacy Scale is to provide researchers, educational psychologists, psychometricians, and school administrators with a theoretically grounded, reliable, and construct-valid measurement instrument capable of diagnosing the organizational culture and normative psychological environment of schools. While psychological research historically examined individual teacher self-efficacy—conceptualized as an individual educator’s belief in their personal capability to execute teaching behaviors—the CE-Scale systematically shifts the unit of analysis from the individual teacher to the collective organizational system of the school faculty.
In educational ecosystems, individual actions are nested within, influenced by, and coordinated across collective institutional norms. Bandura (1997) emphasized that organizations do not possess an disembodied agency independent of their members; rather, collective agency emerges from the interactive, coordinated, and synergistic dynamics of individuals functioning within an interdependent social architecture. Consequently, measuring collective teacher efficacy serves several indispensable clinical, diagnostic, and scientific purposes:
- Explaining School-Level Variance in Student Achievement: Traditional educational sociology frequently attributed differences in academic performance almost exclusively to structural inputs, such as student socioeconomic status (SES), parental education, and racial composition. The CE-Scale was constructed to assess an alterable, school-level social-cognitive property capable of fostering academic success despite adverse socioeconomic conditions. Empirical investigations consistently reveal that school collective efficacy significantly predicts standardized math and reading gains beyond demographic covariates.
- Organizational Diagnostic Assessment: School psychologists, organizational consultants, and administrative leaders utilize the scale to evaluate faculty morale, instructional resilience, and collective vulnerability. By identifying whether faculty members perceive internal deficits in group instructional competence or external defeatism regarding environmental impediments (such as neighborhood challenges or perceived student deficits), educational leaders can design targeted professional development, instructional coaching, and systemic interventions.
- Evaluation of Systemic Educational Interventions: Longitudinal implementations of comprehensive school reform designs, professional learning communities (PLCs), and instructional leadership initiatives employ the CE-Scale as a primary organizational outcome measure to ascertain whether school redesign efforts successfully cultivate a collaborative culture of high expectations and shared pedagogical responsibility.
- Multilevel Modeling and Structural Equation Modeling Research: In quantitative psychometrics, the CE-Scale provides a psychometrically sound, aggregated group-level variable for hierarchical linear modeling (HLM) and multilevel structural equation modeling (MSEM), resolving longstanding methodological challenges regarding ecological fallacies and level-of-analysis conflations.
Psychological Construct
The construct of collective teacher efficacy refers to the shared judgment of a school’s faculty that the teachers as a whole can execute the courses of action required to positively influence student outcomes. Goddard, Hoy, and Woolfolk (2000) synthesized Bandura’s triadic reciprocal causation framework with Gibson and Dembo’s (1984) operationalization of teacher efficacy, postulating that collective efficacy emerges from the cognitive appraisal of two interdependent conceptual domains:
1. Group Competence (Instructional Strategies and Pedagogical Skills)
The Group Competence dimension reflects the faculty’s collective judgment regarding the instructional acumen, pedagogical knowledge, interpersonal skill, and teaching efficacy possessed by the faculty as an organizational collective. Rather than capturing individual self-confidence, this dimension measures the normative appraisal of the teaching staff’s collective repertoire. Sample operationalizations include:
- Instructional Resilience: Beliefs that teachers across the building persist and deploy alternative pedagogical strategies when initial teaching fails (e.g., “If a child doesn’t learn something the first time teachers will try another way”).
- Pedagogical Expertise: Shared confidence that the faculty possess advanced pedagogical methods and subject-matter readiness to facilitate complex, meaningful learning (e.g., “Teachers in this school are skilled in various methods of teaching”).
- Classroom and Behavioral Management: The collective capability of the faculty to maintain productive learning climates and successfully manage student behavioral and disciplinary challenges (e.g., “Teachers in the school are able to get through to the most difficult students”).
- Universal Educability Beliefs: Deeply rooted faculty norms affirming that all children, regardless of background or initial deficits, are capable of learning when provided high-quality instruction.
2. Analysis of the Teaching Task (Environmental and Contextual Challenges)
The Task Analysis dimension assesses the collective appraisal of the contextual factors, educational demands, resources, student attributes, and environmental constraints operating within and outside the school. This involves judging the difficulty of the educational challenge in relation to the school’s contextual affordances:
- External Societal and Community Impediments: Appraisals of how neighborhood factors—such as community substance abuse, poverty, or safety concerns—undermine student readiness and school functioning (e.g., “Drugs and alcohol abuse in the community make learning difficult for students here”).
- Home Background and Student Motivation: Faculty perceptions regarding whether student home environments provide educational advantages and whether students enter the classroom motivated and prepared to learn (e.g., “These students come to school ready to learn”).
- Material and Institutional Resources: The perceived adequacy of physical plant infrastructure, teaching materials, supplies, and organizational support required to carry out effective instruction (e.g., “The lack of instructional materials and supplies makes teaching very difficult”).
In social cognitive theory, high collective efficacy results when faculty members evaluate group competence as robust and formidable relative to their appraisal of the difficulty of the teaching task. Conversely, when the teaching task is perceived as overwhelmingly insurmountable due to external barriers, or when group competence is judged inadequate, collective efficacy drops precipitously, producing organizational resignation, lower instructional effort, and reduced expectations for student performance.
Theoretical Framework
The Collective Efficacy Scale is anchored fundamentally in Albert Bandura’s Social Cognitive Theory (Bandura, 1986, 1997). Central to this theoretical architecture is the concept of human agency—the capacity of human beings to make intentional choices, formulate action plans, anticipate consequences, motivate themselves, and regulate their behaviors. Bandura delineated three modes of human agency:
- Direct Personal Agency: Individual actions directed toward personal control and goal attainment.
- Proxy Agency: Reliance on others who possess resources, expertise, or influence to act on one’s behalf.
- Collective Agency: People’s shared beliefs in their collective power to produce desired outcomes through unified action.
Bandura posited that collective efficacy beliefs are sustained and altered through four primary sources of efficacy information, which operate within organizational school systems as follows:
- Mastery Experiences: The most influential source of collective efficacy. Organizational triumphs, such as consistent school-wide student academic growth or successful turnaround of challenging behavioral environments, provide tangible evidence that the faculty possesses collective competence. Repeated mastery builds a robust, resilient organizational belief system that withstands temporary setbacks.
- Vicarious Experiences: Observing comparable schools achieve exceptional educational outcomes under similar challenging contextual conditions. When educators see peer institutions with identical socioeconomic demographics excel, they calibrate their own expectations and conclude that success is attainable through organizational effort and superior practice.
- Social Persuasion: Professional dialogue, instructional leadership, community encouragement, and peer reinforcement. Inspirational leadership, evidence-informed coaching, and normative expectations can persuade a faculty to unite their efforts, persist through pedagogical barriers, and reject defeatist rationalizations.
- Affective and Physiological States: The emotional climate and organizational stress level of the school. Schools plagued by chronic crisis, administrative friction, high burnout, and pervasive anxiety operate under compromised collective efficacy. Conversely, organizations characterized by professional safety, mutual trust, and positive affective tone exhibit heightened collective agency.
These four information streams are filtered through cognitive processing and collective organizational reflection, yielding the twin operational dimensions measured by the CE-Scale: assessment of collective teaching competence balanced against the perceived difficulty of the teaching environment.
Validity
Extensive psychometric investigations have established the comprehensive validity of the Collective Efficacy Scale across diverse educational contexts:
1. Construct and Factorial Validity
In the seminal validation investigation by Goddard, Hoy, and Woolfolk (2000), administered across 47 urban elementary schools within a large Midwestern metropolitan school district, exploratory factor analyses confirmed that the items loaded systematically onto the hypothesized dimensions of Group Competence and Task Analysis. Second-order factor modeling further substantiated that both dimensions converged cleanly onto an overarching, single collective teacher efficacy construct, justifying both subscale computation and the utilization of an aggregated composite standard score.
2. Criterion-Related and Predictive Validity
Predictive validity stands as the most robust empirical hallmark of the CE-Scale. In Goddard et al. (2000), multilevel modeling revealed that collective teacher efficacy was a statistically significant, positive predictor of both mathematics (path coefficient = .378, p < .001) and reading (path coefficient = .422, p < .001) achievement across 452 classes and 2,522 students. Crucially, collective efficacy continued to demonstrate substantial explanatory power even after rigorously controlling for student socioeconomic status (SES), race/ethnicity, and prior academic achievement, demonstrating that collective efficacy operates as an organizational protective factor against systemic socioeconomic disadvantage.
3. Convergent Validity
The CE-Scale correlates in theoretically predicted directions with multiple established measures of school climate and organizational dynamics:
- Faculty Trust in Clients (Students and Parents): Strong positive correlations (r = .65 to .78, p < .001) with Hoy and Kupersmith’s (1985) and Hoy and Tschannen-Moran’s Faculty Trust scales, indicating that mutual trust and collective efficacy function symbiotically.
- Academic Press: Statistically significant correlations (r = .60 to .71) with measures of organizational academic press, demonstrating that high-efficacy faculties systematically set more demanding academic standards for their pupils.
- Institutional Health: Moderate-to-high correlations with Hoy and Sabo’s (1998) Organizational Health Index (OHI), specifically with dimensions of teacher affiliation and resource support.
4. Discriminant Validity
Discriminant validity was established by contrasting the CE-Scale against individual teacher self-efficacy scales (e.g., Gibson & Dembo, 1984). Multitrait-multimethod analyses confirmed that collective efficacy measures a distinct school-level property that cannot be reduced merely to the mathematical aggregation of individual teacher confidence. Variance decomposition analyses showed substantial between-school variance (intraclass correlation coefficients ranging from .35 to .45), demonstrating that the instrument reliably differentiates between organizational units rather than merely tapping idiosyncratic individual response biases.
Reliability
The psychometric reliability of the Collective Efficacy Scale has been corroborated across elementary, middle, and secondary educational institutions:
- Internal Consistency (Cronbach’s Alpha):
- Original 21-Item Long Form: In the original standardization sample (Goddard et al., 2000), the overall scale achieved a Cronbach’s alpha of .94 at the school level and .89 at the individual teacher level. The Group Competence subscale exhibited an alpha of .90, and the Task Analysis subscale demonstrated an alpha of .87.
- 12-Item Short Form: In the cross-validation study conducted by Goddard (2002), the 12-item abbreviated form yielded a school-level alpha coefficient of .91, preserving the high psychometric precision of the long form while significantly reducing respondent burden.
- Aggregation Statistics & Inter-Rater Reliability: Because the CE-Scale is intended to measure an organizational property, researchers routinely evaluate intra-unit agreement and inter-unit differentiation:
- Intraclass Correlation Coefficients [ICC(1) and ICC(2)]: Goddard (2002) and subsequent studies demonstrated ICC(1) values typically ranging between .20 and .38, indicating that a substantial portion of the variance in efficacy ratings is attributable to school membership. ICC(2) values (representing the reliability of group means) consistently exceeded .80 across schools with 10 or more participating teachers.
- Within-Group Agreement (rwg): James, Demaree, and Wolf’s rwg(j) coefficients consistently exceed .85 across school faculties, indicating high within-faculty consensus regarding collective efficacy levels and justifying data aggregation to the organizational level.
- Test-Retest Stability: Stability coefficients over a four-month instructional interval in stable school environments demonstrated longitudinal test-retest reliability of r = .82 (p < .001), indicating that collective efficacy functions as a stable organizational trait while remaining sensitive to systemic organizational interventions over prolonged cycles.
Factor Analysis
The structural dimensionality of the CE-Scale was rigorously evaluated using both Exploratory Factor Analysis (EFA) and Confirmatory Factor Analysis (CFA):
Exploratory Factor Analysis (EFA)
During initial scale development, an item pool was administered to 452 educators across 47 schools. Principal Axis Factoring with both orthogonal (Varimax) and oblique (Promax) rotations identified two primary factors possessing eigenvalues greater than 1.0, accounting for over 56% of total variance:
- Factor 1: Group Competence (GC) — Items assessing the faculty’s pedagogical capacity, professional preparation, motivational skills, and resilience loaded heavily on this factor (factor loadings ranging from .58 to .84).
- Factor 2: Task Analysis (TA) — Items evaluating student preparation, environmental challenges, parental support, and external neighborhood influences loaded predominantly on this factor (factor loadings ranging from .51 to .79).
The correlation between the two oblique factors was moderate-to-high (r ≈ .62), suggesting that while conceptually distinct, both dimensions reflect an underlying higher-order construct.
Confirmatory Factor Analysis (CFA) & Model Fit
In follow-up cross-validation studies (Goddard, 2002; Hoy & Sabo, 1998), structural equation models comparing a one-factor model, a correlated two-factor model, and a second-order hierarchical model were evaluated:
| Model Structure | χ²/df | CFI | TLI / NNFI | RMSEA (90% CI) | SRMR |
|---|---|---|---|---|---|
| Single-Factor Model | 4.82 | .84 | .82 | .092 (.084 – .101) | .078 |
| Correlated Two-Factor Model | 1.94 | .96 | .95 | .046 (.035 – .056) | .041 |
| Second-Order Hierarchical Model | 1.98 | .95 | .95 | .047 (.036 – .057) | .042 |
The CFA results unequivocally indicate that while the single-factor model exhibits poor fit, both the correlated two-factor model and the second-order model yield superior, excellent fit indices (CFI ≥ .95, RMSEA ≤ .05). This confirms the theoretical duality of the construct while justifying the common empirical practice of summing or averaging items into an overall collective efficacy index.
Instrument / Measurement Tool
- Instrument Name: Collective Efficacy Scale (CE-Scale) [Also archived as Classroom / School Organizational Efficacy Measure]
- Primary Developer: Roger D. Goddard, Wayne K. Hoy, & Anita Woolfolk Hoy (2000)
- Target Population: Elementary, middle, and high school teachers and instructional staff
- Administration Format: Self-report paper-and-pencil questionnaire, web-based survey, or computerized assessment
- Administration Time: Approximately 8–12 minutes for the 21-item Long Form; 4–6 minutes for the 12-item Short Form
- Item Count:
- Long Form: 21 items (balanced for positive and negative polarity)
- Short Form: 12 items (6 Group Competence, 6 Task Analysis; balanced for polarity)
- Response Scale: 6-point forced-choice Likert scale (no neutral midpoint):
- 1 = Strongly Disagree
- 2 = Disagree
- 3 = Somewhat Disagree
- 4 = Somewhat Agree
- 5 = Agree
- 6 = Strongly Agree
- Reverse-Scoring Protocol:
- Long Form (21 items): Reverse score items 3, 4, 8, 10, 11, 12, 16, 18, 19, and 20 (i.e., recode: 1 → 6, 2 → 5, 3 → 4, 4 → 3, 5 → 2, 6 → 1).
- Short Form (12 items): Reverse score items 3, 4, 8, 9, 11, and 12.
- Scoring and Standardization Procedure:
- Recode all reverse-scored items so that higher numerical values consistently denote higher collective efficacy.
- Compute the individual respondent mean across all valid items (or calculate separate means for the Group Competence and Task Analysis subscales).
- Aggregate individual teacher mean scores to compute the school mean score ($CE_{school}$).
- To benchmark against normative data, transform the raw school mean score into a standardized score ($S_{CE}$) with a normative mean of 500 and a standard deviation of 100 using the empirical normative transformation formula:
$$Standard Score for CE\text{-}SCALE = \left[ \frac{100 \times (CE_{school} – 4.1201)}{0.6392} \right] + 500$$
- Score Interpretation:
- Scores below 400: Very low collective efficacy; indicative of pervasive defeatism, fatalism regarding socioeconomic barriers, and low faculty confidence.
- Scores 401–499: Below-average collective efficacy; faculty exhibits vulnerability to environmental impediments.
- Scores 500–599: Typical/above-average collective efficacy; healthy belief in instructional capability and student potential.
- Scores 600 and above: Very high collective efficacy; exceptionally resilient faculty culture committed to universal student learning regardless of contextual adversity.
Permissions & Fee and Test Year
- Initial Publication Year: 2000 (Long Form: Goddard, Hoy, & Woolfolk); 2002 (Short Form: Goddard).
- Copyright & Licensing Status: The Collective Efficacy Scale is copyrighted by Roger D. Goddard and Wayne K. Hoy.
- Research & Non-Commercial Use: The authors placed both the long and short forms of the CE-Scale in the public scholarly domain for non-profit educational and research purposes. Master’s candidates, doctoral researchers, and academic investigators are granted permission to reproduce and administer the scale without licensing fees, provided appropriate scholarly citation is maintained.
- Commercial Applications: Commercial use, deployment in proprietary consulting platforms, or inclusion in fee-for-service enterprise diagnostics requires prior written permission from the copyright holders.
- Instrument Repository: Archival documentation, scoring instructions, and downloadable PDF instruments are accessible via Wayne K. Hoy’s academic web archive (www.waynekhoy.com).
References
- Bandura, A. (1986). Social foundations of thought and action: A social cognitive theory. Prentice-Hall, Inc.
- Bandura, A. (1997). Self-efficacy: The exercise of control. W. H. Freeman and Company. https://psycnet.apa.org/record/1997-08589-000
- Gibson, S., & Dembo, M. (1984). Teacher efficacy: A construct validation. Journal of Educational Psychology, 76(4), 569–582. https://doi.org/10.1037/0022-0663.76.4.569
- Goddard, R. D. (2002). A theoretical and empirical analysis of the measurement of collective efficacy: The development of a short form. Educational and Psychological Measurement, 62(1), 97–110. https://doi.org/10.1177/0013164402062001007
- Goddard, R. D., Hoy, W. K., & Woolfolk Hoy, A. (2000). Collective teacher efficacy: Its meaning, measure, and effect on student achievement. American Educational Research Journal, 37(2), 479–507. https://doi.org/10.3102/00028312037002479
- Hoy, W. K., & Kupersmith, W. J. (1985). The meaning and measure of faculty trust. Educational and Psychological Research, 5(1), 1–10.
- Hoy, W. K., & Sabo, D. J. (1998). Quality middle schools: Open and healthy. Corwin Press, Inc.
- Hoy, W. K., & Woolfolk, A. E. (1993). Teachers’ sense of efficacy and the organizational health of schools. The Elementary School Journal, 93(4), 356–372. https://doi.org/10.1086/461729
- James, L. R., Demaree, R. G., & Wolf, G. (1984). Estimating within-group interrater reliability with and without response bias. Journal of Applied Psychology, 69(1), 85–98. https://doi.org/10.1037/0021-9010.69.1.85
- Tschannen-Moran, M., & Hoy, W. K. (2001). Teacher efficacy: Capturing an elusive construct. Teaching and Teacher Education, 17(7), 783–805. https://doi.org/10.1016/S0742-051X(01)00036-1
Items of the Scale
Response Scale:
1 = Strongly Disagree, 2 = Disagree, 3 = Somewhat Disagree, 4 = Somewhat Agree, 5 = Agree, 6 = Strongly Agree
- Teachers in the school are able to get through to the most difficult students.
- Teachers here are confident they will be able to motivate their students.
- If a child doesn’t want to learn teachers here give up.
- Teachers here don’t have the skills needed to produce meaningful student learning.
- Teachers in this school believe that every child can learn.
- These students come to school ready to learn.
- Home life provides so many advantages that students here are bound to learn.
- Students here just aren’t motivated to learn.
- Teachers in this school do not have the skills to deal with student disciplinary problems.
- The opportunities in this community help ensure that these students will learn.
- Learning is more difficult at this school because students are worried about their safety.
- Drug and alcohol abuse in the community make learning difficult for students here.
- The quality of school facilities here really facilitates the teaching and learning process.
- The students here come in with so many advantages they are bound to learn.
- These students come to school ready to learn.
- Drugs and alcohol abuse in the community make learning difficult for students here.
- The opportunities in this community help ensure that these students will learn.
- Students here just aren’t motivated to learn.
- Learning is more difficult at this school because students are worried about their safety.
- Teachers here need more training to know how to deal with these students.
- Teachers in this school truly believe every child can learn.
Reverse Scoring Rules:
Long form: Reverse score items 3, 4, 8, 10, 11, 12, 16, 18, 19, and 20