How teacher rotation in Japanese high schools affects the clustering of teacher quality: Comparing the distribution of teachers across public and private education sectors Journal website: http://epaa.asu.edu/ojs/ Manuscript received: 3/6/2020 Facebook: /EPAAA Revisions received: 12/14/2020 Twitter: @epaa_aape Accepted: 4/8/2021 education policy analysis archives A peer-reviewed, independent, open access, multilingual journal Arizona State University Volume 29 Number 91 July 5, 2021 ISSN 1068-2341 How Teacher Rotation in Japanese High Schools Affects the Clustering of Teacher Quality: Comparing the Distribution of Teachers across Public and Private Education Sectors1 Ryan Seebruck Independent Scholar United States Citation: Seebruck, R. (2021). How teacher rotation in Japanese high schools affects the clustering of teacher quality: Comparing the distribution of teachers across public and private education sectors. Education Policy Analysis Archives, 29(91). https://doi.org/10.14507/epaa.29.5362 Abstract: I examine a unique facet of Japan’s public education system: jinji idou, a mandatory teacher rotation system governed by the prefectural board of education where teachers are systematically transferred to other schools throughout their careers to appropriately staff schools, facilitate varied career paths, and identify future leaders for administrative roles. Although not a formal goal, this centralized system may also produce a more equal distribution of teacher quality across schools compared to the de centralized teacher labor market found in private schools. Because this system is present in public schools and absent in private schools, comparing sector differences offers a look at its impact on teacher quality distribution. Using a sample of 1,456 tea chers nested in 49 schools, private vs. public group comparison tests indicate that, for most of the teacher 1 This research was funded in part by the following agencies or fellowships: Fulbright Fellowship, National Science Foundation, Japan Society for the Promotion of Science, National Council on Teacher Quality, Boren Fellowship, and the University of Arizona (Social and Behavioral Sciences). http://epaa.asu.edu/ojs/ https://doi.org/10.14507/epaa.29.5362 Education Policy Analysis Archives Vol. 29 No. 91 2 quality traits examined, the public sector distributes teachers more equitably. Furthermore, the public sector has higher mean levels of teacher quality, intimating that education labor markets can be structured in ways that simultaneously minimize variation between schools without hindering quality, findings germane to scholars interested in educational equality. Keywords: structural inequality; education; teacher rotation; labor markets; Japan Cómo la rotación de maestros en las escuelas secundarias japonesas afecta la agrupación de la calidad de los maestros: Comparación de la distribución de maestros en los sectores de educación público y privado Resumen: Examino una faceta única del sistema de educación pública de Japón: jinji idou, un sistema de rotación de maestros obligatorio gobernado por la junta de educación de la prefectura donde los maestros son transferidos sistemáticamente a otras escue las a lo largo de sus carreras para dotar de personal apropiado a las escuelas, facilitar trayectorias profesionales variadas y identificar futuros líderes para roles administrativos. Aunque no es un objetivo formal, este sistema centralizado también puede producir una distribución más equitativa de la calidad de los maestros entre las escuelas en comparación con el mercado laboral descentralizado de los maestros que se encuentra en las escuelas privadas. Debido a que este sistema está presente en las escuelas públicas y ausente en las escuelas privadas, comparar las diferencias sectoriales ofrece una mirada a su impacto en la distribución de la calidad de los docentes. Utilizando una muestra de 1.456 maestros agrupados en 49 escuelas, las pruebas de comparación de grupos públicos y privados indican que, para la mayoría de los rasgos de calidad de los maestros examinados, el sector público distribuye a los maestros de manera más equitativa. Además, el sector público tiene niveles medios más altos de calidad docente, lo que sugiere que los mercados laborales de la educación pueden estructurarse de manera que simultáneamente minimicen la variación entre escuelas sin obstaculizar la calidad, hallazgos relevantes para los académicos interesados en la igualdad educativa. Keywords: desigualdad estructural; educación; rotación de maestros; mercados laborales; Japón Como a rotação de professores nas escolas secundárias japonesas afeta o agrupamento da qualidade dos professores: Comparando a distribuição de professores nos setores de educação pública e privada Resumo: Eu examino uma faceta única do sistema de educação pública do Japão: jinji idou, um sistema de rotação de professores obrigatório governado pelo conselho de educação da prefeitura onde os professores são sistematicamente transferidos para outras escolas ao longo de suas carreiras para escolas com funcionários adequados, facilitando caminhos de carreira variados e identificar futuros líderes para funções administrativas. Embora não seja uma meta formal, esse sistema centralizado também pode produzir uma distribuição mais igualitária da qualidade dos professores entre as escolas, em comparação com o mercado de trabalho descentralizado dos professores encontrado nas escolas privadas. Como esse sistema está presente nas escolas públicas e ausente nas escolas privadas, a comparação das diferenças setoriais oferece uma visão de seu impacto na distribuição da qualidade dos professores. Usando uma amostra de 1.456 professores aninhados em 49 escolas, os testes de comparação de grupos privados vs. públicos indicam que, para a maioria das características de qualidade dos professores examinadas, o setor público distribui os professores de forma mais equitativa. Além disso, o setor público tem níveis How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 3 médios mais elevados de qualidade dos professores, sugerindo que os mercados de trabalho da educação podem ser estruturados de forma que simultaneamente minimizem a variação entre as escolas sem prejudicar a qualidade, conclusões pertinentes a acadêmicos interessados na igualdade educacional. Palavras-chave: desigualdade estrutural; Educação; rotação de professores; mercados de trabalho; Japão How Teacher Rotation in Japanese High Schools Affects the Clustering of Teacher Quality: Comparing the Distribution of Teachers acr oss Public and Private Education Sectors I investigate the organization of education labor markets in Japan: how the careers of educators are managed and how this affects teacher distribution. This is predicated on four claims. First, teacher quality is a predominant causal variable affecting student performance, outweighing others such as class size, student background, or spending (Darling-Hammond, 2000; Rice, 2003). Second, there is an unequal distribution of quality teachers in decentralized education labor markets, such as in the US, that disfavors underprivileged students (Borman & Kimball, 2005). Third, this unequal distribution is caused by policies that encourage quality teachers to gravitate to affluent schools while less qualified teachers remain at poor schools (Peske & Haycock, 2006). Fourth, overall Japan’s education system is more egalitarian in that socioeconomically disadvantaged students have equal access to quality teachers along with their advantaged peers, unlike in the US (PISA, 2010, p. 156)—a country with one of the largest opportunity gaps in student access to qualified teachers (Akiba et al., 2007). Given the importance of teacher quality to student performance, that the U.S. system’s localized structure has produced a disparity in opportunities justifies examining Japan’s centrally controlled system as an alternative organizational tactic to providing equal educational opportunities, as one explanation for the maldistribution of teacher quality in the US is because of its laissez-faire teacher labor market in which educators largely have autonomy over where they work (Prince, 2002). This locally controlled system has produced policies inhibiting educational egalitarianism, such as funding laws that enable schools in wealthy cities to offer higher salaries, teacher transfer restrictions that make it difficult to move to low-achieving districts, or seniority clauses enabling experienced teachers to choose their placements (Prince, 2002). Akiba and LeTendre (2009) argue that the decentralization of education systems in the U.S. has created a disparity in school quality and educational opportunities around socioeconomic status when contrasted with Japan’s centralized public school system, which enables a broader approach to teacher hiring and placement. A possible solution to the maldistribution in teacher quality found in the US is to adopt systematic teacher transfers similar to jinji idou in Japan (Letendre, 2000; White, 1987; Wray, 1999), a system “specifically designed to reduce school-based differences in overall teacher quality” by assigning and periodically rotating public school teachers to other schools throughout their careers (Akiba et al., 2007, p. 381). This system has been lauded by qualitative scholars for creating a more equal distribution of teachers by preventing the clustering of quality teachers at certain schools (Letendre, 2000; White, 1987; Wray, 1999). For instance, Wray (1999, p. 31) asserts that the systematic rotation of teachers prevents “the highest ranked schools from always having the best teachers” and ensures that rural areas have a fair and consistent supply of teachers. These qualitative studies suggest that teacher rotation creates a more equal distribution of teacher quality, which should create more equal educational opportunities for students. Education Policy Analysis Archives Vol. 29 No. 91 4 However, while extant studies on Japan’s teacher rotation system laud its benefits, there is a paucity of comparative, quantitative analyses of its effect on teacher distribution, particularly in English. Japanese studies on the topic tend to focus qualitatively on how the system works (Choshi, 2015) or how it affects teachers’ career paths (Fukuya, 2015) and professional development (Kawakami & Senoh, 2011) rather than its impact on teacher distribution. I therefore complement these studies by quantitatively comparing the distribution of teachers across public and private schools in Japan. Because jinji idou is absent in private schools, a comparison of the distribution of teacher quality across schools in the private sector with its public sector counterpart produces a comparative design that increases the ability to credit it as a causal condition affecting teacher distribution, as there is ultimately only one factor governing where teachers work: either teachers have the autonomy to choose individual schools (as in the private sector) or they do not (as in the public sector). Finally, comparing systems within a prefecture in Japan mitigates concern over externalities as lurking variables. Following that, this paper fills a significant gap in the literature by conducting the first large-scale, quantitative analysis of the distribution of teacher quality in Japan that is written in English. Background Teacher Quality and its Import There is consensus among recent education research that, controlling for socioeconomic status and other student factors, school-level variables impact student performance, with particularly strong effects for teacher characteristics (Darling-Hammond, 2000; Prince, 2002; Rice, 2003; Rivkin et al., 2005; Sanders & Rivers, 1996; Wright et al., 1997). For example, Wright et al. (1997, p. 66) conducted a multivariate, longitudinal, mixed-model analysis of school resources and student achievement across multiple subjects and grades, finding that “teacher effects are dominant factors affecting student academic gain.” Furthermore, in their longitudinal study of a cohort of elementary students, Sanders and Rivers (1996) found the effects of teacher quality on student achievement are additive and cumulative: while individual students did not recover after a year under an ineffective teacher, students who spent a year under an effective teacher experienced benefits up to two years later. The authors conclude that teacher quality is more highly correlated with student achievement than other variables such as student traits like ‘race’ or socioeconomic status. Similarly, in a large- scale study on teacher quality and educational equality, Borman and Kimball (2005) used multilevel models to show that classes taught by higher quality teachers produced higher mean achievement than those taught by lower quality teachers. Most education systems emphasize quality teaching as a means of enhancing student achievement, but there is no accord with what constitutes teaching quality (Croninger et al., 2007; Rice, 2003; Rivkin et al., 2005). Consequently, in this paper I analyze a variety of common teacher quality measures: full-time teacher status, certification status, years of experience, teaching in-field, holding an advanced degree, and the prestige of one’s alma mater. Hereafter I describe each. Full-Time Status: The rationale behind using full-time status as a teacher quality measure is severalfold. First, full-time employees are likely to be more committed to their jobs and employers just as those employers are also more likely to be committed to them—points especially relevant to Japan (Ouchi, 1981). This is evident in how full-time teachers are better remunerated (Ishikida, 2005; Okano & Tsuchiya, 1999). Second, one way to capture unobserved traits such as teacher quality is to capture them with another concept—in this case, full-time teacher status. This is particularly true in the Japanese prefecture studied here where becoming a full-time teacher requires passing a test and an interview for a limited number of slots (Seebruck, 2019). The thought is that the administrators How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 5 know a great deal about the teachers they hire and make hiring decisions based on their judgments of those teachers. This has support in the literature, with Jacob and Lefgren (2008) finding that school principal rankings of teachers are superior predictors of teacher quality than are observed teacher traits (although it is not clear if this proxy is as equally robust in the US and Japan). Certification: Teacher certification is the bestowal of a formal document by an accredited institution recognizing one’s ability to teach in a particular jurisdiction. While many teacher quality measures are thought to contribute to the distinction between a high and low quality teacher, certification status is often considered the most reliable predictor of student achievement as becoming certified as a teacher in most U.S. states requires formal education in a state-approved education program, the completion of either a major or minor in the subject field, plus minimum satisfaction of education credits and student teaching credits (Darling-Hammond, 2000). Certification requirements are similarly stringent in Japan (MEXT, 2016) and certification status factors into teacher rotation decisions (Fukuya, 2015). Because of the strictness and thoroughness of these requirements, certification status is one of the strongest indicators of teacher performance (Seebruck, 2015). Experience: Viewed as a proxy for on-the-job training (Harris & Sass, 2011), studies have consistently found a positive relationship between teacher efficacy and years of experience (Klitgaard & Hall, 1975; Murnane & Phillips, 1981). Darling-Hammond (1995) notes that experienced teachers are more effective at resolving classroom problems, maintaining discipline, motivating students, and adapting to students’ diverse learning needs. Clotfelter et al. (2007), using value-added models to analyze the effects of teacher characteristics on student achievement, found that teacher experience has positive effects. Likewise, Fetler (1999) found that student achievement in mathematics significantly correlates with teacher experience. Pupils of first-year teachers learn less, on average, than pupils of more experience teachers (Boyd et al., 2008). Simply put, experienced teachers are more knowledgeable about curriculum, instruction, and assessment (Prince, 2002). In Japan, years of experience factors into teacher development (MEXT, 2016) and rotation decisions (Seebruck, 2019) and correlates with teacher salaries (Okano & Tsuchiya, 1999). Novice: There is ambiguity regarding the causal relationship between teacher experience and student achievement—not about whether experience matters but about how much experience matters. Buddin and Zamarro (2009) contend that the impact of teacher experience is more of a dichotomy between novices and veterans than a linear progression. Correspondingly, Rice (2003) notes that efficacy gains are particularly salient for those in their first few years of teaching. Although years of experience correlates with student achievement (Fetler, 1999), evidence suggests a flattening slope in their relationship. For instance, while beginning teachers (those with zero to three years of experience) at the high school level are less effective than those with more than three years of experience, there is no difference between teachers with three to five years of experience and those with 25 years of experience (Clotfelter et al., 2010). In-Field: Another common teacher quality measure is whether one teaches in their field of expertise, as teaching in-field positively contributes to education outcomes for students (Rice, 2003). Rowan, Chiang, and Miller (1997) tested the impact of the subject matter of teachers’ degrees, finding that mathematics teachers with a degree in mathematics positively predicted student achievement among high school students. Several other studies have also found that subject-specific training significantly influences high school students’ math and science performance (Monk, 1994; Wenglinsky, 2002). Advanced Degree: Having an advanced degree such as a master’s presumably indicates a higher quality teacher since obtaining one requires years of additional coursework and training; but numerous studies argue that teacher quality appears unrelated to advanced degree status (Croninger Education Policy Analysis Archives Vol. 29 No. 91 6 et al., 2007; Hanushek et al., 2005; Ladd & Sorensen, 2015). One explanation for the mixed findings is the conflation of effects at the elementary and high school levels. For example, Rice (2003) claims that, although the evidence between holding an advanced degree and student achievement is mixed at the elementary school level, at the high school level it is clearer, with evidence showing positive effects on science and math achievement. As this study surveys high schools, I include advanced degree status as a teacher quality measurement. Alma Mater: Many education scholars consider teachers’ alma mater as a stand-in for unobservable traits tied to teacher quality, such as intelligence or ambition. As Rice (2003) notes, studies suggest the selectivity or prestige of one’s alma mater has a positive impact on student performance, particularly for high school students. Summers and Wolfe (1975) examined a random sample of nearly 1,900 urban middle school and high school students and found, after controlling for other factors, that the selectivity of the teacher’s alma mater was significantly related to teacher efficacy: pupils taught by instructors who graduated from higher-ranked universities fared better, on average, than their peers whose instructors attended lower-ranked universities. Teacher Quality and its Distribution As important as teacher quality is to student success, quality teachers are not distributed equally across U.S. school districts (Prince, 2002). This discrepancy in access to quality teachers contributes to achievement gaps between racial and socioeconomic groups in the US (Clotfelter et al., 2010), disadvantaging lower-class and minority youth (Borman & Kimball, 2005; Peske & Haycock, 2006; Seebruck, 2016). Lankford, Loeb, and Wyckoff (2002) stress there is an uneven sorting of teachers across schools, with the least qualified ones clustering at schools with higher shares of disadvantaged and under-achieving students. Not only have these discrepancies in access to equal educational resources remained consistent over time but, in many cases, the gaps are widening despite an elevated priority among politicians in redressing them (Barton & Coley, 2009). Several qualitative researchers have suggested that jinji idou, the mandatory teacher rotation system of Japan’s centrally controlled education labor market, produces a more equal distribution of quality teachers by preventing the clustering of quality teachers at certain schools (Letendre, 2000; White, 1987; Wray, 1999). In this system the prefectural board of education rotates public school teachers to other schools in the prefecture approximately every five years, throughout their entire careers. Kariya (2011, p. 247) describes teacher rotation at the middle school level as a measure to ensure equity—“to equalize quality of teaching between schools.” Teacher rotation occurs at the high school level too but, unlike at the elementary and middle school levels, high schools are more stratified due to the requirement for matriculates to pass entrance examinations (Ono, 2001; Rohlen, 1983). Consequently, I analyze the impact of jinji idou on the distribution of teachers at the high school level, where, because of its tiered nature, it is less clear if the top-down allocation of personnel will remain as equitable as it is at the elementary and middle school levels. This study is therefore motivated by the qualitative studies described above arguing that teacher quality distribution in Japan’s public education system does not suffer from the same clustering seen in U.S. districts but, rather, provides more equal access to qualified teachers on a variety of measures, according to national-level statistics (Akiba et al., 2007). That is, this study aims to examine quantitatively if the expectations of the teacher rotation system’s expected equitability hold true at the high school level, which is particularly important in Japan given the key role high schools play in determining social placement in adulthood (Rohlen, 1983). Following that, I examine the distribution of teachers across high schools in a specific prefecture to determine if the different organizational structures of private and public sector teacher labor markets in Japan explains differences in teacher quality distribution. How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 7 Data Expectations The purpose of this paper is to investigate the internal distribution of high school teachers across sectors in a Japanese prefecture, to determine if the mandatory teacher rotation system present in the public sector but absent in the private sector results in significant differences.2 Within Japan’s 47 prefectures, the prefecture studied here ranks in the top fifteen in land area, population, and population density, and the high school education system comprises over 100,000 students and nearly 7,000 full-time teachers (Statistics Japan, 2013). There are a variety of formal and informal policies governing systematic rotation of public school teachers in this prefecture (for details on teacher rotation in this prefecture, see Seebruck, 2019), but what is important to this paper is that private school teachers have autonomy over where they work whereas public school teachers’ career paths are decided by the prefectural board of education. I have two expectations based on qualitative data. First, the public education sector in the Japanese prefecture surveyed is seen as being of higher quality than the private sector. Given the strong association between teacher quality and student performance, I expect the public sector to have a higher average level of teacher quality. This proposition is stated formally: P1: The public education sector will have higher average teacher quality, compared to the private sector. Second, the mandatory, systematic teacher transfer system that is present in the public sector but absent in the private sector should lead to a less clustered teacher quality distribution in the public sector. The rationale stems from qualitative studies arguing Japan’s mandatory teacher rotation system produces a less clustered distribution of teachers (Letendre, 2000; White, 1987; Wray, 1999). This proposition is stated formally: P2: The public education sector should have a smaller variance in the between-school distribution in teacher quality, compared to the private sector. To collect the quantitative data needed to confirm these propositions, I designed teacher- and school-level questionnaires and surveyed high schools from 2011 to 2012. Survey I surveyed high schools in a Japanese prefecture using a disproportionate stratified random sample without replacement of ‘normal’ high schools, stratified on sector (public, private) and region (east, central, west). By ‘normal’ I mean standard academic high schools, as opposed to vocational, correspondence, part-time, night, or branch schools. There are 83 such schools, 51 public and 32 private. Based on my experience working in the Japanese public education system, along with my pilot research on the jinji idou system in the years prior and my understanding of the literature on teacher quality distribution, I created a multi-faceted questionnaire designed to gather individual- level data on teacher quality traits (described below). The paper-based questionnaire was written in Japanese and edited by native speakers to improve its readability. In addition to a teacher-level questionnaire, I also constructed a school-level questionnaire aimed at providing population data for each school, to aid in statistical weighting procedures. The survey resulted in a final sample of forty-nine schools for a final response rate of 72.1% at the organizational level. Each school averaged nearly 30 respondents, with an average teacher- 2 Employee rotation among conglomerates is possible in the private sector; however, there is only one such school in the sample that has a sister school to which teachers could be rotated. This sister school is a vocational school and therefore is not in the sampling frame. Sensitivity analyses removing the school in question did not significantly change the findings from those reported in the paper. Education Policy Analysis Archives Vol. 29 No. 91 8 level response rate of 60.2%, a minimum of 26%, and a maximum of 100%. Thus, the final analytic sample comprises 1,456 teachers—794 teachers in 23 private schools and 662 teachers in 26 public schools. Prior to surveying, I conducted power analyses to determine the minimum sample size needed to produce a representative sample with a normal distribution (Cohen, 1992). The power analyses revealed that, with a Level 2 sample size of 49 schools, an average Level 1 sample size of 30 teachers per school, and a maximum intraclass correlation coefficient (ICC) of 0.146, the final sample has an estimated power greater than 0.98, which is well above the standard cutoff of 0.80. Operationalization The questionnaire gathered information on the following teacher traits commonly used in the literature as an indicator of teacher quality: full-time status, certification status, years of experience or status as a novice, teaching in-field, holding an advanced degree, and the prestige of one’s alma mater. I test these indicators both separately and as an index (described below). Full-time status is a dichotomous variable coded as 1 for those who are a full-time teacher and coded 0 for those who are not. It stems from the question, “What is your current employment status?” Of the 1,453 teachers who responded to this question, 1,158 were full-time teachers (79.7%), 137 were full-time lecturers (9.4%), and 158 were part-time lecturers (10.9%). To dichotomize these responses, I collapsed the second and third options (lecturers), coding them as 0 and leaving full-time teachers coded as 1.3 Certification status is a dichotomous variable coded as 1 for those who hold the highest level of teacher certification and coded 0 for those who do not. It stems from the question, “What type of teaching certificate to you have?” Of the 1,383 teachers who responded, 216 held a specialized certificate (15.6%), 1,036 held a primary license (74.9%), and 131 held a secondary license (9.5%). To dichotomize these responses, I coded as 1 those who had a specialized certificate, indicating the highest level of certification, and coded as 0 everyone else. Years of Experience is a continuous measure consisting of total years of teaching experience. It stems from the prompt, “Write your total years of teaching experience below (include both public and private school experience, part-time and full-time, as well as experience at different school levels).” Of the 1,449 teachers who responded to this question, the mean years of teacher experienced was 18.2 (with a standard deviation of 11.6 years), a minimum of 0 years of experience (i.e. a first-year teacher) and a maximum of 52 years of experience. The 25th, 50th, and 75th percentiles were 8, 19, and 28 years. Status as a Novice is an alternative, non-linear measure of teacher experience—a dichotomous variable distinguishing beginning teachers and non-beginning teachers, demarcated at having at least three full years of experience. This threshold is based on evidence of when teachers begin to become more effective (Clotfelter et al., 2010; Hanushek et al., 2005; Klitgaard & Hall, 1975; Murnane & Phillips, 1981). This three-year demarcation also fits with qualitative examinations of teacher rotation systems in Japan as an important cutoff (Fukuya, 2015; Seebruck, 2019). Teachers amidst their first, second, or third year of teaching were coded as 1; those amidst their fourth year of teaching or beyond were coded as 0. Of the 1,449 teachers who responded, 159 were defined as novices (11%) whereas 1,290 reported having equal to or greater than three full years of teaching experience (89%). 3 Interviews with Japanese educators indicated they consistently view the demarcation between full-time teachers (i.e. permanent employees) and full-time lecturers (i.e. those on annual contracts) as important. As a robustness check, I also tested the impact of combining full-time lecturers into the full-time teacher category. Compared to the results in Table 2, there are no significant differences, with the p-values for both the LR and PM tests significant at the .001 level regardless of operationalization. How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 9 Teaching In-Field is a dichotomous variable based on teachers’ responses to two questions: “What is the primary subject you currently teach?” and “Is this the same subject as your college major?” Teachers who answered ‘yes’ to the latter were coded as 1 for teaching in their field of expertise and coded as 0 if they answered ‘no.’ Of the 1,420 teachers who responded, 1,262 reported that they were primarily teaching in their field of specialization (88.9%) and 158 reported that the primary subject they were currently teaching was not the same as their field of specialization (11.1%). Advanced Degree is a dichotomous variable collapsed from the question “What is the highest degree you hold?” To dichotomize educational attainment, I coded as 1 those having either a master’s or a doctoral degree (246 respondents, or 17%) and coded as 0 those teachers who had an associate or bachelor’s degree (1,200 respondents, or 83%). Alma Mater Prestige is a continuous measure based on Shimano’s (2009) annual rankings of Japanese universities, running from 1 to 10 (low to high prestige). Shimano’s rankings are based on hensachi— “deviation values” stemming from a university’s acceptance rate, which are considered the most renowned college ranking system in Japan (Masuda, 2003). Shimano relies on the hensachi rankings compiled by Yoyogi Seminar, which is one of the largest yobikou—for-profit, private college preparatory schools (Rohlen, 1983)—in Japan that are officially recognized by the Ministry of Education, Culture, Sports, Science and Technology (Blumenthal, 1992). Shimano improves the Yoyogi Seminar scores via university-wide aggregation and temporal and sectoral adjustments. Shimano adjusts for changes in scores over time—since the range in hensachi scores over the past few decades have contracted—by assigning schools an ordinal ranking based on their temporally contextual hensachi scores. Shimano’s rankings also adjust for the fact that hensachi scores are higher in private universities due to the different types of entrance examinations issued by these schools. Finally, instead of reporting only disparate scores for colleges within universities like other reporting agencies, Shimano unifies his rankings, proving one score for each university. Of the 1,129 teachers who responded, there was a mean score of 8.3, a standard deviation of 1.7, a minimum of 2.5 and a maximum of 10. The 25th, 50th, and 75th percentiles saw prestige scores of 7.0, 9.0, and 9.5.4 Teacher Quality Index (TQI) is an integer-scale index of teacher quality traits, as a means of capturing the combined effects of teacher quality, comprising six dichotomous variables delineated above: full-time, certified, veteran (i.e. not a novice), in-field, advanced degree, and prestigious alma mater (i.e. greater than 8.0 on Shimano’s scale). These dichotomous variables are summed to create a composite score ranging from 0, indicating low teacher quality, to 6, indicating high teacher quality.5 For example, a full-time (1), certified (1), novice (0) with an advanced degree (1) from a non- prestigious institution (0) who is teaching in-field (1) would garner a 4 out of 6 on the index. Of the 1,067 teachers who had scores for all six constituents, 6 had a score of 0 (0.6%), 49 had a score of 1 (4.6%), 117 had a score of 2 (11.0%), 333 had a score of 3 (31.2%), 391 had a score of 4 (36.6%), 83 had a score of 5 (7.8%), and 88 had a score of 6 (8.3%). The average score was a 3.6, and the standard deviation was 1.2. 4 I ran sensitivity analyses for missing data by testing a dichotomized version, where 1 = missing and 0 = non-missing. There were no significant differences between the distributions of these variables across sectors, indicating missing values are approximately uniformly distributed and thereby supporting imputation by mean substitution. Relatedly, I also tested two dichotomous versions of the prestige scores, capturing elite universities with thresholds of 8 and 9 on the alma mater prestige scale. Results corroborated the continuous measure reported in Table 2, with no significant differences in p-values. 5 I employ a summative index for its ease of use (J.-O. Kim & Rabjohn, 1980); see Jackson (2012) for an example of a predicted value-added teacher quality index. Education Policy Analysis Archives Vol. 29 No. 91 10 Methods Overview As the analyses are based on school-level traits, teacher-level responses are aggregated by school. The first analysis examines the sector-level grand mean in teacher quality (i.e. the overall average score for each teacher quality measure in the public sector versus the same in the private sector). The second analysis examines the sector-level, between-school grand coefficient of variation (i.e. the overall relative standard deviation between schools in the public sector versus the same in the private sector). In the case of dichotomous variables, the aggregated mean indicates the proportion of respondents satisfying the outcome. Consequently, this results in a variance that is the product of the probability of the outcome being true and one minus the probability of it not being true (Janda, 2015). Calculating those sector-level grand means and variances enables inter-sector comparisons of these intra-sector, between school differences (see Biggs 1991). This makes it possible to compare both the average teacher quality in the public and private school sectors as well as inter-sector differences in how teacher quality is distributed across schools within these sectors. Selection Bias Because jinji idou is absent in private schools, a comparison with public schools increases the ability to credit it as a causal condition affecting teacher distribution as, ultimately, private school teachers choose where they work whereas public school teachers do not. Comparing systems within a single prefecture mitigates concern of external parameters clouding the results. That said, given the differences between the public and private education labor markets here—such as the compulsory teacher rotation system, the socioeconomic statuses of schools, etc.—there is the potential of self- selection bias obfuscating the results as different types of teachers and students may prefer one sector to the other, which could impact the main variable of interest: teacher quality distribution. However, my research design addresses the issue of self-selection bias in the following ways. As the primary analysis compares the overall variance in the distribution of teacher quality in the public sector to the overall variance in the distribution of teacher quality in the private sector, it is therefore a comparison of the relative variation in each sector. This inter-sector comparison of intra-sector distributions mitigates concerns of self-selection bias across sectors by first obtaining the variance in teacher quality between schools in the same sector, and then comparing those variances across sectors to determine if there is an organizational effect on the variance of the distribution of teacher quality. In other words, my primary analyses are an inter-sector comparison of intra-sector variance, not a sector-level comparison of teacher quality variance. Put differently, regardless of differences between public and private school teachers, this study examines how giving one of those groups autonomy to choose where they work (e.g. an organic distribution) while the other group’s career paths are centrally controlled (e.g. an artificial distribution) impacts how those teachers are dispersed across schools within their respective sectors. Nevertheless, despite the analyses being inter-sector comparisons of intra-sector differences, selection bias could still confound those results if there is a drastic difference in the mean values of teacher quality across sectors, which could make difficult comparisons of the variance since standard deviations are based on the mean. A solution to this is to employ the relative variability—that is, the coefficient of variation—which is the standard deviation divided by the mean and is useful when comparing variation in samples with different means (Marwick, 2018). This standardizes the comparisons of variance across sectors and is the method used here. How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 11 Additionally, I employ supplementary sensitivity analyses via randomized reassignment and difference-in-difference models. These results are available in the appendix and confirm that the different distributions in the public and private sectors are not entirely due to selection bias. Analyses Approach I use Wald tests for the equality of sector-level grand means to test Proposition 1: the public education sector will have higher average teacher quality. I then analyze sector-level differences in within-sector, between-school variation in teacher quality, conducting likelihood-ratio and Pitman- Morgan tests for the equality of sector-level grand coefficients of variation to test Proposition 2: the public education sector will have a smaller variance in the between-school distribution of teacher quality. As the data come from a multi-stage, stratified random sample with disproportionate sampling without replacement, multi-level sampling weights were applied to account for survey design, non-response, and non-coverage, raking on full-time teacher status and gender at the teacher level. For missing data (e.g. 29 of the 1,480 respondents had missing data on gender), I used multiple imputation via chained regression, thereby enabling me to complete the raking process. However, in contrast to typical multilevel weights, in which the Stage 1 final weight is multiplied with the Stage 2 final weight, because the analyses involve sector-level comparisons of aggregated, school-level data, the multi-stage weights were calculated sequentially. That is, first, the Level 1 teacher weights were employed while aggregating the school-level data, and then the Level 2 school weights were subsequently employed to calculate the sector-level values. Difference in Means There are eight analyses, one for each measure of teacher quality: full-time status, certification status, years of experience, status as a beginning teacher, teaching in-field, holding an advanced degree, the prestige of one’s alma mater, and an integer-based Teacher Quality Index (TQI) composed of dichotomous versions of six teacher quality traits: full-time, certification, experience, in-field, advanced degree, and alma mater prestige. Sector-level summary statistics for the eight measures of teacher quality (not shown) reveal potential support for Proposition 1 as the public sector has preferable mean scores on every measure. For employment status, 81.8% of public school teachers, on average, are full-time teachers compared to only 66.7% in the private sector. For credentialization, 16.0% of public school teachers are highly credentialed compared to 11.0% in private schools. The average years of experience in the public sector is 19.1 versus 15.1 in the private sector; and for the percentage of novices per school the public sector has an average of 11.9% compared to 16.5% in the private sector (n.b. a lower percentage of novices is preferable). The public sector also has a higher percentage of teachers teaching in their field of expertise at 92.0%, compared to 87.3 in the private sector. In contrast to expectations that private schools would entice educated teachers by financially rewarding those with advanced degrees, the public sector has a higher percentage of teachers on this measure (15.9%) than the private sector (12.9%). The public sector also has higher average prestige scores for teachers’ alma mater, at 8.7 compared to 7.7. Finally, the public sector also has a higher mean TQI at 3.7 versus 3.1. To test whether these differences in means are statistically significant, I employ Wald tests for the equality of means (Gupta & Ma, 1996; Shoukri et al., 2008). The results of these analyses are in Table 1. The means calculated are sector-level grand averages—that is, they are the mean scores per sector, of the means scores per school, for schools in that sector. Using two-tailed tests, the F- Education Policy Analysis Archives Vol. 29 No. 91 12 values for the following teacher quality measures all have statistically significant p-values: the percentage of full-time teachers at a school, mean years of experience at a school, the percentage of teachers teaching in-field at a school, the mean prestige rating at a school of teachers’ alma mater, and the mean TQI score at a school. For all five of those measures, the public sector has a significantly higher grand mean, thereby supporting Proposition 1. Conversely, Proposition 1 was not supported for the following teacher quality traits: the percentage of highly credentialed teachers at a school, the percentage of novices teaching at a school, and the percentage of teachers at a school with advanced degrees. Table 1 Inter-Sector Comparison of Intra-Sector Mean Teacher Quality Wald Test for the Equality of Sector-Level Grand Means Teacher Quality Vars. Sector Mean Std. Dev. F- value p-value Outcome of Sig. Test Full-Time % Private 0.6673 0.1692 13.57 0.0006*** Public mean is significantly higher Public 0.8181 0.1284 Highly Credentialed % Private 0.1102 0.1253 3.12 0.0843 No significant difference between sectors Public 0.1595 0.078 Years of Experience Private 15.1197 5.1328 8.29 0.0062** Public mean is significantly higher Public 19.0823 4.5833 Novices % Private 0.1648 0.1313 1.75 0.1923 No significant difference between sectors Public 0.1187 0.1045 In-Field % Private 0.8732 0.0641 7.38 0.0095** Public mean is significantly higher Public 0.9197 0.0504 Graduate Degree % Private 0.1286 0.1174 1.47 0.2318 No significant difference between sectors Public 0.1586 0.073 Alma Mater Prestige Private 7.6873 0.8796 26.10 0.0000*** Public mean is significantly higher Public 8.7082 0.5266 Teacher Quality Index Private 3.0965 0.6151 20.88 0.0000*** Public mean is significantly higher Public 3.7225 0.3464 Note. Weighted data, 2-tailed tests, N = 49 (23 Private, 26 Public), df = 43, p < .05* p < .01** p < .001*** That most teacher quality measures were higher in the public sector suggests that the qualitative evidence proclaiming the public sector in the sampled Japanese prefecture is of higher quality seems to have merit in terms of teacher quality. This intimates that there is some selection bias in this prefecture’s teacher labor markets, with higher quality teachers, on average, selecting into How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 13 the public sector. There may be many reasons for this—such as differences in salaries or working conditions, or the desire to teach more motivated or proficient students (Seebruck, 2016)—but, for now, this is an interesting finding because it means the jinji idou system is not so off-putting to quality teachers that they eschew the public sector. (In fact, many public school teachers I interviewed stated that the teacher rotation system was a positive factor that increased job satisfaction). It is also important to reiterate that this finding of the public sector generally seeing higher levels of teacher quality is a generalization based on sector-level data; there is variation in teacher quality between schools in both sectors. Difference in Variation That most of the teacher quality traits analyzed here have higher averages in the public sector is interesting, but the core analysis is the inter-sector comparison of intra-sector, between- school differences in average teacher quality, as the primary research question is whether the different organizational structures of the public- and private-sector teacher labor markets affect the distribution of teacher quality. In the US, there is a pervasive maldistribution of teacher quality (Borman & Kimball, 2005), with high quality teachers clustering at certain schools, at the expense of others. This maldistribution has been attributed to laissez-faire policies that largely permit school teachers in the US to work where they want (Peske & Haycock, 2006). This decentralized organizational setup is also found among the private-sector education labor market in Japan. In contrast, public sector education labor markets in Japan are centrally governed by prefectural boards of education, with all public high school teachers subjected to regular, compulsory transfers within each prefecture. This teacher transfer system likely affects the distribution of teacher quality. To test this I employ likelihood-ratio tests (Bhoj & Ahsanullah, 1993; Gupta & Ma, 1996; Liu et al., 2011; Shoukri et al., 2008; Verrill, 2009) and Pitman-Morgan tests (Morgan, 1939; Pitman, 1939; Shoukri et al., 2008; Snedecor & Cochran, 1989) on the equality of grand coefficients of variation across sectors. The Pitman-Morgan test for the equality of coefficients of variation has been employed in a variety of disciplines (Y. K. Kim, Kim, Park, & Lee, 2017) as it is favored for being a powerful, unbiased test compared to a typical F-test (Haynes, 1981; Mudholkar et al., 2003), particularly when there is a possibility of non-independence among subjects (Cochran, 1965), such as may be the case when analyzing overlapping labor market pools (many teachers in the sample have worked in both sectors). The Pitman-Morgan test excels across many types of distributions, particularly normal or lightly tailed ones (Wilcox, 2015). However, it encounters difficulty with heavier tailed ones (Mudholkar et al., 2003). Consequently, in line with Shoukri et al. (2008), in addition to the Pitman-Morgan tests I also employ likelihood-ratio tests as a measure of robustness. Each sector’s grand coefficient of variation—that is, relative standard variation—is compared because the previous subsection demonstrated significant differences between mean teacher quality scores across sectors (Sokal & Braumann, 1980). Comparing the coefficient of variation instead of standard deviation mitigates concerns of self-selection bias in each sector’s teacher labor market impacting the results, instead standardizing the variation measure to allow for inter-sector comparisons of intra-sector, inter-school differences in teacher quality. The results of the likelihood-ratio (LR) and Pitman-Morgan (PM) tests for the equality of sector-level grand coefficients of variation are listed in Table 2. Larger coefficients of variation indicate higher variation. The tests are two-tailed. Overall, the LR and PM tests largely support each other, with the p-values for both tests agreeing on six of the eight measures. On five of those six measures, both tests demonstrate significant differences between the grand coefficients of variation across sectors for the following measures: the percentages of full-time teachers at a school, highly credentialed teachers at a school, Education Policy Analysis Archives Vol. 29 No. 91 14 and teachers at a school who hold an advanced degree as well as the average prestige rating of teachers’ alma mater and average Teacher Quality Index (TQI) score at a school. The results for these five unambiguous results support Proposition 2. Table 2 Inter-Sector Comparison of Intra-Sector, Between-School Differences in Teacher Quality Likelihood Ratio and Pitman-Morgan Tests for the Equality of Sector-Level Grand Coefficients of Variation Teacher Quality Vars. Sector CV LR Test p-value PM Test p-value Outcome of Sig. Test Full-Time % Private 0.2535 0.0000*** 0.0011** Private sector is significantly more variable Public 0.1569 Highly Credentialed % Private 1.1368 0.0048** 0.0000*** Private sector is significantly more variable Public 0.4893 Years of Experience Private 0.3395 0.1158 0.0164* Mixed support: Private sector may be more variable Public 0.2402 Novices % Private 0.7966 0.7414 0.4779 No significant difference between sectors Public 0.8802 In-Field % Private 0.0734 0.1554 0.0415* Mixed support: Private sector may be more variable Public 0.0548 Graduate Degree % Private 0.9129 0.0136* 0.0000*** Private sector is significantly more variable Public 0.4602 Alma Mater Prestige Private 0.1144 0.0022** 0.0000*** Private sector is significantly more variable Public 0.0605 Teacher Quality Index Private 0.1986 0.0004*** 0.0000*** Private sector is significantly more variable Public 0.0930 Note. Weighted data, 2-tailed tests, N = 49 (23 Private, 26 Public), df = 43, p < .05* p < .01** p < .001*** Both the LR and PM tests are also in accord that the there is no significant difference in the variation of the percentage of novices teaching at a school, across sectors. Thus, Proposition 2 is not supported for this measure. One possible reason for this null finding is that the public sector has more rural schools, which are more difficult to staff, particularly with veteran teachers. This is the only one of the eight measures to be unambiguously non-significant. There are two measures where the LR and PM tests are not in accord: years of experience and the percentage of in-field teachers at a school. For these two variables, the PM test finds a statistically significant difference between sectors, with the between-school variation in the private sector significantly larger at the .05 level. However, the p-values for the LR tests of these two measures are not statistically significant, creating an ambiguous situation. In this sense, Proposition 2 is How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 15 supported for these two variables, but the results are not as unambiguously robust as the other teacher quality measures. Data Visualization Although statistical significance tests are useful in hypothesis testing, particularly when testing for differences in means and variation, they should not be relied on alone, as it is possible for data sets dissimilar in reality to appear similar when comparing summary statistics (Anscombe, 1973). I graph the statistics formally tested in the previous sections to determine if the distribution of the data reaffirms the results of the statistical significance tests, contradicts them, or unveils differences not previously identified. Regarding the inter-sector comparisons of intra-sector, between-school variation in mean teacher quality, results of the Pitman-Morgan (PM) tests found that the private sector had significantly larger between-school, relative variation for seven out of the eight measures of teacher quality: full-time percentage, highly credentialed percentage, years of experience, in-field percentage, graduate degree percentage, alma mater prestige, and TQI. The only variable not found to significantly differ by sector, according to the PM tests, was the percentage of novices. Likelihood-ratio (LR) tests largely reaffirmed these findings, except for years of experience and in-field percentage, which were non-significant in the LR tests. To examine these relationships, I graph between-school differences using mean plots, depicted in Figure 1. Private and public schools in the sample are demarcated, with private schools on the left and public schools on the right in each figure. Every ring represents a school, and a ring’s location on the y-axis depicts the school-level mean score for the teacher quality trait listed. The solid circle on either side of the graph represents the sector-level, grand mean of those school-level means. The solid line traversing the graph offers a visual comparison of the difference in grand means for each sector. The dotted envelope slopes demarcate one grand standard deviation above and below the grand mean. Therefore, narrowing slopes across the graph intimate a shrinking variation in between-school, mean teacher quality. First, I will address the five teacher quality measures that were found to be unambiguously significantly higher in the private sector. Regarding the percentage of full-time teachers, it is clear from the mean plot that the public sector has a higher grand mean and slightly narrower envelope slopes, mostly due to the wider range on the y-axis between schools in the private sector. For the percentage of certified teachers, the mean plot reveals more dispersion among private schools as well, with noticeably wider envelope slopes compared to the public sector. Graduate degree percentage likewise has narrower envelope slopes in the private sector, as do alma mater prestige and the TQI. Second, the visual depictions of the sole variable that was unanimously non- significant—the percentage of novices—seems to support that null finding, as it is difficult to discern any difference in the width of the envelope slopes for that variable. These are all in line with the LR and PM tests in Table 2. The two variables where the LR and PM tests differed are more interesting cases. Examining the mean plot for years of teaching experience reveals a higher grand mean for the public sector, but differences in the grand standard deviation are unclear, as the general dispersion for both are similar, as are the corresponding widths of their envelope slopes. Thus, visualizing this variable did not clearly support either the LR or PM test. In contrast, examining the mean plot for the percentage of in-field teachers is more definitive, with a noticeably larger range in mean teacher quality among private schools, with one having a 100% score and two all the way down near the 75% mark. In contrast, the range of public schools is much smaller. As such, the envelope slopes are wider in the private sector. Together, the visual evidence reaffirms the PM test of a significant difference in sector-level variation for one of the two previously unclear variables: in-field teachers. Education Policy Analysis Archives Vol. 29 No. 91 16 Figure 1 Mean Plots of School-Level Teacher Quality Measures, By Sector Note: Y-axis location of each ring is a school’s mean score for a teacher quality trait. Dotted lines demarcate one grand standard deviation from sector-level grand means (i.e., the solid lines endpoints). How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 17 Figure 1 (Cont’d) Mean Plots of School-Level Teacher Quality Measures, By Sector Note: Y-axis location of each ring is a school’s mean score for a teacher quality trait. Dotted lines demarcate one grand standard deviation from sector-level grand means (i.e., the solid lines endpoints). Education Policy Analysis Archives Vol. 29 No. 91 18 Conclusion This paper tested (1) the difference in average high school teacher quality between the public and private education sectors in a Japanese prefecture, and (2) the differences in the distribution of high school teacher quality across these two sectors. The former analysis answers the question of which sector, on average, has higher levels of teacher quality. Remembering that variation exists within sectors as well (meaning many private schools have higher mean levels of teacher quality than many public schools), Wald tests for the equality of sector-level grand means reveal that the public sector has higher levels of average teacher quality for the following five measures: the percentage of full-time teachers, average years of teaching experience, the percentage of teachers primarily teaching their subject of expertise, the average prestige rating of teachers’ alma mater, and the average score on the composite measure (the Teacher Quality Index). These findings support Proposition 1 for those traits. Proposition 1 was not supported for three teacher quality measures: the percentage of highly credentialed teachers, the percentage of novices, and the percentage of teachers with an advanced degree, as no significant difference was found between sectors for those measures. The second set of analyses investigated sector-level differences in within-sector, between- school variation in teacher quality, conducting likelihood-ratio (LR) and Pitman-Morgan (PM) tests for the equality of sector-level grand coefficients of variation. These inter-sector analyses of intra- sector, between-school differences in teacher quality support the argument that the public sector in the sampled Japanese prefecture has a more equal distribution of quality teachers than does the private sector. The PM tests found that, for seven out of the eight measures, the private sector had significantly more clustering of teacher quality, thereby supporting Proposition 2 with one exception: the percentage of novices, which was not significantly different across sectors. The LR tests corroborated five of those seven differences, with visualizations of the data adjudicating and finding some support for the PM tests. Collectively, these examinations reveal that the distribution of quality teachers in the public sector is more equal compared to the distribution of quality teachers in the private sector. As the coefficient of variation was used to essentially control for differences in the populations of teachers in each sector (Sokal & Braumann, 1980), this suggests that the most conspicuous organizational difference between these two education labor markets—that is, the centrally controlled, large-scale, systematic rotation of teachers across schools that is present in the public sector but absent in the private sector—likely accounts for their different teacher distributions. Akin to the maldistribution of teacher quality in the US, the private sector in this Japanese prefecture has notable variation in teacher quality across schools. This is interesting because the education labor market in the Japanese private sector, being locally controlled, is organized similarly to both private and public teacher labor markets in the US. In contrast, the Japanese public sector, with its centrally controlled education labor market marked by a mandatory, systematic, and career- long geographic relocation of teachers across schools, has a more even distribution of teacher quality. These findings are important since the negative effects of being taught by ineffective teachers is substantial. Consider Rivers and Sanders’s (2002, p. 18) finding that the “cumulative and residual effects of teachers on the academic progress of students are huge” and that the “extreme variability of teachers’ effectiveness” in the US has a “dramatic effect” on student progress. Centralizing control over teachers’ career paths like Japan’s public school system may be one solution to the issue of educational inequality in the US (Akiba & Letendre, 2009), where school districts are locally governed and financed and where there is a trend of well-off areas trying to break away from larger districts, at the expense of poorer areas (Newkirk, 2014). Increased decentralization such as this could exacerbate the maldistribution of educational resources whereas centralization of How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 19 resources, particularly governance over teacher labor markets, could alleviate inequalities in educational opportunities. Such changes surely would need to be modified to fit sociopolitical nuances and geographic challenges in the US—a difficult endeavor given the federalist nature of its education systems—but, even then, narrowing the opportunity gap in students’ access to quality teachers does not guarantee a narrowing of the achievement gap (Akiba et al., 2007). That said, there are some limitations to this study that could be addressed in future research. First, this study examines only high schools. This reduces generalizability, as there are qualitative differences between education levels when it comes to how jinji idou operates—such as the geographic range of teacher transfers, which spans the entire prefecture for high schools but are limited to smaller geographical municipalities in elementary and middle schools (Seebruck, 2019). Moreover, unlike high schools in Japan, which are very much tiered academically by student capability, this is less of an issue at elementary and middle schools, where students do not have to pass entrance examinations to gain enrollment (Kariya, 2011; Ono, 2001; Rohlen, 1983). As such, there should be a less variation in student achievement across elementary and middle schools in Japan, making for an interesting scenario for a replication of the analyses in this paper. Second, this study examines a single prefecture and other scholars have shown jinji idou policies vary somewhat across prefectures in Japan (Kawakami, 2013). Replication studies in other prefectures are needed to determine the generalizability of the findings here. Finally, while this paper examined the distribution of various teacher quality traits, it does not examine the extent to which those teacher quality traits impact student achievement. Future research could build off these findings by ascertaining the extent to which the traits studied here matter to student performance and by including student and family voices as well. Acknowledgements I thank Dr. Joe Galaskiewicz, Dr. Jane Zavisca, Dr. Jeff Sallaz, and Dr. Hirohisa Takenoshita for their comments on this research. References Abadie, A., Diamond, A., & Hainmueller, J. (2010). Synthetic control methods for comparative case studies: Estimating the effect of California's Tobacco Control Program. Journal of the American Statistical Association, 105(490), 493-505. https://doi.org/10.1198/jasa.2009.ap08746 Akiba, M., & Letendre, G. K. (2009). Improving teacher quality: The U.S. teaching force in global context. Teachers College Press. Akiba, M., Letendre, G. K., & Scribner, J. P. (2007). Teacher quality, opportunity gap, and national achievement in 46 countries. Educational Researcher, 36(7), 369-387. https://doi.org/10.3102/0013189X07308739 Anscombe, F. J. (1973). Graphs in statistical analysis. American Statistician, 27, 17-21. https://doi.org/10.1080/00031305.1973.10478966 Barton, P. E., & Coley, R. J. (2009). Parsing the achievement gap II. http://www.ets.org/Media/Research/pdf/PICPARSINGII.pdf Bhoj, D. S., & Ahsanullah, M. (1993). Testing equality of coefficients of variation of two populations. Biometrical Journal, 3, 355-359. https://doi.org/10.1002/bimj.4710350311 Education Policy Analysis Archives Vol. 29 No. 91 20 Biggs, M. F. (1991). Methodology for calculating a “grand mean” and a “grand standard deviation” within 640K DOS spreadsheet software. Quality Engineering, 3(4), 455-459. https://doi.org/10.1080/08982119108918874 Blumenthal, T. (1992). Japan's juken industry. Asian Survey, 32(5), 448-460. https://doi.org/10.2307/2644976 Borman, G. D., & Kimball, S. M. (2005). Teacher quality and educational equality. Elementary School Journal, 106(1), 3-20. https://doi.org/10.1086/496904 Boyd, D., Lankford, H., Loeb, S., Rockoff, J., & Wyckoff, J. (2008). The narrowing gap in New York City teacher qualifications and its implications for student achievement in high- poverty schools. Journal of Policy Analysis and Management, 27(4), 739-818. https://doi.org/10.1002/pam.20377 Buckley, J., & Shang, Y. (2003). Estimating policy and program effects with observational data: The “differences-in-differences” estimator. Practical Assessment, Research & Evaluation, 8(24), 1-10. Buddin, R., & Zamarro, G. (2009). Teacher qualifications and student achievement in urban elementary schools. Journal of Urban Economics, 66(2), 103-115. https://doi.org/10.1016/j.jue.2009.05.001 Card, D., & Krueger, A. B. (1994). Minimum wages and employment: A case study of the fast- food industry in New Jersey and Pennsylvania. The American Economic Review, 84(4), 772- 793. https://doi.org/10.3386/w4509 Choshi, D. (2015). Kyoiku keieini okeru kyoin jinji idouno kenkyuu (Research of teacher personnel changes in educational management). Tokyo University Graduate School of Education Bulletin, 55, 471-480. Clotfelter, C., Ladd, H., & Vigdor, J. (2007). How and why do teacher credentials matter for student achievement? (NBER Working Paper Series Vol. 12828 Working Paper 13443). https://doi.org/10.3386/w12828 Clotfelter, C., Ladd, H., & Vigdor, J. (2010). Teacher credentials and student achievement in high school: A cross-subject analysis with student fixed effects. Journal of Human Resources, 45(3), 655-681. https://doi.org/10.1353/jhr.2010.0023 Cochran, W. G. (1965). Query 12: Testing two correlated variances. American Society for Quality, 7(3), 447-449. https://doi.org/10.2307/1266603 Cohen, J. (1992). Statistical power analysis. Current Directions in Psychological Science, 1(3), 98-101. https://doi.org/10.1111/1467-8721.ep10768783 Croninger, R. G., Rice, J. K., Rathbun, A., & Nishio, M. (2007). Teacher qualifications and early learning: effects of certification, degree, and experience on first-grade student achievement. Economics of Education Review, 26(3), 312-324. https://doi.org/10.1016/j.econedurev.2005.05.008 Darling-Hammond, L. (1995). Inequality and access to knowledge. In E. J. A. Banks & C. A. M. Banks (Eds.), Handbook of research on multicultural (pp. 465-483). Simon & Schuster Macmillan. Darling-Hammond, L. (2000). Teacher quality and student achievement: A review of state policy evidence. Education Policy Analysis Archives, 8(1), 1-44. https://doi.org/10.14507/epaa.v8n1.2000 Fetler, M. (1999). High school staff characteristics and mathematics test results. Education Policy Analysis Archives, 7(9), 1-19. https://doi.org/10.14507/epaa.v7n9.1999 Fukuya, K. (2015). Tokai chihono jichitaini okeru shochugakkono kyoinno jinji idono keiko (Tendency of the regular personnel transfers for elementary and junior high school How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 21 teachers in the Tokai district). Journal of the School of Education, Sugiyama Jogakuen University, 8, 207-216. Gupta, R. C., & Ma, S. (1996). Testing the equality of coefficients of variation in K normal populations. Communications in Statistics: Theory and Methods, 25(1), 115-132. https://doi.org/10.1080/03610929608831683 Hanushek, E. A., Kain, J. F., O'Brien, D. M., & Rivkin, S. G. (2005). The market for teacher quality (NBER Working Paper Series Working Paper 11154). http://www.nber.org/papers/w11154.pdf https://doi.org/10.3386/w11154 Harris, D. N., & Sass, T. R. (2011). Teacher training, teacher quality and student achievement. Journal of Public Economics, 95, 798-812. https://doi.org/10.1016/j.jpubeco.2010.11.009 Haynes, J. D. (1981). Statistical simulation study of new proposed uniformity requirement for bioequivalency studies. Journal of Pharmaceutical Sciences, 70(6), 673-675. https://doi.org/10.1002/jps.2600700625 Imberman, S. (2009). The effect of charter schools on achievement and behavior of public school students. https://doi.org/10.2139/ssrn.1031693 Ishikida, M. (2005). Japanese education in the 21st century. iUniverse, Inc. Jackson, C. K. (2012). School competition and teacher labor markets: Evidence from charter school entry in North Carolina. Journal of Public Economics, 96, 431-448. https://doi.org/10.1016/j.jpubeco.2011.12.006 Jacob, B. A., & Lefgren, L. (2008). Can principals identify effective teachers? Evidence on subjective performance evaluation in education. Journal of Labor Economics, 26(1), 101-136. https://doi.org/10.1086/522974 Janda, K. (2015). Computing standard deviations for proportions. Northwestern University. http://www.janda.org/c10/Lectures/topic07/stdevproportions.htm Kariya, T. (2011). Japanese solutions to the equity and efficiency dilemma? Secondary schools, inequity and the arrival of “universal” higher education. Oxford Review of Education, 37(2), 241-266. https://doi.org/10.1080/03054985.2011.559388 Kawakami, Y. (2013). Kouritsu gakkou no kyouin jinji shisutemu (Public schools' teacher personnel system). Gakujutsu Shuppankai. Kawakami, Y., & Senoh, W. (2011). Kyouin no idou kenshuu ga nouryokukaihatsu ni oyobosu chokusetsuteki kansetsutekikeiro ni tsuiteno kousatsu (The direct and indirect effects of training and personnel changes on the capacity building of teachers). Journal of the Faculty of Culture and Education, Saga-University, 16(1), 1-20. Kim, J.-O., & Rabjohn, J. (1980). Binary variables and index construction. Sociological Methodology, 11, 120-159. https://doi.org/10.2307/270861 Kim, Y. K., Kim, A., Park, S. J., & Lee, H. (2017). New tablet formulation of tacrolimus with smaller interindividual variability may become a better treatment option than the conventional capsule formulation in organ transplant patients. Drug Design, Development and Therapy, 11, 2861-2869. https://doi.org/10.2147/DDDT.S142201 Klitgaard, R. E., & Hall, G. R. (1975). Are there unusually effective schools? Journal of Human Resources, 10(1), 90-106. https://doi.org/10.2307/145121 Ladd, H. F., & Sorensen, L. C. (2015). Do master's degrees matter? Advanced degrees, career paths, and the effectiveness of teachers (CALDER Working Paper No. 136). https://caldercenter.org/sites/default/files/WP 136_0.pdf Lankford, H., Loeb, S., & Wyckoff, J. (2002). Teacher sorting and the plight of urban schools: A descriptive analysis. Educational Evaluation and Policy Analysis, 24(1), 37-62. https://doi.org/10.3102/01623737024001037 Education Policy Analysis Archives Vol. 29 No. 91 22 Letendre, G. K. (2000). Learning to be adolescent: Growing up in U.S. and Japanese middle schools. Yale University Press. Liu, X., Xu, X., & Zhao, J. (2011). A new generalized P-value approach for testing equality of coefficients of variation in K normal populations. Journal of Statistical Computation and Simulation, 81(9), 1121-1130. https://doi.org/10.1080/00949651003724790 Marwick, B. (2018). Testing coefficients of variation from multiple samples. The Comprehensive R Archive Network. https://cran.r- project.org/web/packages/cvequality/vignettes/how_to_test_CVs.html Masuda, M. (2003). Daigaku-wa gakusei-ni nani-ga dekiru-ka? "Gakusei-wo genki-ni saseru" daigaku kaikaku-towa [What can universities do for students? What university reforms can "energize students"?]. President Inc. MEXT. (2016). Improving the quality and ability of teachers. Ministry of Education, Culture, Sports, Science, and Technology. https://www.mext.go.jp/en/policy/education/elsec/title02/detail02/__icsFiles/afieldfil e/2015/08/11/1303528_01_1.pdf Monk, D. H. (1994). Subject matter preparation of secondary mathematics and science teachers and student achievement. Economics of Education Review, 13(2), 125-145. https://doi.org/10.1016/0272-7757(94)90003-5 Morgan, W. A. (1939). A test for the significance of the difference between the two variances in a sample from a normal bivariate population. Biometrika, 31(1-2), 13-19. https://doi.org/10.1093/biomet/31.1-2.13 Mudholkar, G. S., Wilding, G. E., & Mietlowski, W. L. (2003). Robustness properties of the Pitman-Morgan Test. Communication in Statistics: Theory and Methods, 32(9), 1801-1816. https://doi.org/10.1081/STA-120022710 Murnane, R. J., & Phillips, B. R. (1981). Learning by doing, vintage, and selection: Three pieces of the puzzle relating teaching experience and teaching performance. Economics of Education Review, 4(1), 453-465. https://doi.org/10.1016/0272-7757(81)90015-7 Newkirk, M. (2014, February 6). Baton Rouge's rich want new town to keep poor pupils out: Taxes. Bloomberg. Retrieved from https://www.bloomberg.com/news/articles/2014 -02- 06/baton-rouge-s-rich-want-new-town-to-keep-poor-pupils-out-taxes Okano, K., & Tsuchiya, M. (1999). Education in contemporary Japan: Inequality and diversity. Cambridge University Press. Ono, H. (2001). Who goes to college? Features of institutional tracking in Japanese higher education. American Journal of Education, 109(2), 161-195. https://doi.org/10.1086/444265 Ouchi, W. G. (1981). Theory Z: How American business can meet the Japanese challenge. Addison- Wesley. https://doi.org/10.1016/0007-6813(81)90031-8 Peske, H. G., & Haycock, K. (2006). Teaching inequality: How poor and minority students are shortchanged on teacher quality. Education Trust. PISA. (2010). Results, Volume II, Overcoming social background: Equity in learning opportunities and outcomes. Organisation for Economic Co-operation and Development (OECD): Programme for International Student Assessment . https://www.oecd- ilibrary.org/education/pisa-2009-results-overcoming-social- background_9789264091504-en Pitman, E. J. G. (1939). A note on normal correlation. Biometrika, 31(1-2), 9-12. https://doi.org/10.1093/biomet/31.1-2.9 How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 23 Prince, C. D. (2002). The challenge of attracting good teachers and principals to struggling schools. American Association of School Administrators. Puhani, P. A. (2012). The treatment effect, the cross difference, and the interaction term in nonlinear “difference-in-differences” models. Economics Letters, 115(1), 85-87. https://doi.org/10.1016/j.econlet.2011.11.025 Rice, J. K. (2003). Teacher quality: Understanding the effectiveness of teacher attributes. Economic Policy Institute. http://www.epi.org/publication/books_teacher_quality_execsum_intro/ Rivers, J. C., & Sanders, W. L. (2002). Teacher quality and equity in educational opportunity: Findings and policy implications. In L. T. Izumi & W. M. Evers (Eds.), Teacher Quality (pp. 13-23). Hoover Institution Press. http://www.hoover.org/sites/default/files/uploads/documents/0817929320_13.pdf Rivkin, S. G., Hanushek, E. A., & Kain, J. F. (2005). Teachers, schools, and academic achievement. Econometrica, 73(2), 417-458. https://doi.org/10.1111/j.1468- 0262.2005.00584.x Rohlen, T. (1983). Japan's high schools. University of California Press. https://doi.org/10.1525/9780520341302 Rowan, B., Chiang, F., & Miller, R. J. (1997). Using research on employees' performance to study the effects of teachers on students' achievement. Sociology of Education, 70, 256-284. https://doi.org/10.2307/2673267 Sanders, W. L., & Rivers, J. C. (1996). Cumulative and residual effects of teachers on future student academic achievement. University of Tennessee, Value-Added Research and Assessment Center. Seebruck, R. (2015). Teacher quality and student achievement: A multilevel analysis of California's school districts. McGill Sociological Review, 5(July), 1-18. Seebruck, R. (2016). Examining the maldistribution in teacher quality: A spatial analysis of the distribution of credentialed educators in California schools. California Geographer, 55, 55- 82. Seebruck, R. (2019). A case study of how education labor markets are organized in Japan: Mandatory teacher rotation in Japanese public schools. Asia Pacific Journal of Education, 39(3), 323-337. https://doi.org/10.1080/02188791.2019.1598847 Shimano, K. (2009). Abunai daigaku kieru daigaku '10 (Going broke universities, disappearing universities '10). YELL books. Shoukri, M. M., Colak, D., Kaya, N., & Donner, A. (2008). Comparison of two dependent within subject coefficients of variation to evaluation the reproducibility of measurement devices. BMC Biomedical Research Methodology, 8, 24-35. https://doi.org/10.1186/1471- 2288-8-24 Snedecor, G. W., & Cochran, W. G. (1989). Statistical methods (8th ed.). Iowa State University Press. Sokal, R. R., & Braumann, C. A. (1980). Significance tests for coefficients of variation and variability profiles. Systematic Zoology, 29(1), 50-66. https://doi.org/10.2307/2412626 Statistics Japan. (2013). Japan statistical yearbook 2013. Statistics Bureau of Japan. http://www.stat.go.jp/english/data/nenkan/back62/index.htm Summers, A. A., & Wolfe, B. L. (1975). Which school resources help learning? Efficiency and equality in Philadelphia public schools. Federal Reserve Bank of Philadelphia. Retrieved from http://files.eric.ed.gov/fulltext/ED102716.pdf Education Policy Analysis Archives Vol. 29 No. 91 24 Verrill, S. (2009). A likelihood ratio test of the equality of the coefficients of variation of k normally distributed populations. Forest Products Laboratory, U.S. Forest Service, U.S. Department of Agriculture. http://www1.fpl.fs.fed.us/covtestk.html Wenglinsky, H. (2002). How schools matter: The link between teacher classroom practices and student academic performance. Education Policy Analysis Archives, 10(12), 1-30. https://doi.org/10.14507/epaa.v10n12.2002 White, M. (1987). The Japanese educational challenge: A commitment to children. The Free Press. Wilcox, R. (2015). Comparing the variances of two dependent variables. Journal of Statistical Distributions and Applications, 2(7), 1-8. https://doi.org/10.1186/s40488-015-0030-z Wray, H. (1999). Japanese and American education: Attitudes and practices. Bergin and Garvey. Wright, S. P., Horn, S. P., & Sanders, W. L. (1997). Teacher and classroom context effects on student achievement: Implications for teacher evaluation. Journal of Personnel Evaluation in Education, 11, 57-67. https://doi.org/10.1023/A:1007999204543 How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 25 Appendix: Sensitivity Analyses Introduction When comparing two groups composed of potentially different subjects there is concern that population differences may affect the comparison. There are ways to address this concern, such as comparing across groups differences which exist within groups or by using relative variation measures like the coefficient of variation. I employed both techniques above. Here I also employ randomized reassignment and difference-in-difference (DiD) models. Past education research has employed random assignment methods (Imberman, 2009), and the implementation here is similar to DiD models common to experimental research (Abadie et al., 2010; Buckley & Shang, 2003; Card & Krueger, 1994; Puhani, 2012). A DiD model is essentially a fixed-effects model that “allows for the presence of unobserved confounders but restricts the effect of those confounders to be constant” (Abadie et al., 2010, p. 495). A common set up is where outcomes are recorded for two groups across two time periods. The treatment group is exposed to an intervention between Time 1 and Time 2, but the control group is not. Because the same subjects are recorded at the start and end of the experiment, the change in scores from Time 1 to Time 2 in the control group can be subtracted from the change in scores for the treatment group. This mitigates concerns of biases stemming from group differences and enables researchers to isolate the effect of the treatment. As an example of this in the education literature, Jackson (2012) utilizes DiD models to compare the change in outcomes within schools after they become a charter school to the change in outcomes within schools that remain non- charter schools over that same period. In this sense, becoming a charter school is the treatment, those schools that became charter schools compose the treatment group and those schools that did not become charter schools compose the control group. Accordingly, in this study, to conduct these sensitivity analyses requires four data sets, each with two groups representing the treatment and control: 1. the original data set (i.e. the actual data for the public and private sectors); 2. the public sector data set (i.e. the actual data for public schools plus a randomized version of them), 3. the private sector data set (i.e. the actual data for private schools plus a randomized version of them); and 4. the randomized data set (i.e. the randomized versions of the public and private school data sets). The first data set was used in the main analysis above (see Table 2). The second and third ones are used in the DiD analyses, and the fourth one is used in the randomized reassignment analyses. Here the treatment is the different organizational structures of the teacher labor markets. In the public school data set this is the top-down allocation of teachers to schools by the prefectural government, whereas in the private school data set the treatment is ‘teacher choice’. That is, compared to a random pool of teachers in that labor market (i.e. Time 1), private school teachers choosing where they work represents the outcome (i.e. Time 2) whereas in the public sector, compared to a random pool of teachers in that labor market (i.e. Time 1), the prefectural board of education deciding where teachers work represents the outcome (i.e. Time 2). Following that, comparing the different distributions of these randomized and non-randomized groups sheds light on how these different treatments impact teacher distribution. Difference-in-Differences To setup the DiD analyses, I first compare the between-school differences in teacher quality variation among the randomized and actual samples in each sector, similar to the analyses in Table 2. However, unlike the analyses in Table 2, because the randomized and actual data sets comprise the exact same pool of teachers it is not necessary to use relative standard variation. Instead, the regular Education Policy Analysis Archives Vol. 29 No. 91 26 standard deviation is used, and with that, a shift from the likelihood-ratio (LR) and Pitman-Morgan (PM) significance tests to Bartlett’s test for the homogeneity of sector-level grand standard deviations. The Bartlett tests find that, for every teacher quality measure, the actual distribution of teachers in both the public and private sector are significantly (p < .000) more variable than random (results not shown). These results are not surprising, as humans do not randomly decide where to work (private sector) or where others should work (public sector). Instead, these results serve to justify the DiD analyses comparing inter-sector differences in the magnitude of these intra-sector variances in teacher quality distribution. The DiD analyses shown in Table 3 are similar to past research (see Card & Krueger, 1994, p. 780). In short, they reveal the inter-sector (i.e. public vs. private) difference between the intra- sector (i.e. randomized vs. actual) differences in the between-school variation of each teacher quality measure. The results show, with one exception, the differences between the actual and the randomized public school sample are more favorable, from an educational egalitarianism perspective, compared to the differences between the actual and the randomized private school sample. The one exception is the percent of novices, which has a larger DiD in the public sector— something unsurprising given the need for the public sector to staff rural schools (Wray, 1999). Table 3 Magnitude of the Difference in Differences between Measures of Variation in the Actual and Randomized Samples Teacher Quality Vars. Measures ABS (PrivRRANDOMR - PrivRACTUALR) ABS (PubRRANDOMR - PubRACTUALR) PrivRDIFFR - PubRDIFF Outcome of CV Comparison Full-Time % SD 0.1097 0.0921 0.0176 Private more variable than expected, compared to public CV 0.1500 0.1043 0.0457 Highly Credentialed % SD 0.0994 0.0622 0.0372 Private more variable than expected, compared to public CV 0.8144 0.3827 0.4317 Years of Experience SD 3.3731 3.8891 -0.5160 Private more variable than expected, compared to public CV 0.2088 0.2011 0.0077 Novices % SD 0.0950 0.0682 0.0268 Public more variable than expected, compared to private CV 0.6467 0.7860 -0.1393 In-Field % SD 0.0513 0.0391 0.0122 Private more variable than expected, compared to public CV 0.0592 0.0426 0.0166 Graduate Degree % SD 0.0927 0.0703 0.0224 Private more variable than expected, compared to public CV 0.6443 0.4266 0.2177 Alma Mater Prestige SD 0.6649 0.4307 0.2342 Private more variable than expected, compared to public CV 0.0862 0.0495 0.0367 Teacher Quality Index SD 0.4531 0.2801 0.1730 Private more variable than expected, compared to public CV 0.1410 0.0731 0.0679 Note. Absolute Values (ABS) of the Difference (DIFF) between the standard deviations (SD) and coefficient of variations (CV) for the randomized and actual samples, for the private (Priv) and public (Pub) sectors; unweighted data, N = 49 (23 Private, 26 Public), proportionally allocated groups. How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 27 Overall, the DiD analyses support the proposition that the centrally controlled public sector, with its mandatory teacher rotation system, distributes teachers more equitably across schools than the decentralized private sector. Randomized Reassignment In addition to the DiD analyses, randomized reassignment analyses can provide further illumination. These analyses mimic those in Table 2, but instead of comparing the actual public and private school data, they compare the randomized versions of them. If self-selection bias is a concern, the outcomes of the randomized data set comparisons (Table 4) should be identical to those of the actual data sets (Table 2). As a comparison of different labor pools, LR and PM tests are once again used. Recall that seven out of the eight measures in the actual analyses revealed that the private sector was significantly more variable than the public sector. If these differences were due to self-selection bias rather than the treatment (i.e. jinji idou), then the randomized analyses should corroborate these results. However, as seen in Table 4, some of the results are counter to this expectation, meaning that self-selection bias is not a concern for these measures. Table 4 Inter-Sector Comparison of Intra-Sector, Between-School Differences in Teacher Quality among Randomized Samples Likelihood Ratio and Pitman-Morgan Tests for the Equality of Sector-Level Grand Coefficients of Variation Teacher Quality Vars. Random CV LR Test p-value PM Test p-value Outcome of Sig. Test Full-Time % Private 0.0460 0.0000*** 0.0000*** Private sector is significantly more variable Public 0.0147 Highly Credentialed % Private 0.0978 0.7111 0.5783 No significant difference between sectors Public 0.0904 Years of Experience Private 0.0342 0.9588 0.9258 No significant difference between sectors Public 0.0338 Novices % Private 0.1399 0.3907 0.2130 No significant difference between sectors Public 0.1668 In-Field % Private 0.0136 0.6040 0.4451 No significant difference between sectors Public 0.0122 Graduate Degree % Private 0.0775 0.1443 0.0367* Mixed support: Public sector may be more variable Public 0.1046 Alma Mater Prestige Private 0.0226 0.0016** 0.0000*** Private sector is significantly more variable Public 0.0117 Teacher Quality Index Private 0.0288 0.0001*** 0.0000*** Private sector is significantly more variable Public 0.0930 Note. Private sector randomized sample vs. public sector randomized sample; unweighted data, 2-tailed tests, N = 49 (23 Private, 26 Public), df = 43, p < .05* p < .01** p < .001*** Education Policy Analysis Archives Vol. 29 No. 91 28 Note that three of the eight measures are significantly more variable in the private sector for the randomized analyses: full-time teacher status, alma mater prestige, and the Teacher Quality Index (TQI). This suggests that the differences found in the actual analyses may be—but are not necessarily—due to self-selection bias for these measures. In contrast, the remaining five teacher quality measures reveal that self-selection bias is unlikely, as the randomized analyses saw either no significant difference between sectors—as is the case with credentialization, years of experience, novices, and in-field teaching status—or, even more interestingly, showed the randomized version of the public sector had a significantly larger between-school variance in the distribution of teachers holding an advanced degree, which is the opposite of the actual analyses. These results suggest that, compared to the more laissez-faire education labor market of the private sector, the teacher rotation system in the public sector distributes teachers more equitably across schools for most of the teacher quality variables identified in the main analyses as having different distributions (credentialization, experience, in-field teaching, and advanced degree). Consequently, given the results of the DiD and randomized reassignment analyses, it is unlikely that the results of the inter-sector comparative analyses—which largely found that teacher quality distribution was more clustered in the private sector compared to the public sector—were the consequence of self-selection bias rather than organizational differences such as jinji idou. How t eacher rotation in Japanese high schools affects the clustering of t eacher quality 29 About the Author Ryan Seebruck Independent Scholar ryan.seebruck@outlook.com https://orcid.org/0000-0002-3461-5658 Ryan is a statistician and independent scholar. He earned his PhD in sociology from the University of Arizona in 2015. His research primarily focuses on equality in educational systems and has appeared in journals such as Social Problems, Sociological Quarterly, Sociological Focus, Sociological Inquiry, and Sociological Spectrum, among others. education policy analysis archives Volume 29 Number 91 July 5, 2021 ISSN 1068-2341 Readers are free to copy, display, distribute, and adapt this article, as long as the work is attributed to the author(s) and Education Policy Analysis Archives, the changes are identified, and the same license applies to the derivative work. More details of this Creative Commons license are available at https://creativecommons.org/licenses/by-sa/4.0/. EPAA is published by the Mary Lou Fulton Institute and Graduate School of Education at Arizona State University Articles are indexed in CIRC (Clasificación Integrada de Revistas Científicas, Spain), DIALNET (Spain), Directory of Open Access Journals, EBSCO Education Research Complete, ERIC, Education Full Text (H.W. Wilson), QUALIS A1 (Brazil), SCImago Journal Rank, SCOPUS, SOCOLAR (China). About the EPAA/AAPE Editorial Team: https://epaa.asu.edu/ojs/about/editorialTeam Please send errata notes to Audrey Amrein-Beardsley at audrey.beardsley@asu.edu Join EPAA on Facebook https://www.facebook.com/EPAAAAPE and Twitter @epaa_aape mailto:ryan.seebruck@outlook.com https://orcid.org/0000-0002-3461-5658 https://creativecommons.org/licenses/by-sa/4.0/ http://www.doaj.org/ http://www.doaj.org/ https://epaa.asu.edu/ojs/about/editorialTeam mailto:audrey.beardsley@asu.edu https://www.facebook.com/EPAAAAPE