Conference Agenda
Overview and details of the sessions of this conference. Please select a date or location to show only sessions at that day or location. Please select a single session for detailed view (with abstracts and downloads if available).
|
Daily Overview |
| Session | ||
09 SES 15 C: Grading Fairness, Assessment Load and Student Perceptions
Paper Session | ||
| Presentations | ||
09. Assessment, Evaluation, Testing and Measurement
Paper How do Gender, Cognitive Ability and Parental Education Predict Students' Perceptions of Being Graded University of Gothenburg, Sweden Presenting Author:This study aims to increase understanding of how grades affect students’ motivation and well-being by examining students’ own perceptions of how grades influence them in school. The study investigates (1) the role of student background characteristics (gender, cognitive ability, and parental education) in perceptions of grades as motivating or as sources of stress and anxiety, and (2) the extent to which these perceptions are related to students’ academic achievement. Grades constitute a central component of all European education systems, although their design vary. When policies concerning how and when grades should be applied in schools are established, they are partly based on assumptions about how grades affect students’ behavior. However, research presents mixed conclusions regarding the impact of grades on students. Some studies conclude that grades, particularly those of a high-stakes character, have a positive effect on students’ academic achievement; students are expected to exert greater effort and devote more time to schoolwork when they are aware that their knowledge will be evaluated (e.g. Hvidman & Sievertsen, 2021). Other studies, however, suggest that grades are negatively related to future achievements (Klapp, 2015) and associated with reduced well-being and lower academic self-concept among students (Cashman et al., 2023; Linder et al., 2025). Drawing on Pekrun’s (2006) control-value theory, this study focuses on the emotions and perceptions that students themselves attribute to grades. Such emotions are a key part of understanding how assessments can be expected to affect students’ learning and well-being in school. In control-value theory, achievement emotions arise before, during, and after engaging in activities that are assessed. Positive activating emotions (e.g., enjoyment) are believed to facilitate students’ future performance by strengthening academic self-concept and intrinsic motivation, while successful performance, in turn, promotes positive activating emotions. Negative emotions (e.g., anxiety) are expected to impair performance, partly because coping with these emotions requires cognitive resources or leads students to allocate their resources to domains other than schoolwork (Pekrun et al., 2023). Qualitative research examining students’ perceptions of grading shows substantial variation between individuals. Some students perceive grades as motivating, “a push in the right direction”, whereas others primarily experience them as distracting or as sources of stress and anxiety. Furthermore, these perceptions may coexist within the same individual (Costa et al., 2024; Liljeröd et al., 2025). However, how common these perceptions are, and how they differ across groups of students, is not well understood. A large body of research examining the effects of assessment on students has shown that differences can be related to student background characteristics. For example, associations between grades and reduced perceived well-being as well as academic self-concept appear to be stronger among girls than among boys (Linder et al., 2025). Studies identifying negative associations between assessment and motivation often find these associations to be more pronounced among low-achieving students (Harlen & Deakin Crick, 2002; Klapp, 2015). It is therefore reasonable to expect that students’ perceptions of how grades affect them vary across student groups. Methodology, Methods, Research Instruments or Sources Used The study uses survey and register data from upper secondary school students born in 2004 (n = 9437). The data were obtained from the Evaluation Through Follow-up database, a longitudinal research project that has collected data from several birth cohorts since 1961 (Härnqvist, 2000). The database also includes results from a cognitive ability test as well as register data from Statistics Sweden, including parental education (measured in five categories) and students’ grade point average (GPA) at the completion of upper secondary education. Confirmatory factor analysis was used to establish two latent factors based on six survey items measuring students’ perceptions of being affected by grades. Students reported the extent to which grades (1) stressed them, (2) made them feel bad, (3) made them lose their desire to learn, (4) encouraged them to perform better, (5) made them work harder to learn, and (6) made them “sharpen up” and focus more. Items 1–3 constituted the factor Negative Emotions, while items 4–6 constituted the factor External Motivator. Structural equation modelling was subsequently used to examine whether the two factors were associated with students’ gender, cognitive ability, and parental education. These background variables were related to the two factors individually in separate models and jointly in a combined model. To examine potential interactions between the background characteristics, interaction terms (Gender × Cognitive Ability, Gender × Parental Education, Cognitive Ability × Parental Education, and Gender × Cognitive Ability × Parental Education) were created and related to the two factors. In a final step, associations between the two factors and students’ GPA were examined, with the three background variables included as control variables. Conclusions, Expected Outcomes or Findings The models showed that girls were clearly overrepresented in the Negative Emotions factor (β = .361). Negative Emotions was also weakly negatively related to cognitive ability (β = −.119) and parental education (β = −.102). For the External Motivator factor, no significant associations were found with gender or cognitive ability, and only a weak association with parental education (β = −.119). However, the models indicated complex interaction effects for the association between cognitive ability and External Motivator, which seemed to vary by gender and by levels of parental education. The strongest positive association with External Motivator was observed among girls with high cognitive ability and highly educated parents, whereas the corresponding association appeared to be negative among girls with high cognitive ability but lower-educated parents. Furthermore, the models showed that Negative Emotions was negatively associated with GPA (β = −.278), whereas External Motivator was positively associated with GPA. However, the cross-sectional design does not allow conclusions about whether students’ perceptions of grades predict achievement or vice versa. As shown by Pekrun et al. (2023), these associations are likely to be reciprocal. The results can be understood in light of previous research demonstrating the differentiating effects of grading systems on students (Klapp, 2015). Students who possess combinations of characteristics that are already associated with higher academic achievement are those most likely to perceive grades as motivating, and this perception is positively related to academic performance. References Cashman, M., Strandh, M., & Högberg, B. (2023). Have performance-based educational reforms increased adolescent school-pressure in Sweden? A synthetic control approach. International Journal of Educational Development, 103, 102922. https://doi.org/10.1016/j.ijedudev.2023.102922 Costa, S., Norton, L. S., & Pirchio, S. (2024). Discourses about grades and competency-based evaluation: Exploring communicative and situated meanings at an Italian high school. Social Psychology of Education, 27(5), 2177–2198. https://doi.org/10.1007/s11218-024-09911-5 Harlen, W., & Deakin Crick, R. (2002). A systematic review of the impact of summative assessment and tests on students’ motivation for learning. Research Evidence in Education Library, 2002(1). Härnqvist, K. (2000). Evaluation through follow-up: A longitudinal program for studying education and career development. In C.-G. Janson (Ed.), Seven Swedish longitudinal studies in behavioral science (pp. 76–114). Swedish Council for Planning and Coordination of Research. Hvidman, U., & Sievertsen, H. (2021). High-Stakes Grades and Student Behavior. Journal of Human Resources, 56(3), 821–849. https://doi.org/10.3368/jhr.56.3.0718-9620R2 Klapp, A. (2015). Does grading affect educational attainment? A longitudinal study. Assessment in Education: Principles, Policy & Practice, 22(3), 1–22. https://doi.org/10.1080/0969594X.2014.988121 Liljeröd, H., Jönsson, A., Klapp, A., & Jonsson, A.-C. (2025). Students’ perceptions of how grades influence their motivation: Voices of upper secondary school students in Norway and Sweden. Educational Assessment, Evaluation and Accountability, 1–25. https://doi.org/10.1007/s11092-025-09454-z Linder, A., Gerdtham, U.-G., & Heckley, G. (2025). Adolescent Mental Health: Impact of Introducing Earlier Compulsory School Grades. Health Economics, 34(9), 1731–1746. https://doi.org/10.1002/hec.4982 Pekrun, R. (2006). The control-value theory of achievement emotions: Assumptions, corollaries, and implications for educational research and practice. Educational Psychology Review, 18(4), 315–341. https://doi.org/10.1007/s10648-006-9029-9 Pekrun, R., Marsh, H. W., Suessenbach, F., Frenzel, A. C., & Goetz, T. (2023). School grades and students’ emotions: Longitudinal models of within-person reciprocal effects. Learning and Instruction, 83, 101626. https://doi.org/10.1016/j.learninstruc.2022.101626 09. Assessment, Evaluation, Testing and Measurement
Paper Are Assessments in Higher Education Fair for All? Investigating the Association be-tween Assessment Load and Type and Students' Academic Performance University of Leeds, United Kingdom Presenting Author:Background In contemporary European higher education, the pursuit of “social justice” has transcended the initial goal of widening participation to focus on the equity of student outcomes. The European Higher Education Area (EHEA), through successive ministerial communiqués, has consistently emphasised the “social dimension” of higher education, positing that the student body should reflect the diversity of our populations. However, persistent awarding gaps remain a pan-European challenge. Marginalised ethnic (BAME) groups and students from lower socio-economic status (SES) backgrounds continue to experience systemic underperformance (Mountford-Zimdars et al., 2015; OfS, 2023). This study argues that these disparities are not the result of individual student “deficits” but are structurally embedded in the assessment infrastructures of modern universities. While extensive literature documents performance disparities (Canal & Child, 2025; Yu et al., 2023), there is a dearth of large-scale empirical evidence exploring how the underlying assessment types are associated with students' educational outcomes. This study challenges the “neutrality” of assessment, suggesting that different ways of evaluating knowledge act as structural filters that exclude certain groups of students (McArthur, 2016). Furthermore, it questions the current international policy trend toward reducing assessment loads. In contemporary discourses, lightening the load is advocated to improve student well-being, yet the impact of such reductions on equity remains under-theorised and empirically contested. Theoretical Framework This study is grounded in the tripartite taxonomy of assessment standards discourses proposed by Ajjawi et al. (2021):
This research explores how institutional habitus shapes the awarding gap. The central research questions are: 1. To what extent does assessment load and different assessment discourses (representational, sociocultural, and sociomaterial) associate with students’ academic performance across the university? 2. How does the association between assessment characteristics and academic performance vary systematically across different student subgroups, specifically in terms of ethnicity and socioeconomic status? Methodology, Methods, Research Instruments or Sources Used Methods This study conducts a systematic empirical enquiry using an administrative longitudinal dataset from a research-intensive Russell Group university in the UK. The dataset tracks 52,420 students across 3,606 modules over an eight-year period (2017/18–2024/25). This dataset is comprehensive with information of anonymised student characteristics (sex, ethnicity, POLAR4 quintiles, UCAS Tariff points), and module-level details (credits, assessment types, syllabus and contact hours), which is suitable for a student-module-year level analysis. Data Operationalisation To bridge theory and data, we manually coded 60 unique institutional assessment labels (e.g., timed examination, portfolio, fieldwork report) into the three discourses defined by Ajjawi et al. (2021). We conducted the Inter-Rater Reliability to ensure the rigours of classification. Assessment load was normalised by calculating the number of assessment points per 10 module credits, ensuring a standardised measure for cross-disciplinary comparison. Demographic variables included sex, ethnicity (7-way classification), and SES (proxied by POLAR4 quintiles). Statistical Modelling Given the nested multilevel structure of educational data, where students are nested within modules, and modules within academic schools, we employed Linear Mixed-Effects Models (LMM). This approach is methodologically necessary to partition variance accurately. Intraclass Correlation Coefficients (ICCs) revealed that 33.1% of outcome variation was situated at the student level and 9.9% at the module level, justifying the multilevel approach. The significant results of Likelihood Ratio Tests (LRT) also confirmed the necessity to employ multilevel modelling. Intersectional Analysis A core strength of this methodology is the use of interaction terms between assessment predictors and subgroup indicators. This allowed us to test for heterogeneity of effects, specifically, how the impact of a particular assessment discourse reverses or amplifies depending on a student’s ethnic and socioeconomic background. This rigorous analytical framework ensures that our findings are accessible and robust for an international audience of educational researchers. Conclusions, Expected Outcomes or Findings Findings and Conclusion The findings identify a striking “direction-reversal” effect. While representational assessments, such as exams, yield marginal advantages for White and high-SES students, they significantly undermine the performance of BAME and lower-SES groups. By rewarding inherited cultural capital, these designs act as a double-gatekeeper that reinforces systemic awarding gaps across the academy. Conversely, sociomaterial discourse offers a transformative “amplification effect”. The positive association between practice-based, relational tasks and performance is significantly stronger for Black students and those from lower-participation backgrounds. This suggests that when assessment focuses on material interaction rather than abstract knowledge, it validates alternative ways of knowing and alleviates the “belonging uncertainty often felt by marginalised learners (Walton & Cohen, 2007). One of the most counterintuitive outcomes is the scaffolding effect of load. While modern policies often advocate for reducing assessment, our results show that high assessment frequency yields significantly greater benefits for disadvantaged students. For underrepresented groups, frequent assessment functions as a navigational roadmap (Harland et al., 2015), providing structured feedback loops that help them decode the “rules of the game” in a high-stakes environment (Jorre & Boud, 2022). These findings contribute to the European dialogue on Inclusive Assessment. They suggest that a unilateral emphasis on reducing burden may inadvertently remove the feedback structures that support equity. Furthermore, it is also vital to advocate for a pedagogical balance that maintains rigorous standards while dismantling structural barriers (Tai et al., 2023). This research provides a robust evidence base for European policymakers to move beyond student-focused support toward a critical redesign of institutional assessment infrastructures, ensuring that assessment serves as a bridge, not a barrier, to success. References References •Ajjawi, S., et al. (2021). Assessment and the sociomaterial. Assessment & Evaluation in Higher Education. •Bourdieu, P. (1986). The forms of capital. In J. G. Richardson (Ed.), Handbook of Theory and Research for the Sociology of Education. •Canal, M. M., & Child, R. (2025). Impact of the type of assessment on awarding gaps in bioscience undergraduate degrees. Assessment & Evaluation in Higher Education. •Fenwick, T., & Edwards, R. (2013). Performative ontologies: Sociomaterial approaches to researching adult education and lifelong learning. European Journal for Research on the Education and Learning of Adults. •Harland, T., et al. (2015). An assessment arms race and its fallout: high-stakes grading and the case for slow scholarship. Assessment & Evaluation in Higher Education. •Jorre de St Jorre, T., & Boud, D. (2022). Understanding the game: The role of assessment in equity. Teaching in Higher Education. •McArthur, J. (2016). Assessment for social justice: the role of assessment in achieving social justice. Assessment & Evaluation in Higher Education. •McArthur, J. (2023). Rethinking authentic assessment: work, well-being, and society. Higher Education. •Mountford-Zimdars, A., et al. (2015). Causes of Differences in Student Outcomes. HEFCE. •Reay, D. (2018). Miseducation: Inequality, education and the working classes. Policy Press. •Tai, J., et al. (2023). Assessment for inclusion: rethinking contemporary strategies in assessment design. Higher Education Research & Development. •Walton, G. M., & Cohen, G. L. (2007). A question of belonging: Race, social fit, and achievement. Journal of Personality and Social Psychology. •Yu, D., et al. (2023). Do assessment loads affect student academic success? South African Journal of Higher Education. 09. Assessment, Evaluation, Testing and Measurement
Paper Moderating the Big-Fish-Little-Pond Effect: The Role of School Grades in the Absence of External Testing Charles University, Faculty of Education, Czech Republic (Czechia) Presenting Author:Academic self-concept (ASC), i.e., self-related thoughts and beliefs about the ability to succeed in academic tasks and achievement situations (Bong & Skaalvik, 2003), is considered a prerequisite of achievement motivation (Shavelson et al., 1976) and is, therefore, of utmost importance for students’ academic functioning. Research has shown that positive ASC is associated with higher educational achievement, more ambitious career aspirations (Bong & Skaalvik, 2003), and contributes to students’ well-being and school satisfaction (Marcionetti & Rossier, 2016). Theories of ASC have identified some key processes that shape the formation of ACS. The primary source of ASC are reflected experiences from past performance situations (Bong & Skaalvik, 2003; Shavelson et al., 1976) on the basis of which students form beliefs about their competence. Longitudinal studies have demonstrated that early academic achievement explains gains in later ASC and early ASC is associated with later gains in educational achievement, according to a so-called reciprocal effect model (Wu et al., 2021). As academic performance cannot be easily evaluated against an objective accomplishment standard, ASC is heavily influenced by a frame of reference the students use when judging their academic ability. The same achievement may, therefore, lead to different ASC levels, depending on which frame of reference the students choose. Students can use social, dimensional, or temporal frames of reference, but social comparison often serves as prominent source of information about one’s perceived competence in educational settings (Bong & Skaalvik, 2003). Marsh (1984) called the conceptual model of ASC formation in social comparison with significant others, usually classmates, the big-fish-little-pond effect (BFLPE). This effect describes that students develop poorer ASC when they attend schools or classes with higher ability level as compared to equally able students who are placed in schools or classes with lower average achievement. While there is strong evidence for the universality of BFLPE across education systems and cultural contexts (Guo et al., 2018; Seaton et al., 2009), less is known about potential variables that could moderate the negative effect of average school or class achievement on student ASC. Because ASC results from evaluative processes occurring in interactions between individual students and their learning environment, two categories of moderators can be distinguished (Schwabe et al., 2019): (1) student individual characteristics, such as gender, ethnicity, personality traits or ability level, and (2) classroom-related characteristics, such as socioemotional climate or teachers’ pedagogical approaches. Recent research suggests that both categories may play a role in moderating BFLPE (Stockus & Zell, 2024), but classroom-related factors are of particular practical relevance given their malleability and openess to change. In particular, teacher-student relationships (Schwabe et al., 2019) and individualized instruction (Roy et al., 2015) have been found to moderate BFLPE in European education systems, but moderation effect was not confirmed for teacher grading practices and the use of individualized frame of reference, i.e., evaluating student performance in terms of individual learning progress (Stockus & Zell, 2024). However, this line of research remains limited, restricting the generalizability of the findings. This submission contributes to research on BFLPE moderators by providing empirical evidence from mainstream lower secondary schools in Czechia. Within Europe, Czech education system is among the most decentralized, with schools enjoying a high degree of autonomy over curriculum organization, instructional methods, assessment practices, and their adaptation to students’ needs. In addition, external student assessment using standardized tests does not take place at primary and lower secondary levels, meaning that students rely on classroom-related feedback regarding their academic performance. Methodology, Methods, Research Instruments or Sources Used This study uses a nationally representative sample of sixth-grade students in mainstream track lower secondary schools to investigate the role of peer relationships, teacher support and grading practices as BFLPE moderators in the context of an education system where students receive feedback on their performance primarily from social comparison with their classmates and from their teachers. Data collection took place in 2023 as part of SYRI School Panel Study. Students were administered achievement tests in mathematics and Czech language, and background questionnaires that contained questions about their school experience including self-concept in mathematics and Czech language, perceived teacher support, peer relationships, and grades. The analysis excludes students with missing data on all questionnaire items and classes in which the number of students with valid questionnaire data was lower than 10. The final analytical sample comprises around 2,400 students from 133 classes (average class size is 17.5). This contribution focuses on ASC in mathematics. It was measured by four items adapted from shortened version of the Students' Approaches to Learning (SAL) questionnaire (Ropovik & Greger, 2023). Five items taken from PISA school belonging scale were used to measure peer relationships within classrooms. Nine items taken from the study by Fauth et al. (2014) were used to measure teacher support. Student achievement in mathematics was measured by a 30-item test that employed a combination of item types and student scores were estimated by a mixed 2PL/3PL IRT model. Gender and socioeconomic status were used as control variables. Doubly latent two-level structural equation models (Marsh et al., 2009) with ASC as dependent variable are used to estimate BFLPE and effects associated with student- and classroom-related variables. In the first step, individual student achievement and classroom average achievement are entered to replicate BFLPE. Next, contextual variables are introduced at the student level to evaluate their effects on ASC. Finally, interactions between classroom average achievement and contextual variables are tested. All analyses are performed in Mplus version 8.7. Conclusions, Expected Outcomes or Findings Although the sample excludes selective long academic track schools (8-year academic schools), BFLPE was replicated in the current sample, indicating that even less obvious and more subtle forms of ability grouping within the mainstream school track pose risks to students’ ASC. When individual and classroom achievement were accounted for, ASC was further predicted by positive peer-relationships, teacher support, male gender and better grades. Results of the moderation analysis are not yet available and will be presented at the conference. Preliminary analyses suggest that in addition to their main effect on ASC, grades moderate the strength of the relationship between ASC and classroom average achievement, while a similar moderation effect is not confirmed for teacher support and peer relationships. The findings suggest that in the absence of external standardized testing, school grades are particularly powerful tools in shaping students’ ASC. Apart from practical implications for the schools’ pedagogical practice, this study suggests that characteristics of education systems may help explain inconsistent findings in existing research on BLFPE moderators. References Bong, M., & Skaalvik, E. M. (2003). Academic self-concept and self-efficacy: how different are they really? Educational Psychology Review, 15(1), 1–40. Guo, J., Marsh, H. W., Parker, P. D., & Dicke, T. (2018). Cross-cultural generalizability of social and dimensional comparison effects on reading, math, and science self-concepts for primary school students using the combined PIRLS and TIMSS data. Learning and Instruction, 58, 210–219. Fauth, B., Decristan, J., Rieser, S., Klieme, E., & Büttner, G. (2014). Student ratings of teaching quality in primary school: Dimensions and prediction of student outcomes. Learning and Instruction, 29, 1–9. Marcionetti, J., & Rossier, J. (2016). Global life satisfaction in adolescence: The role of personality traits, self-esteem, and self-efficacy. Journal of Individual Differences, 37(3), 135–144. Marsh, H. W. (1984). Self-concept, social comparison, and ability grouping: A reply to Kulik and Kulik. American Educational Research Journal, 21(4), 799–806. Marsh, H. W., Lüdtke, O., Robitzsch, A., Trautwein, U., Asparouhov, T., Muthén, B., & Nagengast, B. (2009). Doubly-latent models of school contextual effects: Integrating multilevel and structural equation approaches to control measurement and sampling error. Multivariate Behavioral Research, 44, 764–802 Ropovik, I., & Greger, D. (2023). The measurement of motivation and self-concept within the Students’ approaches to learning framework. Psychology in the Schools, 60(9), 3351–3371. Seaton, M., Marsh, H. W., & Craven, R. G. (2009). Earning its place as a pan-human theory: Universality of the big-fish–little-pond effect across 41 culturally and economically diverse countries. Journal of Educational Psychology, 101(2), 403–419. Shavelson, R. J., Hubner, J. J., & Stanton, G. C. (1976). Self-concept: Validation of construct interpretations. Review of Educational Research, 46(3), 407–441. Wu, H., Guo, Y., Yang, Y., Zhao, L., & Guo, C. (2021). A meta-analysis of the longitudinal relationship between academic self-concept and academic achievement. Educational Psychology Review, 33, 1749–1778. 09. Assessment, Evaluation, Testing and Measurement
Paper Grade Inflation in Sweden and the US: A Comparative Analysis. 1: University of Turku, Finland; 2: University of Gothenburg, Sweden Presenting Author:Grades are intended to function as meaningful and legitimate indicators of students’ subject-specific knowledge. In a criterion-referenced grading system, grades should accurately reflect the extent to which students meet predefined criteria for different performance levels and thereby provide a trustworthy measure of achievement. However, an expanding body of research suggests that grade inflation has become increasingly prevalent across educational systems, raising concerns about the validity and comparability of grades as measures of student achievement. When grades no longer reliably distinguish between levels of performance, their role as signals for selection, evaluation, and accountability is weakened. Several explanations for grade inflation have been proposed, including heightened accountability metrics imposed on schools and teachers, increased competition between schools for enrollment (especially commercially managed independent schools dependent on revenue for continued operation), growing pressure from parents and students for higher grades given significant labor-market consequences associated with grades, and shifts toward more subjective assessment practices. Grade inflation accordingly has important implications at both individual and system levels. For students, inflated grades may foster unrealistic self-perceptions and complicate transitions to higher levels of education. At the system level, grade inflation can undermine equity and public trust, particularly when grading practices vary systematically across schools or student groups. Consequently, understanding the causes, patterns, and consequences of grade inflation remains a central concern in both educational research and policy. In Sweden, the introduction of a criterion-referenced grading system in the early 1990s was met with criticism, partly because all students could, in principle, attain the highest grade if the criteria were met. Grades in the Swedish system are assigned by teachers, and there is no formal external mechanism to ensure uniform interpretation of grading criteria. As a result, decisions regarding whether students meet the requirements for specific grade levels rest largely with individual teachers, increasing the risk of inequity in grading practices (Abrams, 2016; Vlachos, 2019). In the United States, grading is likewise defined by a criterion-referenced system. However, since 1968, the federal government has administered an exam every two years in reading, writing, math, and science to a sample of fourth- and eighth-grades in all 50 states. Results on this exam, called the National Assessment for Educational Progress (NAEP), or the Nation’s Report Card, have slid significantly while student grades at the secondary level, according to surveys by the National Center for Education Statistics (NCES), have climbed substantially (Sanchez & Moore, 2022). A similar pattern has emerged at the tertiary level, with results on university-entrance exams declining significantly over the past three decades while student grades at universities have climbed substantially (Fair, 2023; Nierenberg, 2023; Radavoi, Quadrelli, & Collins, 2025; Arsenault, 2026). The overarching aim of the present study is to document grade inflation in Sweden and the United States at every level where data are available, to explore explanations for this grade inflation, and to assess implications in Sweden and the United States. For comparative purposes, data for nations with little, if any, grade inflation will be presented. Methodology, Methods, Research Instruments or Sources Used We use several data sources. For Sweden, we use data from the longitudinal database Evaluation Through Follow-up (UGU), which comprises data from 11 birth cohorts, spanning individuals born between 1948 and 2010. The cohorts 1992, 1998 and 2004 will be the focus. We also use data from Gothenburg Educational Longitudinal Database (GOLD), which comprises population data for individuals born in 1972 and onwards, from age 16. In addition, we utilise data published by the Swedish National Agency for Education. The data from these different data sources include measures of cognitive ability, national test results, and teacher-assigned grades in all school subjects and for different levels in the education system, from primary to upper-secondary education. Comprehensive demographic information is also available, including school governance, parental education and income, immigration background, and student gender. Beyond achievement measures, the UGU database includes a broad range of questionnaire items capturing constructs such as cognitive and psychological well-being, academic self-concept, motivation, and interest. The survey data can be used for analysing how the different constructs are associated with subgroups of students considering their achievement levels. In the case of the United States, data from NCES and NAEP along with individual reports from universities will be deployed to document the phenomenon of grade inflation. Descriptive statistics for the measures included will be estimated for several cohorts of students in Sweden and United States and the distribution of the different grades will be analysed. Analyses will be made of UGU data to investigate how much variance in cognitive ability and national test results explain differences in specific subjects. Confirmatory factor analysis (CFA) will be conducted to create factors of the questionnaire items (indicators). Structural equation models (SEM) will be estimated to reflect the structure of the relations between the factors, cognitive ability, national test results, and GPA in order to present the amount of variance that can be explained and whether it varies over time/cohort. Conclusions, Expected Outcomes or Findings This study will present results on patterns for how grades have fluctuated over time. Preliminary results suggest that grade inflation exists in most educational systems (Radavoi, Quadrelli, & Collins, 2025). Grade inflation may exist for different reasons and may reflect macrocosmic forces having little to do with education itself and much to do with anxieties about employment, remuneration, and status. This interpretive portion of this research project constitues the heart of this undertaking and will necessitate much investigation. References Abrams, S.E. Education and the Commercial Mindset. Harvard University Press, 2016. Arsenault,M. “Harvard Considers a Proposal to Add an A+ to Help Rein in Grade Inflation,” New York Times, February 1, 2026. Fair, R. Grade Report Update: 2022-23. Yale University (November 2023). Nierenberg, A. “Excellence at Yale Doubted as Nearly Everyone Gets A’s,” New York Times, December 6, 2023. Radavoi, C., Quadrelli C., & Collins, P. “Moral Responsibility for Grade Inflation: Where Does It Lie?” Journal of Academic Ethics, Vol. 23 (2025): 1781-1798. Sanchez, E., & Moore, R. Grade Inflation Continues to Grow in the Past Decade, ACT Research, May 2022. Vlachos, J. “Trust-based evaluation in a market-oriented school system,” Neoliberalism and Market Forces in Education. Routledge, 2019. | ||