INTEGRATING AUTHENTIC AND MODEL-BASED ASSESSMENT IN ENGLISH LANGUAGE LEARNING: A COMPREHENSIVE EVALUATION FRAMEWORK FOR HIGHER EDUCATION

Penulis

  • Saiyidinal Firdaus Jakarta State University

DOI:

https://doi.org/10.30822/8nvk6676
Crossmark

Kata Kunci:

Authentic assessment; evidence-centered design; argument-based validity; program evaluation; higher education English learning

Abstrak

Current debates in higher education English language assessment center on the tension between pedagogically valued authentic assessment and the growing demand for transparency, defensibility, and program-level accountability. While authentic tasks are widely promoted for representing real-world language use, their capacity to support valid inferences and systematic evaluation remains contested. This study engages with theories of Evidence-Centered Design, argument-based validity, and learning-oriented assessment to address this tension through an integrative program evaluation lens. Drawing on an empirical examination of authentic assessment practices in a higher education English program, the study demonstrates how model-based assessment theory can inform and strengthen current practice without undermining pedagogical authenticity. The proposed framework articulates key theoretical features, including explicit construct and evidence modeling, validity argumentation, and alignment across task, course, and program levels. Findings show that authentic tasks generate rich language performance evidence, but only become evaluatively actionable when embedded within explicit inferential and evaluative structures. The study concludes by arguing that integrating authenticity with model-based assessment and program evaluation enables assessment systems that are simultaneously learning-oriented, inferentially defensible, and institutionally usable, offering a principled pathway for best practice in higher education English language assessment.

Unduhan

Data unduhan tidak tersedia.

Referensi

Ajjawi, R., Tai, J., Dollinger, M., Dawson, P., Boud, D., & Bearman, M. (2024). From authentic assessment to authenticity in assessment: broadening perspectives. Assessment & Evaluation in Higher Education, 49(4), 499–510. https://doi.org/10.1080/02602938.2023.2271193

Barno, E., & Phelps, G. (2025). Using a Multi-Agent System and Evidence-Centered Design to Integrate Educator Expertise Within Generated Feedback. Education Sciences, 15(10), 1273. https://doi.org/10.3390/educsci15101273

Bearman, M., Nieminen, J. H., & Ajjawi, R. (2023). Designing assessment in a digital world: an organizing framework. Assessment & Evaluation in Higher Education, 48(3), 291–304. https://doi.org/10.1080/02602938.2022.2069674

Boud, D., & Bearman, M. (2024). The assessment challenge of social and collaborative learning in higher education. Educational Philosophy and Theory, 56(5), 459–468. https://doi.org/10.1080/00131857.2022.2114346

Buchan, M.C., Bhawra, J. & Katapally, T.R. (2024). Navigating the digital world: development of an evidence-based digital literacy program and assessment tool for youth. Smart Learn. Environ, 11(8). https://doi.org/10.1186/s40561-024-00293-x

Charlton, N., & Newsham-West, R. (2024). A conceptual model for program-level assessment. Higher Education Research & Development, 43(8), 1721–1736. https://doi.org/10.1080/07294360.2024.2364094

Dai, D. W., Vu, T., Knoch, U., Lim, A. S., Malone, D. T., & Mak, V. (2024). Expanding Kane's argument-based validity framework: What can validation practices in language assessment offer health professions education?. Medical education, 58(12), 1462–1468. https://doi.org/10.1111/medu.15452

Fawns, T., Bearman, M., Dawson, P., Nieminen, J. H., Ashford-Rowe, K., Willey, K., Press, N. (2025). Authentic assessment: from panacea to criticality. Assessment & Evaluation in Higher Education, 50(3), 396–408. https://doi.org/10.1080/02602938.2024.2404634

Hadzhikoleva, S., Hadzhikolev, E., Gaftandzhıeva, S., & Pashev, G. (2025). A conceptual framework for multi-component summative assessment in an e-learning management system. Frontiers in Education, 10. https://doi.org/10.3389/feduc.2025.1656092

Hu, A., Liu, Q., & Daniel, B. (2025). Digital Technologies in Authentic Assessment in Higher Education: A Systematic Literature Review and Narrative Synthesis. Sage Open, 15(3). https://doi.org/10.1177/21582440251357198

Huang, S.-H. (2025). Evaluating the English for General Purposes (EGP) program at a Taiwanese university: A CIPP (context, input, process, and product) model study. Evaluation and Program Planning, 112, 102662. https://doi.org/10.1016/j.evalprogplan.2025.102662

Imsa-ard, P. (2025). Learning-oriented assessment and L2 argumentative writing: effects on writing ability and academic resilience. Discov Educ 4, 396. https://doi.org/10.1007/s44217-025-00779-x

Jones, Daniel & Cheng, Liying & Tweedie, Gregory. (2023). Automated Scoring of Speaking and Writing: Starting to Hit Its Stride. Canadian Journal of Learning and Technology. 48. 1-22. 10.21432/cjlt28241.

Krooi, M., Whittingham, J., & Beausaert, S. (2024). Introducing the 3P conceptual model of internal quality assurance in higher education: A systematic literature review. Studies in Educational Evaluation, 82, 101360. https://doi.org/10.1016/j.stueduc.2024.101360

Kubsch, Marcus., Grimm, Adrian., Neumann, Knut., & Drachsler, Hendrik. (2024). Using Evidence-Centered Design to Develop an Automated System for Tracking Students’ Physics Learning in a Digital Learning Environment. Uses of Artificial Intelligence in STEM Education, 11, 230—249. https://doi.org/10.1093/oso/9780198882077.003.0011

Lam, D. M. K., & Gayton, A. M. (2025). "6.5 sounds a bit better than 6.0?": a case for embedding language assessment literacy in university teaching staff professional development. Higher Education Research & Development, 44(4), 976–991. https://doi.org/10.1080/07294360.2024.2439854

Li, J., Huang, J., & Sheeran, T. (2025). ChatGPT4o as an AI Peer Assessor in EFL Speaking Classrooms: Examining Scoring Reliability and Feedback Effectiveness. Sage Open, 15(3). https://doi.org/10.1177/21582440251369938

Liu, C., Hwang, G., Yu, P., Tu, Y., & Wang, Y. (2025). Effects of an automated corrective feedback-based peer assessment approach on students’ learning achievement, motivation, and self-regulated learning conceptions in foreign language pronunciation. Educational Technology Research and Development, 73(4), 2403-2424. https://doi.org/10.1007/s11423-025-10484-z

Liu, Xu Jared; Wang, Jingwen; and Zou, Bin. (2025). Evaluating an AI speaking assessment tool: score accuracy, perceived validity, and oral peer feedback as feedback enhancement. Journal of English for Academic Purposes, 75, 101505. doi:10.1016/j.jeap.2025.101505.

Newton, S., Edwards, D., Usselman, M., Hernandez, D., Helms, M., Alemdar, M., & Rutstein, D. (2021). Utilizing Evidence-Centered Design to Develop Assessments: A High School Introductory Computer Science Course. Frontiers in Education, 6. https://doi.org/10.3389/feduc.2021.695376

Nguyen, T. M. H., Gu, P., & Coxhead, A. (2023). Argument-based validation of academic collocation tests. Language Testing, 41(3), 459-505. https://doi.org/10.1177/02655322231198499

Nieminen, Juuso & Yan, Zi & Boud, David. (2025). Self-assessment design in a digital world: centring student agency. Assessment & Evaluation In Higher Education. 50. 10.1080/02602938.2025.2467647.

Tubis, A. A. (2023). Digital Maturity Assessment Model for the Organizational and Process Dimensions. Sustainability, 15(20), 15122. https://doi.org/10.3390/su152015122

Viswanathan, N., Meacham, S., & Adedoyin, F. F. (2022). Enhancement of the online education system by using a multi-agent approach. Computers and Education: Artificial Intelligence, 3, 100057. https://doi.org/10.1016/j.caeai.2022.100057

Vlachopoulos, Dimitrios & Makri, Agoritsa. (2024). A systematic literature review on authentic assessment in higher education: Best practices for the development of 21st-century skills and policy considerations. Studies in Educational Evaluation, 83. https://doi.org/10.1016/j.stueduc.2024.101425

Vogt, K., Bøhn, H., & Tsagari, D. (2024). Language assessment literacy. Language Teaching, 57(3), 325–340. doi:10.1017/S0261444824000090

Wakid, M., Sofyan, H., Widowati, A., & Zaida Ilma, A. (2024). Learning-oriented assessment: a systematic literature network analysis. Cogent Education, 11(1). https://doi.org/10.1080/2331186X.2024.2366075

Zapata-Rivera, D, Forsyth, C, Graf, E, & Jiang, Y. (2024). Designing and Evaluating Evidence-Centered Design-Based Conversations for Assessment with LLMs. Proceedings of EDM 2024 Workshop: Leveraging Large Language Models for Next Generation Educational Technologies, 1—9. https://par.nsf.gov/biblio/10591735.

Zhan, Y., Boud, D. & Du, Z. (2025). Designing for authentic assessment: a scoping review. High Educ. https://doi.org/10.1007/s10734-025-01588-9

Zhang, J., Yu, G., & Browne, W. (2025). Teachers’ language assessment literacy: Exploring its construct and contextual factors. Studies in Educational Evaluation, 87, 101525. https://doi.org/10.1016/j.stueduc.2025.101525

Diterbitkan

2026-09-12

Cara Mengutip

INTEGRATING AUTHENTIC AND MODEL-BASED ASSESSMENT IN ENGLISH LANGUAGE LEARNING: A COMPREHENSIVE EVALUATION FRAMEWORK FOR HIGHER EDUCATION. (2026). Lectio: Journal of Language and Language Teaching, 6(2), 1-16. https://doi.org/10.30822/8nvk6676