Cognitive Processes and Foreign Language Reading: Investigating Students' Test Taking Behaviour with a Reflective Questionnaire

Commonly, models of reading emphasize the relevance of cognitive processes and metacognitive strategies when it comes to reading comprehension. Although these models are derived from research into first language (L1) reading, they often serve as the basis for operationalizing foreign language (FL) reading comprehension tests. This is also true for the FL Reading Test (E8 Reading Test) administered in 2013 and 2019 in Austria. However, little is known about test takers’ behaviour regarding their cognitive processes and metacognitive strategies when taking the test. Undoubtedly, such information could be useful for the FL classroom and to inform foreign language reading instruction as well as test development. Therefore, this paper investigates which cognitive processes and metacognitive strategies students (n=106) apply when taking the E8 Reading Test. Data was collected using a reflective questionnaire. The results show that compared to weaker readers, strong readers apply the expected cognitive processes more frequently. There is no statistically significant difference, however, between stronger and weaker readers regarding metacognitive strategies applied. Furthermore, the data revealed that some good readers arrive at the correct answer by making use of different/other cognitive processes and metacognitive strategies than expected. These findings emphasize the importance of explicit, guided foreign language reading instruction focussing not only on the product of reading (comprehension) but also on the processes involved. However, more research is needed to better understand what the absence or presence of skills and strategies mean regarding individual learner abilities.

2025-10-10

The Design and Validation of an Online Speaking Test for Young Learners in Uruguay: Challenges and Innovations

This research presents the development of an online speaking test of English for students at the end of primary and beginning of secondary school education in state schools in Uruguay. Following the success of the Plan Ceibal one computer-tablet per child initiative, there was a drive to further utilize technology to improve the language ability of students, particularly in speaking, where the majority of students are at CEFR levels pre-A1 and A1. The national concern over a lack of spoken communicative skills amongst students led to a decision to develop a new speaking test, specifically tailored to local needs. This paper provides an overview of the speaking test development and validation project designed with the following objectives in mind: to establish, track, and report annually learners’ achievements against the Common European Framework of Reference for Languages (CEFR) targeting CEFR levels pre-A1 to A2, to inform teaching and learning, and to promote speaking practice in classrooms. Results of a three-phase mixed-methods study involving small-scale and large-scale trials with learners and examiners as well as a CEFR-linking exercise with expert panelists will be reported. Different sources of evidence will be brought together to build a validity argument for the test. The paper will also focus on some of the challenges involved in assessing young learners and discuss how design decisions, local knowledge and expertise, and technological innovations can be used to address such challenges with implications for other similar test development projects.

2025-10-10

Equating Rasch Values and Expert Judgement Through Externally-Referenced Anchoring

This paper reports on the use of externally-referenced anchoring by LanguageCert as a methodology for calibrating language test materials and aligning test forms. The datasets used are taken from tests at each of the six levels of LanguageCert IESOL suite, all of which have been aligned to the CEFR through expert judgement. We illustrate in this paper the extent to which externally-referenced anchoring, using Item Response Theory (IRT) but based on expert judgement, can be used as an effective, reliable and valid methodology. The approach is based on the premise that successful anchoring may be achieved by reference to well-targeted, expertly-written test forms aligned to the underlying traits of a particular CEFR level by expert judgement and verified through the use of IRT. This study focuses on the analysis of 18 LanguageCert test forms, three at each CEFR level. The LanguageCert Item Difficulty (LID) scale, which underlies all LanguageCert test materials, is linked empirically to the CEFR, and each test was placed on the LID scale based at the midpoint of its distribution. This midpoint setting was then set as the externally-referenced anchor for a given CEFR level. The findings of this study indicate that, while the match between the distribution of items in the selected LanguageCert IESOL tests and the LID scale was not perfect, in general, a relatively close match between the items in the tests and the LID scale was found and, as a consequence, the corresponding CEFR level. For each test, most of the items fell between the 25th and 75th percentile of any given level: this range representing the lower and upper bounds of LID scale values for each CEFR level. These results demonstrate that LanguageCert IESOL test items are well set and appropriately positioned at respective CEFR levels on the basis of expert judgement. The study illustrates that externally-referenced anchoring based on expert judgement may be used as a methodology for aligning test forms to an external frame of reference, in this case the CEFR.

2025-10-10