<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ArticleSet PUBLIC "-//NLM//DTD PubMed 2.7//EN" "https://dtd.nlm.nih.gov/ncbi/pubmed/in/PubMed.dtd">
<ArticleSet>
<Article>
<Journal>
<PublisherName>OICC Press</PublisherName>
<JournalTitle>Journal of New Trends in English Language Learning (JNTELL)</JournalTitle>
<Issn>2821-2290</Issn>
<Volume>5</Volume>
<Issue>1</Issue>
<PubDate PubStatus="epublish">
<Year>2026</Year>
<Month>03</Month>
<Day>31</Day>
</PubDate>
</Journal>
<ArticleTitle>Implementing the Rasch Rating Scale Model for Assessing Writing Performance</ArticleTitle>
<VernacularTitle></VernacularTitle>
<FirstPage></FirstPage>
<LastPage></LastPage>
<ELocationID EIdType="doi">10.57647/jntell.2026.0501.03</ELocationID>
<Language>EN</Language>
<AuthorList>
<Author>
<FirstName>Sareh</FirstName>
<LastName>Sahebalam</LastName>
<Affiliation>Department of Teaching English as a Foreign Language, NT.C., Islamic Azad University, Tehran, Iran</Affiliation>
<Identifier Source="ORCID"></Identifier>
</Author>
<Author>
<FirstName>Purya</FirstName>
<LastName>Baghaei</LastName>
<Affiliation>Department of Teaching English as a Foreign Language, Ma.C., Islamic Azad University, Mashhad, Iran</Affiliation>
<Identifier Source="ORCID"></Identifier>
</Author>
<Author>
<FirstName>Mojgan</FirstName>
<LastName>Rashtchi</LastName>
<Affiliation>Department of Teaching English as a Foreign Language, NT.C., Islamic Azad University, Tehran, Iran</Affiliation>
<Identifier Source="ORCID">https://orcid.org/0000-0001-7713-9316</Identifier>
</Author>
</AuthorList>
<PublicationType>Journal Article</PublicationType>
<History>
<PubDate PubStatus="received">
<Year>2026</Year>
<Month>03</Month>
<Day>31</Day>
</PubDate>
</History>
<Abstract>Rater-mediated assessments require raters to make complex judgments, often expressed through ordinal rating scale categories, about test-takers’ performances. When the focus is on the quality of these judgments, it becomes essential to evaluate the ratings for their psychometric soundness, including validity, reliability, and fairness. In applied linguistics and second-language assessment research, the Many-Facet Rasch Model (MFRM; Linacre, 1989) has been widely used to investigate rater performance. However, the MFRM’s complexity, combined with the limited accessibility of user-friendly software, presents challenges for researchers and practitioners. In this study, the researchers propose using the Rating Scale Model (RSM), a more straightforward and widely recognized framework, as an alternative for analyzing judged performance. To explore its utility, the RSM was applied to scores assigned by five raters to 156 compositions written by English as a Foreign Language Learners (EFL). The findings indicate that the RSM can successfully convert ordinal ratings into interval-level measures for both test-takers and tasks, while also allowing examination of item fit, person fit, rating scale functioning, and dimensionality. Moreover, the model proved capable of investigating interaction effects, such as the influence of rater and examinee gender, as well as the interplay between raters and specific items. Overall, this study demonstrates that the RSM, while more limited than the MFRM, offers a practical and accessible tool for evaluating the quality of assessor judgments in performance assessment. We also highlight both the promise and the constraints of relying on RSM in this context, with implications for future inquiry and implementation.</Abstract>
<ObjectList>
<Object Type="keyword">
<Param Name="value">Rasch Rating Scale Model</Param>
</Object>
<Object Type="keyword">
<Param Name="value">Writing Performance</Param>
</Object>
<Object Type="keyword">
<Param Name="value">Assessment</Param>
</Object>
</ObjectList>
</Article>
</ArticleSet>