Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

On improving Kaiwa (会話) assessment

Incorporating JF Standard descriptors and MFRM

Bibliographic Data

ID21240110
AuthorsRina Supriatnaningsih (0000-0002-0329-0563, State University of Semarang), Ahmad Yulianto (State University of Semarang), Lispridona Diner (0000-0001-6576-4819, State University of Semarang)
Year2023
Volume13
Issue2
Pages407-417
Publication date2023-09-30
Peer ReviewedYes
Open AccessYes
TypeARTICLE
VenueIndonesian Journal of Applied Linguistics (JOURNAL)
Journal identifiersISSN: 2301-9468 • E-ISSN: 2502-6747
PublisherUniversitas Pendidikan Indonesia (UPI) (PUBLISHER)
DOI10.17509/ijal.v13i2.63075
OpenAlexW4391367815
LanguageEN
Citations received1
References cited2

Despite the available rubrics, assessing speaking objectively has been a debatable issue to language assessment experts mostly due to the dependence on the raters’ authority. Scoring speaking performance often results in unfairness since subjectivity may come into play. Kaiwa (Speaking) is one of the four competencies examined in Japanese language assessment. On one hand, objective and accurate speaking assessment is badly needed. On the other hand, raters tend to overrate or underrate at times. Using Many-Facets Rasch Measurement (MFRM) and JF Standard descriptors, this study aimed to evaluate the Japanese speaking (Kaiwa) assessment. To this end, a cohort of 75 freshmen, consisting of 28 males (37%) and 47 females (63%), were assessed on the five-rubric scale (comprehension, vocabulary, structure, fluency, and pronunciation). These students’ age ranged from 18 to 20 years of age and their Japanese proficiency level was equal to N5. Two raters were involved in the assessment. The result revealed that: (1) 29 biases were found in rater-student interaction and rater-component interaction; (2) different patterns of rating behaviour were discovered. Rater 1 was more lenient than rater 2 but rater 2 was more consistent; (3) pronunciation and fluency are components that contributed the most to bias while structure was the most objective component being scored. For examiners, this result implies that scoring moderation should be held before grading students. For policymakers, the implication of the study suggests that modifications in the assessment rubric and statistical control be made so that fairer ratings could be achieved

Geography · Educational Technology and Assessment · EFL/ESL Teaching and Learning · Mathematics · Student Assessment and Feedback

  • Breakthrough, Overload, or Stability

    Open Access•Ai Sumirah Setiawati, Fathur Rokhman et al.•Languages•2026

  • Applying the Rasch Model

    Open Access•Trevor G Bond, Christine Fox et al.•Journal of Educational Measurement•2003

  • A comprehensive review of Rasch measurement in language assessment

    Open Access•Vahid Aryadoust, Li Ying Ng et al.•Language Testing•2021

Unique citing works1
Citations per year1
Citation span2026 - 2026 (1)
Citation velocitycurrent
Highly citedNo

Tools

Open DOIOpen Access
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae