Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model

Two-tier multiple-choice (TTMC) items are used to assess students’ knowledge of a scientific concept for tier 1 and their reasoning about this concept for tier 2. But are the knowledge and reasoning involved in these tiers really distinguishable? Are the tiers equally challenging for students? The a...

Full description

Bibliographic Details
Main Authors: Fulmer, G., Chu, H., Treagust, David, Neumann, K.
Format: Journal Article
Published: 2015
Online Access:http://hdl.handle.net/20.500.11937/43349
_version_ 1848756666815741952
author Fulmer, G.
Chu, H.
Treagust, David
Neumann, K.
author_facet Fulmer, G.
Chu, H.
Treagust, David
Neumann, K.
author_sort Fulmer, G.
building Curtin Institutional Repository
collection Online Access
description Two-tier multiple-choice (TTMC) items are used to assess students’ knowledge of a scientific concept for tier 1 and their reasoning about this concept for tier 2. But are the knowledge and reasoning involved in these tiers really distinguishable? Are the tiers equally challenging for students? The answers to these questions influence how we use and interpret TTMC instruments. We apply the Rasch measurement model on TTMC items to see if the items are distinguishable according to different traits (represented by the tier), or according to different content sub-topics within the instrument, or to both content and tier. Two TTMC data sets are analyzed: data from Singapore and Korea on the Light Propagation Diagnostic Instrument (LPDI), data from the United States on the Classroom Test of Scientific Reasoning (CTSR). Findings for LPDI show that tier-2 reasoning items are more difficult than tier-1 knowledge items, across content sub-topics. Findings for CTSR do not show a consistent pattern by tier or by content sub-topic. We conclude that TTMC items cannot be assumed to have a consistent pattern of difficulty by tier—and that assessment developers and users need to consider how the tiers operate when administering TTMC items and interpreting results. Researchers must check the tiers’ difficulties empirically during validation and use. Though findings from data in Asian contexts were more consistent, further study is needed to rule out differences between the LPDI and CTSR instruments.
first_indexed 2025-11-14T09:15:50Z
format Journal Article
id curtin-20.500.11937-43349
institution Curtin University Malaysia
institution_category Local University
last_indexed 2025-11-14T09:15:50Z
publishDate 2015
recordtype eprints
repository_type Digital Repository
spelling curtin-20.500.11937-433492017-09-13T14:00:16Z Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model Fulmer, G. Chu, H. Treagust, David Neumann, K. Two-tier multiple-choice (TTMC) items are used to assess students’ knowledge of a scientific concept for tier 1 and their reasoning about this concept for tier 2. But are the knowledge and reasoning involved in these tiers really distinguishable? Are the tiers equally challenging for students? The answers to these questions influence how we use and interpret TTMC instruments. We apply the Rasch measurement model on TTMC items to see if the items are distinguishable according to different traits (represented by the tier), or according to different content sub-topics within the instrument, or to both content and tier. Two TTMC data sets are analyzed: data from Singapore and Korea on the Light Propagation Diagnostic Instrument (LPDI), data from the United States on the Classroom Test of Scientific Reasoning (CTSR). Findings for LPDI show that tier-2 reasoning items are more difficult than tier-1 knowledge items, across content sub-topics. Findings for CTSR do not show a consistent pattern by tier or by content sub-topic. We conclude that TTMC items cannot be assumed to have a consistent pattern of difficulty by tier—and that assessment developers and users need to consider how the tiers operate when administering TTMC items and interpreting results. Researchers must check the tiers’ difficulties empirically during validation and use. Though findings from data in Asian contexts were more consistent, further study is needed to rule out differences between the LPDI and CTSR instruments. 2015 Journal Article http://hdl.handle.net/20.500.11937/43349 10.1186/s41029-015-0005-x fulltext
spellingShingle Fulmer, G.
Chu, H.
Treagust, David
Neumann, K.
Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model
title Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model
title_full Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model
title_fullStr Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model
title_full_unstemmed Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model
title_short Is it harder to know or to reason? Analyzing two-tier science assessment items using the Rasch measurement model
title_sort is it harder to know or to reason? analyzing two-tier science assessment items using the rasch measurement model
url http://hdl.handle.net/20.500.11937/43349