How to construct and implement script concordance tests: insights from a systematic review | Zendy

Dory Valérie | Zendy; Gag Robert | Zendy; Vanpee Dominique | Zendy; Charlin Bernard | Zendy

AI Assistant Blog Pricing

Home ZAIA Blog

Premium

How to construct and implement script concordance tests: insights from a systematic review

Author(s) -

Dory Valérie,

Gag Robert,

Vanpee Dominique,

Charlin Bernard

Publication year - 2012

Publication title -

medical education

Language(s) - English

Resource type - Journals

SCImago Journal Rank - 1.776

H-Index - 138

eISSN - 1365-2923

pISSN - 0308-0110

DOI - 10.1111/j.1365-2923.2011.04211.x

Subject(s) - concordance , medline , construct (python library) , psycinfo , construct validity , competence (human resources) , medical education , formative assessment , psychology , computer science , medicine , psychometrics , social psychology , mathematics education , clinical psychology , political science , law , programming language

Medical Education 2012: 46 : 552–563 Context Programmes of assessment should measure the various components of clinical competence. Clinical reasoning has been traditionally assessed using written tests and performance‐based tests. The script concordance test (SCT) was developed to assess clinical data interpretation skills. A recent review of the literature examined the validity argument concerning the SCT. Our aim was to provide potential users with evidence‐based recommendations on how to construct and implement an SCT. Methods A systematic review of relevant databases (MEDLINE, ERIC [Education Resources Information Centre], PsycINFO, the Research and Development Resource Base [RDRB, University of Toronto]) and Google Scholar, medical education journals and conference proceedings was conducted for references in English or French. It was supplemented by ancestry searching and by additional references provided by experts. Results The search yielded 848 references, of which 80 were analysed. Studies suggest that tests with around 100 items (25–30 cases), of which 25% are discarded after item analysis, should provide reliable scores. Panels with 10–20 members are needed to reach adequate precision in terms of estimated reliability. Panellists’ responses can be analysed by checking for moderate variability among responses. Studies of alternative scoring methods are inconclusive, but the traditional scoring method is satisfactory. There is little evidence on how best to determine a pass/fail threshold for high‐stakes examinations. Conclusions Our literature search was broad and included references from medical education journals not indexed in the usual databases, conference abstracts and dissertations. There is good evidence on how to construct and implement an SCT for formative purposes or medium‐stakes course evaluations. Further avenues for research include examining the impact of various aspects of SCT construction and implementation on issues such as educational impact, correlations with other assessments, and validity of pass/fail decisions, particularly for high‐stakes examinations.

This content is not available in your region!

Continue researching here.

Having issues? You can contact us here

Empowering knowledge with every search

About

About Careers Publisher Partners Contact Us

Learn

FAQs Blog Terms of Use Privacy Policy

About

Learn

Discover

Explore