Examining the relationship between differential item functioning and differential test functioning

2006 ◽  
Vol 23 (4) ◽  
pp. 475-496 ◽  
Author(s):  
T.-I. Pae ◽  
G.-P. Park
2021 ◽  
pp. 001316442110015
Author(s):  
Dimiter M. Dimitrov ◽  
Dimitar V. Atanasov

This study offers an approach to testing for differential item functioning (DIF) in a recently developed measurement framework, referred to as D-scoring method (DSM). Under the proposed approach, called P–Z method of testing for DIF, the item response functions of two groups (reference and focal) are compared by transforming their probabilities of correct item response, estimated under the DSM, into Z-scale normal deviates. Using the liner relationship between such Z-deviates, the testing for DIF is reduced to testing two basic statistical hypotheses about equal variances and equal means of the Z-deviates for the reference and focal groups. The results from a simulation study support the efficiency (low Type error and high power) of the proposed P–Z method. Furthermore, it is shown that the P–Z method is directly applicable in testing for differential test functioning. Recommendations for practical use and future research, including possible applications of the P–Z method in IRT context, are also provided.


2013 ◽  
Vol 34 (3) ◽  
pp. 170-183 ◽  
Author(s):  
Eunike Wetzel ◽  
Benedikt Hell

Large mean differences are consistently found in the vocational interests of men and women. These differences may be attributable to real differences in the underlying traits. However, they may also depend on the properties of the instrument being used. It is conceivable that, in addition to the intended dimension, items assess a second dimension that differentially influences responses by men and women. This question is addressed in the present study by analyzing a widely used German interest inventory (Allgemeiner Interessen-Struktur-Test, AIST-R) regarding differential item functioning (DIF) using a DIF estimate in the framework of item response theory. Furthermore, the impact of DIF at the scale level is investigated using differential test functioning (DTF) analyses. Several items on the AIST-R’s scales showed significant DIF, especially on the Realistic, Social, and Enterprising scales. Removal of DIF items reduced gender differences on the Realistic scale, though gender differences on the Investigative, Artistic, and Social scales remained practically unchanged. Thus, responses to some AIST-R items appear to be influenced by a secondary dimension apart from the interest domain the items were intended to measure.


2019 ◽  
Vol 38 (5) ◽  
pp. 627-641
Author(s):  
Beyza Aksu Dunya ◽  
Clark McKown ◽  
Everett Smith

Emotion recognition (ER) involves understanding what others are feeling by interpreting nonverbal behavior, including facial expressions. The purpose of this study is to evaluate the psychometric properties of a web-based social ER assessment designed for children in kindergarten through third grade. Data were collected from two separate samples of children. The first sample included 3,224 children and the second sample included 4,419 children. Data were calibrated using Rasch dichotomous model. Differential item and test functioning were also evaluated across gender and ethnicity. Across both samples, we found consistent item fit, unidimensional item structure, and adequate item targeting. Analyses of differential item functioning (DIF) found six out of 111 items displaying DIF across gender and no items demonstrating DIF across ethnicity. The analyses of person measure calibrations with and without DIF items yielded no evidence of differential test functioning (DTF) across gender and ethnicity groups in both samples.


2021 ◽  
Vol ahead-of-print (ahead-of-print) ◽  
Author(s):  
Julio César Acosta-Prado ◽  
Arnold Alejandro Tafur-Mendoza ◽  
Rodrigo Arturo Zárate-Torres ◽  
Geli Mercedes Pautt-Torres

Purpose Job satisfaction and leadership behavior are recognized by the organizational world as fundamental elements that influence the overall effectiveness of a company. However, as the first step for an adequate intervention on any of these variables, it is the evaluation. The purpose of this paper is to develop and validate two brief measures on job satisfaction and leadership behavior. Design/methodology/approach The sample was made up of 246 workers located in Bogota, Colombia. The study was an instrumental research. To collect validity evidence, the internal structure and the relationship with other variables were used. For the evaluation of equity, the differential item functioning was analyzed according to the sex of the participants. Reliability was estimated through the ordinal omega coefficient. Findings Both brief measures presented a unifactorial structure, where job satisfaction was measured by five items and leadership behavior by four items. On the other hand, only one item of leadership behavior showed differential item functioning; however, its magnitude was trivial. Also, convergent and discriminant evidence was provided for both measures, and the reliability levels were adequate. Originality/value The measures developed represents an effort to briefly measure job satisfaction and leadership behavior. Likewise, it constitutes two of the few instruments to measure job satisfaction and leadership behavior in Latin American, representing a good alternative for the measurement of the referred constructs in an organizational context.


2020 ◽  
pp. 073428292094552
Author(s):  
Maryellen Brunson McClain ◽  
Bryn Harris ◽  
Sarah E. Schwartz ◽  
Megan E. Golson

Although the racial/ethnic demographics in the United States are changing, few studies evaluate the cultural and linguistic responsiveness of commonly used autism spectrum disorder screening and diagnostic assessment measures. The purpose of this study is to evaluate item and test functioning of the Autism Spectrum Rating Scales (ASRS) in a sample of racially/ethnically diverse parents of children (nonclinical) between the ages of 6–18 ( N = 806). This study is a follow-up to a prior publication examining the factor structure of the ASRS among a similar sample. The present study furthers the examination of measurement invariance of the ASRS in racially/ethnically diverse populations by conducting differential item functioning and differential test functioning with a larger sample. Results indicate test-level invariance; however, five items are noninvariant across parent reporters from different racial/ethnic groups. Implications for practice and directions for future research are discussed.


Sign in / Sign up

Export Citation Format

Share Document