Measuring Skin Color: Consistency, Comparability, and Meaningfulness of Rating Scale Scores and Handheld Device Readings

Measuring Skin Color: Consistency, Comparability, and Meaningfulness of Rating Scale Scores and Handheld Device Readings
复制标题

测量肤色:评定量表分数和手持设备读数的一致性、可比性和意义

DOI:
10.1093/jssam/smab046
复制
发表时间:
2022
影响因子:
2.1
通讯作者:
Gordon, Rachel A.
Gordon, Rachel A.
中科院分区:
数学3区
文献类型:
--
作者:
Gordon, Rachel A.

文献摘要

相似文献

随着美国社会继续多样化,对种族主义外表的更好测量的呼声越来越高,调查研究人员需要在实地研究中指导评估肤色的有效策略。这项研究考察了两种最广泛使用的肤色分级标准(Massey-Martin和Perla)和两种便携且廉价的手持肤色测量设备(Nix色度计和Labby分光光度计)的一致性、可比性和意义。我们亲自使用这四种仪器收集了46名大学生的数据,这些大学生被选为反映四个种族群体(亚洲人、黑人、拉丁裔和白人)广泛肤色的人。这些大学生、五名研究人员和来自在线样本的459名成年人也对40张库存照片进行了评级,再次选择了肤色多样性。我们的结果-基于在受控条件下收集的数据-显示出不同评分者和读数的高度一致性。梅西-马丁和佩拉的评分彼此高度线性相关,尽管佩拉在肤色最浅的人中更好地区分开来。除了显示种族内部和种族之间的预期差异外,NIX和Labby从暗到明(L*)的读数也同样与彼此以及与Massey-Martin和Perla的分数线性相关。此外,较暗的Massey-Martin和Perla评分与在线评分者的预期相关,即被拍摄者经历了更大的歧视。相比之下,红色(a*)和黄色(b*)的暗音在评级量表得分的中值范围内最高,并且在种族之间表现出更大的重叠。总体而言,当实施得当(例如,不需要记忆)时,每种工具都表现出足够的一致性、可比性和在实地调查中使用的意义。然而,在代表肤色最浅的个人的研究中,Perla可能比Massey-Martin更受欢迎,而当研究只能收集单一评级时,手持设备可能更受青睐,以减少测量误差。
As US society continues to diversify and calls for better measurements of racialized appearance increase, survey researchers need guidance about effective strategies for assessing skin color in field research. This study examined the consistency, comparability, and meaningfulness of the two most widely used skin tone rating scales (Massey–Martin and PERLA) and two portable and inexpensive handheld devices for skin color measurement (Nix colorimeter and Labby spectrophotometer). We collected data in person using these four instruments from forty-six college students selected to reflect a wide range of skin tones across four racial-ethnic groups (Asian, Black, Latinx, White). These college students, five study staff, and 459 adults from an online sample also rated forty stock photos, again selected for skin tone diversity. Our results—based on data collected under controlled conditions—demonstrate high consistency across raters and readings. The Massey–Martin and PERLA scale scores were highly linearly related to each other, although PERLA better differentiated among people with the lightest skin tones. The Nix and Labby darkness-to-lightness (L*) readings were likewise linearly related to each other and to the Massey–Martin and PERLA scores, in addition to showing expected variation within and between race ethnicities. In addition, darker Massey–Martin and PERLA ratings correlated with online raters’ expectations that a photographed person experienced greater discrimination. In contrast, the redness (a*) and yellowness (b*) undertones were highest in the mid-range of the rating scale scores and demonstrated greater overlap across race-ethnicities. Overall, each instrument showed sufficient consistency, comparability, and meaningfulness for use in field surveys when implemented soundly (e.g., not requiring memorization). However, PERLA might be preferred to Massey–Martin in studies representing individuals with the lightest skin tones, and handheld devices may be preferred to rating scales to reduce measurement error when studies could gather only a single rating.