A Machine Learning Model Helps Process Interviewer Comments in Computer-assisted Personal Interview Instruments: A Case Study

Catherine Billington,Jiating (Kristin) Chen,Andrew Jannett,Gonzalo Rivero

doi:10.1177/1525822x221107053

Catherine Billington, Jiating (Kristin) Chen + Show 2 more

https://doi.org/10.1177/1525822x221107053

Copy DOI

Journal: Field Methods	Publication Date: Jun 21, 2022
Citations: 1

Affiliation: Westat (United States)

Abstract

During data collection, field interviewers often append notes or comments to a case in open text fields to request updates to case-level data. Processing these comments can improve data quality, but many are non-actionable, and processing remains a costly manual task. This article presents a case study using a novel application of machine learning tools to assist in the evaluation of these comments. Using over 5,000 comments from the Medical Expenditure Panel Survey, we built features that were fed to a machine learning model to predict a grouping category for each comment as previously assigned by data technicians to expedite processing. The model achieved high top-3 accuracy and was incorporated into a production tool for editing. A qualitative evaluation of the tool also provided encouraging results. This application of machine learning tools allowed a small but worthwhile increase in processing efficiency, while maintaining exacting standards for data quality.

Full Text