Collecting and Characterizing Natural Language Utterances for Specifying Data Visualizations
Collecting and Characterizing Natural Language Utterances for Specifying Data Visualizations
复制标题
DOI:
10.1145/3411764.3445400
复制
发表时间:
2021-05
期刊:
影响因子:
--
通讯作者:
Arjun Srinivasan;Nikhila Nyapathy;Bongshin Lee;S. Drucker;J. Stasko
中科院分区:
文献类型:
--
作者:
Arjun Srinivasan;Nikhila Nyapathy;Bongshin Lee;S. Drucker;J. Stasko
Natural language interfaces (NLIs) for data visualization are becoming increasingly popular both in academic research and in commercial software. Yet, there is a lack of empirical understanding of how people specify visualizations through natural language. We conducted an online study (N = 102), showing participants a series of visualizations and asking them to provide utterances they would pose to generate the displayed charts. From the responses, we curated a dataset of 893 utterances and characterized the utterances according to (1) their phrasing (e.g., commands, queries, questions) and (2) the information they contained (e.g., chart types, data aggregations). To help guide future research and development, we contribute this utterance dataset and discuss its applications toward the creation and benchmarking of NLIs for visualization.