A simulated dataset for proactive robot task inference from streaming natural language dialogues

1Citations
Citations of this article
9Readers
Mendeley users who have this article in their library.

Abstract

This paper introduces a dataset designed to support research on proactive robots that infer human needs from natural language conversations. Unlike traditional human-robot interaction datasets focused on explicit commands, this dataset captures implicit task requests within multi-party dialogues. It simulates realistic workplace environments, spanning 10 diverse scenarios, such as biotechnology research centers, legal consulting firms, and game development studios, among others. The dataset includes 10,000 synthetic dialogues generated using a large language model-based pipeline, covering a wide range of topics, including task-related discussions and casual conversations. The dataset focuses on common workplace tasks, such as borrowing, distributing, and processing items. It provides a resource for advancing proactive robotic systems, enabling research in natural language understanding, intent recognition, and autonomous task inference.

Cite

CITATION STYLE

APA

Xu, H., Li, C., Yuan, X., Zhi, T., & Liu, H. (2025). A simulated dataset for proactive robot task inference from streaming natural language dialogues. Scientific Data , 12(1). https://doi.org/10.1038/s41597-025-05727-w

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free