| 英文摘要 |
Task-oriented dialog systems require labeled corpus for model training. However, in the face of new services, how to effectively collect dialogue corpus is a problem that must be faced in the construction of dialogue systems. Existing task-oriented systems mainly focus on reservations for restaurants, hotels, and airline tickets. There is no dialogue corpus for virtual assistants that could provide transactional services such as sending messages and creating events. This paper imitates the method of collecting dialogue datasets from CrossWOZ to allow annotators to simulate user and virtual assistant dialogue scenarios through a dialogue website interface, creating a dialogue dataset that can handle three services: email management, calendar management, and message delivery. It is expected that this corpus will lay the foundation for the development of Chinese virtual assistant dialogue system. The annotation system and dataset have been open-sourced at https://github.com/TedYeh/messageWOZ. |