Hi, I'm the NTCIR Dialogue Evaluation Task. I have been accepted for #ntcir15 .
My mother is @ntcirstc . I have two subtasks: Dialogue Quality (DQ) and ND (Nugget Detection). These subtasks are the same as the #ntcir14 STC-3 DQ and ND subtasks.
Draft overview + official results for the Dialogue Quality (DQ) and Nugget Detection (ND) subtasks have been sent to the participants!
https://t.co/qUIJ2OYD9Z
The training and test data (in Chinese and English) for
DQ (Dialogue Quality) and ND (Nugget Detection) have just been released!
https://t.co/t2vcw8BXpH
The English data are manual translations of the original Chinese Helpdesk-Customer dialogues on Weibo.
We already have 14 registered teams but we have decided to extend the task registration deadline to Sunday 9th Steptember!
Please register! Have fun with our new tasks!
https://t.co/OFfXzBJaWc
Please participate in the #ntcir14 Short Text Conversation task (STC-3) if you are interested in dialogue systems and/or dialogue evaluation!
Slides (updated June 6):
https://t.co/gYdsEcUip3
DQ: given a customer-helpdesk dialogue, return an estimated distribution over user ratings e.g. customer satisfaction.
ND: given a dialogue, for each customer or helpdesk utterance block, return an estimated distribution over utterance types, e.g. trigger, goal, not-a-nugget.
We have updated the slides for the Dialogue Quality subtask and the Nugget Detection subtask (May 2). Please visit our website!
https://t.co/qUIJ2OYD9Z
#ntcir14
#sigir2018 short paper accepted: Comparing Two Binned Probability Distributions for Information Access Evaluation
The measures will be used in the #ntcir14@ntcirstc task.
Short Text Conversation, THE largest task of NTCIR, is back!
Check out the tentative info of the Chinese Emotional Conversation Generation, Dialogue Quality, and Nugget Detection subtasks at https://t.co/qUIJ2OYD9Z
#ntcir14
64 runs are retrieval-based just like STC1, while 56 runs are generation-based.
As many as 15 teams tried the generation-based approach! https://t.co/v5emh8Ubkn