The Second Conversational Intelligence Challenge (ConvAI2)
Emily Dinan
V. Logacheva
Valentin Malykh
Alexander H. Miller
Kurt Shuster
Jack Urbanek
Douwe Kiela
Arthur Szlam
Iulian Serban
Ryan J. Lowe
Shrimai Prabhumoye
A. Black
Alexander I. Rudnicky
Jason Williams
Joelle Pineau
Mikhail Burtsev
Jason Weston

Abstract
We describe the setting and results of the ConvAI2 NeurIPS competition that aims to further the state-of-the-art in open-domain chatbots. Some key takeaways from the competition are: (i) pretrained Transformer variants are currently the best performing models on this task, (ii) but to improve performance on multi-turn conversations with humans, future systems must go beyond single word metrics like perplexity to measure the performance across sequences of utterances (conversations) -- in terms of repetition, consistency and balance of dialogue acts (e.g. how many questions asked vs. answered).
View on arXivComments on this paper