ECAsT: a large dataset for conversational search and an evaluation of metric robustness

Al-Thani, Haya; Jansen, Bernard J.; Elsayed, Tamer

المؤلف	Al-Thani, Haya
المؤلف	Jansen, Bernard J.
المؤلف	Elsayed, Tamer
تاريخ الإتاحة	2024-11-05T06:05:19Z
تاريخ النشر	2023
اسم المنشور	PeerJ Computer Science
المصدر	Scopus
المعرّف	http://dx.doi.org/10.7717/PEERJ-CS.1328
الرقم المعياري الدولي للكتاب	23765992
معرّف المصادر الموحد	http://hdl.handle.net/10576/60872
الملخص	The Text REtrieval Conference Conversational assistance track (CAsT) is an annual conversational passage retrieval challenge to create a large-scale open-domain conversational search benchmarking. However, as of yet, the datasets used are small, with just more than 1,000 turns and 100 conversation topics. In the first part of this research, we address the dataset limitation by building a much larger novel multi-turn conversation dataset for conversation search benchmarking called Expanded-CAsT (ECAsT). ECAsT is built using a multi-stage solution that uses a combination of conversational query reformulation and neural paraphrasing and also includes a new model to create multiturn paraphrases. The meaning and diversity of paraphrases are evaluated with human and automatic evaluation. Using this methodology, we produce and release to the research community a conversational search dataset that is 665% more extensive in terms of size and language diversity than is available at the time of this study, with more than 9,200 turns. The augmented dataset not only provides more data but also more language diversity to improve conversational search neural model training and testing. In the second part of the research, we use ECAsT to assess the robustness of traditional metrics for conversational evaluation used in CAsT and identify its bias toward language diversity. Results show the benefits of adding language diversity for improving the collection of pooled passages and reducing evaluation bias. We found that introducing language diversity via paraphrases returned up to 24% new passages compared to only 2% using CAsT baseline.
اللغة	en
الناشر	PeerJ Inc.
الموضوع	Conversational evaluation Conversational search Open-domain Paraphrasing
العنوان	ECAsT: a large dataset for conversational search and an evaluation of metric robustness
النوع	Article
رقم المجلد	9
dc.accessType	Open Access

الملفات في هذه التسجيلة

الاسم:: peerj-cs-1328.pdf
الحجم:: 4.463Mb
الصيغة:: PDF

عرض / فتح

هذه التسجيلة تظهر في المجموعات التالية

علوم وهندسة الحاسب [‎2484‎ items ]

عرض بسيط للتسجيلة

ECAsT: a large dataset for conversational search and an evaluation of metric robustness

الملفات في هذه التسجيلة

هذه التسجيلة تظهر في المجموعات التالية

Video