الملخص

Conversational Query Rewriting (CQR) turns a context dependent user turn into a standalone query for a retriever, and most methods do this in a single step from the dialogue history before retrieving once. The rewrite is therefore fixed before any corpus evidence is available to correct its reference resolution or its vocabulary. We recast CQR as a sequential retrieval problem: an agent rewrites the current turn, retrieves, and conditions its next rewrite on the returned passages. The agent acts in a typed space of three rewriting operations, resolving conversational intent into a standalone query, generating lexical reformulations, or synthesizing pseudo-documents for document-to-document matching, together with a stop action that ends the episode. We train the policy with supervised fine-tuning followed by reinforcement learning against a single retrieval-quality reward, using no human rewrite annotations. Across TopiOCQA and QReCC, the agent outperforms several retrieval-aligned baselines, while remaining effective across retrieval backends and generalizing to the CAsT benchmarks without additional training. Further analysis shows that, through retrieval-reward optimization alone, the learned policy develops a behavior of grounding pseudo-documents in passages retrieved by earlier steps, substantially improving retrieval.

الكلمات المفتاحية

الموضوع

بيانات النشر

المجلة
غير متاح
وصول مفتوح
وصول مفتوح أخضر

اقتبس هذه المقالة

APA 7

Coelho, J., Wang, H., Yuan, J., Wang, Z., Koelle, S., & Niu, W. (2026). Learning Multi-Step Query Rewriting via Corpus Feedback for Conversational Search. https://omanscience.com/ar/articles/learning-multi-step-query-rewriting-via-corpus-feedback-for-conversational-search

MLA 9

Coelho, João, et al. "Learning Multi-Step Query Rewriting via Corpus Feedback for Conversational Search." https://omanscience.com/ar/articles/learning-multi-step-query-rewriting-via-corpus-feedback-for-conversational-search.

شيكاغو (المؤلف–التاريخ)

Coelho, João, Hong Wang, Jie Yuan, Zhuoer Wang, Samson Koelle, and Wei Niu. 2026. "Learning Multi-Step Query Rewriting via Corpus Feedback for Conversational Search." https://omanscience.com/ar/articles/learning-multi-step-query-rewriting-via-corpus-feedback-for-conversational-search.

هارفارد

Coelho, J., Wang, H., Yuan, J., Wang, Z., Koelle, S. and Niu, W. (2026) 'Learning Multi-Step Query Rewriting via Corpus Feedback for Conversational Search', Available at: https://omanscience.com/ar/articles/learning-multi-step-query-rewriting-via-corpus-feedback-for-conversational-search.

فانكوفر

Coelho J, Wang H, Yuan J, Wang Z, Koelle S, Niu W. Learning Multi-Step Query Rewriting via Corpus Feedback for Conversational Search. https://omanscience.com/ar/articles/learning-multi-step-query-rewriting-via-corpus-feedback-for-conversational-search

IEEE

J. Coelho, H. Wang, J. Yuan, Z. Wang, S. Koelle, and W. Niu, "Learning Multi-Step Query Rewriting via Corpus Feedback for Conversational Search," https://omanscience.com/ar/articles/learning-multi-step-query-rewriting-via-corpus-feedback-for-conversational-search.