有条件产生的提问蓝图

论文标题

有条件产生的提问蓝图

Conditional Generation with a Question-Answering Blueprint

论文作者

Narayan, Shashi, Maynez, Joshua, Amplayo, Reinald Kim, Ganchev, Kuzman, Louis, Annie, Huot, Fantine, Sandholm, Anders, Das, Dipanjan, Lapata, Mirella

论文摘要

传达相关和忠实信息的能力对于有条件生成的许多任务至关重要，但对于神经seq-to-seq模型仍然难以捉摸，这些模型通常会揭示幻觉，并且无法正确涵盖重要细节。在这项工作中，我们主张规划作为有用的中间表示，以使有条件的一代减少不透明，扎根。我们的作品提出了将文本计划作为一系列提问（QA）对的新概念化。我们用QA蓝图作为内容选择的代理（即〜要说的）和计划（即〜以什么顺序）来增强现有数据集（例如，用于摘要）。我们通过利用最先进的问题生成技术并将输入输出对自动获取蓝图，并将其转换为输入薄荷输出输出元素。我们开发了基于变压器的模型，每个模型都在它们如何将蓝图结合到生成的输出中（例如，作为全局计划或迭代）。跨指标和数据集的评估表明，蓝图模型比不采取计划并允许对生成输出进行更严格控制的替代方案更为事实。

The ability to convey relevant and faithful information is critical for many tasks in conditional generation and yet remains elusive for neural seq-to-seq models whose outputs often reveal hallucinations and fail to correctly cover important details. In this work, we advocate planning as a useful intermediate representation for rendering conditional generation less opaque and more grounded. Our work proposes a new conceptualization of text plans as a sequence of question-answer (QA) pairs. We enhance existing datasets (e.g., for summarization) with a QA blueprint operating as a proxy for both content selection (i.e.,~what to say) and planning (i.e.,~in what order). We obtain blueprints automatically by exploiting state-of-the-art question generation technology and convert input-output pairs into input-blueprint-output tuples. We develop Transformer-based models, each varying in how they incorporate the blueprint in the generated output (e.g., as a global plan or iteratively). Evaluation across metrics and datasets demonstrates that blueprint models are more factual than alternatives which do not resort to planning and allow tighter control of the generation output.

下载PDF全文

下载文献需遵守相关版权规定

论文标题