Plug-and-Play Conversational Models

Andrea Madotto; Etsuko Ishii; Zhaojiang Lin; Sumanth Dathathri,; Pascale Fung

arXiv:2010.04344·cs.CL·October 12, 2020

Plug-and-Play Conversational Models

Andrea Madotto, Etsuko Ishii, Zhaojiang Lin, Sumanth Dathathri,, Pascale Fung

PDF

1 Repo

TL;DR

This paper introduces plug-and-play methods for controllable conversational response generation that do not require fine-tuning or dialogue-specific datasets, balancing control, fluency, and computational efficiency.

Contribution

It proposes novel plug-and-play techniques enabling attribute-controlled response generation without fine-tuning large language models or requiring dialogue datasets.

Findings

01

Effective control over response attributes demonstrated

02

High fluency maintained in generated responses

03

Approach reduces computational overhead during decoding

Abstract

There has been considerable progress made towards conversational models that generate coherent and fluent responses; however, this often involves training large language models on large dialogue datasets, such as Reddit. These large conversational models provide little control over the generated responses, and this control is further limited in the absence of annotated conversational datasets for attribute specific generation that can be used for fine-tuning the model. In this paper, we first propose and evaluate plug-and-play methods for controllable response generation, which does not require dialogue specific datasets and does not rely on fine-tuning a large model. While effective, the decoding procedure induces considerable computational overhead, rendering the conversational model unsuitable for interactive usage. To overcome this, we introduce an approach that does not require…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

andreamad8/PPCM
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.