2000 character limit reached
Modeling Real-Time Interactive Conversations as Timed Diarized Transcripts (2405.13203v1)
Published 21 May 2024 in cs.LG and cs.CL
Abstract: Chatbots built upon LLMs have exploded in popularity, but they have largely been limited to synchronous, turn-by-turn dialogues. In this paper we present a simple yet general method to simulate real-time interactive conversations using pretrained text-only LLMs, by modeling timed diarized transcripts and decoding them with causal rejection sampling. We demonstrate the promise of this method with two case studies: instant messenger dialogues and spoken conversations, which require generation at about 30 tok/s and 20 tok/s respectively to maintain real-time interactivity. These capabilities can be added into LLMs using relatively little data and run on commodity hardware.
- Garrett Tanzer (11 papers)
- Gustaf Ahdritz (5 papers)
- Luke Melas-Kyriazi (22 papers)