Language as a latent sequence: Deep latent variable models for semi-supervised paraphrase generation

Aduragba, Olanrewaju Tahir; Cristea, Alexandra I; Harit, Anoushka; Moubayed, Noura Al; Shi, Lei; Sun, Zhongtian; Yu, Jialin

Language as a latent sequence: Deep latent variable models for semi-supervised paraphrase generation

Authors: Olanrewaju Tahir Aduragba
Alexandra I Cristea
Anoushka Harit
Noura Al Moubayed
Lei Shi
Zhongtian Sun
Jialin Yu
Publication date: 1 January 2023
Publisher: 'Elsevier BV'

Abstract

This paper explores deep latent variable models for semi-supervised paraphrase generation, where the missing target pair for unlabelled data is modelled as a latent paraphrase sequence. We present a novel unsupervised model named variational sequence auto-encoding reconstruction (VSAR), which performs latent sequence inference given an observed text. To leverage information from text pairs, we additionally introduce a novel supervised model we call dual directional learning (DDL), which is designed to integrate with our proposed VSAR model. Combining VSAR with DDL (DDL+VSAR) enables us to conduct semi-supervised learning. Still, the combined model suffers from a cold-start problem. To further combat this issue, we propose an improved weight initialisation solution, leading to a novel two-stage training scheme we call knowledge-reinforced-learning (KRL). Our empirical evaluations suggest that the combined model yields competitive performance against the state-of-the-art supervised baselines on complete data. Furthermore, in scenarios where only a fraction of the labelled pairs are available, our combined model consistently outperforms the strong supervised model baseline (DDL) by a significant margin ( ; Wilcoxon test). Our code is publicly available at https://github.com/jialin-yu/latent-sequence-paraphrase

Similar works

Full text

Open in the Core reader

Download PDF

Available Versions

UCL Discovery

oai:eprints.ucl.ac.uk.OAI2:101...

Last time updated on 19/06/2023