Disentangling Semantics and Syntax in Sentence Embeddings with Pre-trained...

Disentangling semantics and syntax is an effective approach in text style transfer as well as unsupervised sentence embedding learning. This paper uses separate encoders for semantics and syntax to learn paraphrasing on ParaNMT dataset, resulting in a sentence semantics encoder disentangled from syntax.

The syntax encoder operates on linearized and de-lexicalized constituency trees. They further applies adversarial learning to encourage disentanglement by adding a classifier to predict BoW of level constituent tags from the semantic embedding, with the following objective:

\min\limits_{E_{sem}, E_{syn}, D_{dec}} \left(\max\limits_{D_{dis}} \left(\mathcal{L}_{para}-\lambda_{adv}\mathcal{L}_{adv}\right)\right)

where \lambda_{adv} is a hyperparameter to balance loss terms. In each iteration, the D_{dis} is updated by considering the inner optimization so \mathcal{L}_{para} is constant to D_{dis} (that’s why it can be placed inside the \max), and then update E_{sem}, E_{syn} and D_{dec} by considering the outer optimization.


  • Overall quality is good, especially the adversarial part.
  • The syntax encoder is quite crude. There exists many tree RNNs for constituent trees.
