Hi authors:
This is a great work and I have some questions about the AE training process. In the paper, you mentioned that the training contains two stages: unsupervised training with unlabeled data and supervised training with labeled data. For the first stage, what exact loss do you use? Is there only the reconstruction loss for the autoencoder training, or like the paper you cited 'Toward Controlled Generation of Text', you have an additional discrimitator to do the GAN-style training? Which part of the code is used for the unsupervised training?
Best,
Guangyu
Hi authors:
This is a great work and I have some questions about the AE training process. In the paper, you mentioned that the training contains two stages: unsupervised training with unlabeled data and supervised training with labeled data. For the first stage, what exact loss do you use? Is there only the reconstruction loss for the autoencoder training, or like the paper you cited 'Toward Controlled Generation of Text', you have an additional discrimitator to do the GAN-style training? Which part of the code is used for the unsupervised training?
Best,
Guangyu