Guided Dialog Policy Learning: Reward Estimation for Multi-Domain Task-Oriented Dialog

Ryuichi Takanobu, Hanlin Zhu, Minlie Huang · 2019

Ryuichi Takanobu, Hanlin Zhu, Minlie Huang. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP). 2019.

Read the paper · More papers on PaperTik