Abstract
One of the core challenges in Visual Dialogue problems is asking the question that will provide the most useful information towards achieving the required objective. Encouraging an agent to ask the right questions is difficult because we don't know a-priori what information the agent will need to achieve its task, and we don't have an explicit model of what it knows already. We propose a solution to this problem based on a Bayesian model of the uncertainty in the implicit model maintained by the visual dialogue agent, and in the function used to select an appropriate output. By selecting the question that minimises the predicted regret with respect to this implicit model the agent actively reduces ambiguity. The Bayesian model of uncertainty also enables a principled method for identifying when enough information has been acquired, and an action should be selected. We evaluate our approach on two goal-oriented dialogue datasets, one for visual-based collaboration task and the other for a negotiation-based task. Our uncertainty-aware information-seeking model outperforms its counterparts in these two challenging problems.
| Original language | English |
|---|---|
| Title of host publication | Proceedings - 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2019 |
| Editors | Abhinav Gupta, Derek Hoiem, Gang Hua, Zhuowen Tu |
| Place of Publication | Piscataway NJ USA |
| Publisher | IEEE, Institute of Electrical and Electronics Engineers |
| Pages | 4150-4159 |
| Number of pages | 10 |
| ISBN (Electronic) | 9781728132938 |
| ISBN (Print) | 9781728132945 |
| DOIs | |
| Publication status | Published - 2019 |
| Externally published | Yes |
| Event | IEEE Conference on Computer Vision and Pattern Recognition 2019 - Long Beach, United States of America Duration: 16 Jun 2019 → 20 Jun 2019 Conference number: 32nd http://cvpr2019.thecvf.com/ https://ieeexplore.ieee.org/xpl/conhome/8938205/proceeding (Proceedings) |
Conference
| Conference | IEEE Conference on Computer Vision and Pattern Recognition 2019 |
|---|---|
| Abbreviated title | CVPR 2019 |
| Country/Territory | United States of America |
| City | Long Beach |
| Period | 16/06/19 → 20/06/19 |
| Internet address |
Keywords
- Deep Learning
- Vision + Language
Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver