Image-Chat: Engaging Grounded Conversations

doi:10.18653/V1/2020.ACL-MAIN.219

Open AccessProceedings Article10.18653/V1/2020.ACL-MAIN.219

Image-Chat: Engaging Grounded Conversations

Kurt Shuster, +3 more

- 01 Jul 2020

- pp 2414-2429

81

TL;DR: Automatic metrics and human evaluations of engagingness show the efficacy of this approach, and state-of-the-art performance on the existing IGC task is obtained, and the best performing model is almost on par with humans on the Image-Chat test set.

Chat with Paper

AI Agents for this Paper

Find similar papers on Google Scholar, PubMed and Arxiv
Write a critical review of this paper
Analyze citations of this paper to find unaddressed research gaps

Citations

•Journal Article•10.1162/coli_a_00426

Deep Learning for Text Style Transfer: A Survey

Olga Vechtomova, +4 more

- 01 Jan 2022

- Computational Linguistics

TL;DR: Text style transfer is an important task in natural language generation, which aims to control certain attributes in the generated text, such as politeness, emotion, humor, and many others as mentioned in this paper .

...read moreread less

135

•Posted Content

Recent Advances in Deep Learning Based Dialogue Systems: A Systematic Survey.

Jinjie Ni, +5 more

- 10 May 2021

- arXiv: Computation and Language

TL;DR: In this paper, a survey of state-of-the-art research outcomes in dialogue systems is presented, focusing mainly on the deep learning-based dialogue systems, and the authors comprehensively review the evaluation methods and datasets for dialogue systems.

...read moreread less

135

•Posted Content•10.1145/1122445.1122456

Text is NOT Enough: Integrating Visual Impressions into Open-domain Dialogue Generation

Lei Shen, +4 more

- 13 Sep 2021

- arXiv: Computation and Language

TL;DR: In this paper, a co-attention encoder is used to generate a post representation with both visual and textual information, and then the response is generated based on the post and RVIs.

...read moreread less

110

•Posted Content

The Adapter-Bot: All-In-One Controllable Conversational Model

Andrea Madotto, +3 more

- 28 Aug 2020

- arXiv: Computation and Language

TL;DR: The Adapter-Bot is proposed, a dialogue model that uses a fixed backbone conversational model such as DialGPT and triggers on-demand dialogue skills via different adapters, thus allowing a continual integration of skills without retraining the entire model.

...read moreread less

73

•Posted Content

Open-Domain Conversational Agents: Current Progress, Open Problems, and Future Directions.

Stephen Roller, +15 more

- 22 Jun 2020

- arXiv: Computation and Language

TL;DR: The properties of continual learning, providing engaging content, and being well-behaved are discussed -- and how to measure success in providing them and their recommendations to the community are discussed.

...read moreread less

55

...

Expand

References

•Proceedings Article•10.1109/CVPR.2016.90

Deep Residual Learning for Image Recognition

Kaiming He, +3 more

- 27 Jun 2016

TL;DR: In this article, the authors proposed a residual learning framework to ease the training of networks that are substantially deeper than those used previously, which won the 1st place on the ILSVRC 2015 classification task.

...read moreread less

198.7K

•Proceedings Article

Attention is All you Need

Ashish Vaswani, +7 more

- 12 Jun 2017

TL;DR: This paper proposed a simple network architecture based solely on an attention mechanism, dispensing with recurrence and convolutions entirely and achieved state-of-the-art performance on English-to-French translation.

...read moreread less

94.2K

Preprint•10.48550/arxiv.1706.03762

Attention Is All You Need

Ashish Vaswani, +7 more

- 01 Jan 2017

Abstract: The dominant sequence transduction models are based on complex recurrent or convolutional neural networks in an encoder-decoder configuration. The best performing models also connect the encoder and decoder through an attention mechanism. We propose a new simple network architecture, the Transformer, based solely on attention mechanisms, dispensing with recurrence and convolutions entirely. Experiments on two machine translation tasks show these models to be superior in quality while being more parallelizable and requiring significantly less time to train. Our model achieves 28.4 BLEU on the WMT 2014 English-to-German translation task, improving over the existing best results, including ensembles by over 2 BLEU. On the WMT 2014 English-to-French translation task, our model establishes a new single-model state-of-the-art BLEU score of 41.8 after training for 3.5 days on eight GPUs, a small fraction of the training costs of the best models from the literature. We show that the Transformer generalizes well to other tasks by applying it successfully to English constituency parsing both with large and limited training data.

...read moreread less

51.8K

•Journal Article•10.1007/S11263-015-0816-Y

ImageNet Large Scale Visual Recognition Challenge

Olga Russakovsky, +11 more

- 01 Dec 2015

- International Journal of Computer Vision

TL;DR: The ImageNet Large Scale Visual Recognition Challenge (ILSVRC) as mentioned in this paper is a benchmark in object category classification and detection on hundreds of object categories and millions of images, which has been run annually from 2010 to present, attracting participation from more than fifty institutions.

...read moreread less

41.6K

•Proceedings Article•10.1109/CVPR.2017.634

Aggregated Residual Transformations for Deep Neural Networks

Saining Xie, +4 more

- 21 Jul 2017

TL;DR: ResNeXt as discussed by the authors is a simple, highly modularized network architecture for image classification, which is constructed by repeating a building block that aggregates a set of transformations with the same topology.

...read moreread less

11.2K