Start Over

Learning Bidirectional Action-Language Translation with Limited Supervision and Incongruent Input

Authors :: Özdemir, Ozan
Kerzel, Matthias
Weber, Cornelius
Lee, Jae Hee
Hafez, Muhammad Burhan
Bruns, Patrick
Wermter, Stefan
Source :: Applied Artificial Intelligence Volume 37, 2023 - Issue 1
Publication Year :: 2023
Abstract: Human infant learning happens during exploration of the environment, by interaction with objects, and by listening to and repeating utterances casually, which is analogous to unsupervised learning. Only occasionally, a learning infant would receive a matching verbal description of an action it is committing, which is similar to supervised learning. Such a learning mechanism can be mimicked with deep learning. We model this weakly supervised learning paradigm using our Paired Gated Autoencoders (PGAE) model, which combines an action and a language autoencoder. After observing a performance drop when reducing the proportion of supervised training, we introduce the Paired Transformed Autoencoders (PTAE) model, using Transformer-based crossmodal attention. PTAE achieves significantly higher accuracy in language-to-action and action-to-language translations, particularly in realistic but difficult cases when only few supervised training samples are available. We also test whether the trained model behaves realistically with conflicting multimodal input. In accordance with the concept of incongruence in psychology, conflict deteriorates the model output. Conflicting action input has a more severe impact than conflicting language input, and more conflicting features lead to larger interference. PTAE can be trained on mostly unlabelled data where labeled data is scarce, and it behaves plausibly when tested with incongruent input.<br />Comment: Published in: Applied Artificial Intelligence, 37:1, 2179167

Subjects :: Computer Science - Computation and Language
Computer Science - Artificial Intelligence
Computer Science - Neural and Evolutionary Computing
Computer Science - Robotics

Details

Database :: arXiv
Journal :: Applied Artificial Intelligence Volume 37, 2023 - Issue 1
Publication Type :: Report
Accession number :: edsarx.2301.03353
Document Type :: Working Paper
Full Text :: https://doi.org/10.1080/08839514.2023.2179167

Full Text Access

View/download PDF

Tools

Email
Cite

Printer

Authors Abstract Subjects Details

Searchworks

Select search scope, currently: Articles

Catalog

books, media & more in Jio Institute collections

Articles

journal articles & other e-resources

Learning Bidirectional Action-Language Translation with Limited Supervision and Incongruent Input

Abstract

Subjects

Details

Tools

Searchworks

Select search scope, currently: Articles Catalog books, media & more in Jio Institute collections Articles journal articles & other e-resources

Learning Bidirectional Action-Language Translation with Limited Supervision and Incongruent Input

Abstract

Subjects

Details

Tools

Select search scope, currently: Articles

Catalog

books, media & more in Jio Institute collections

Articles

journal articles & other e-resources