PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 20192,336 citationsOpen Access

LXMERT: Learning Cross-Modality Encoder Representations from Transformers

HTHao TanMBMohit Bansal

Key Points

  • The study aims to develop a model that improves cross-modality learning for natural language processing tasks.
  • Introduced LXMERT, a model leveraging transformer architectures for cross-modality tasks.
  • Evaluated on various natural language processing benchmarks to demonstrate effectiveness.
  • Utilized multimodal inputs to enhance representation learning.
  • Showed significant improvements in benchmark performance metrics over existing models.
  • Achieved superior results in tasks involving both visual and textual data integration.

Abstract

Hao Tan, Mohit Bansal. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP). 2019.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Tan et al. (2019) studied this question.

synapsesocial.com/papers/6a0024fc4716aad0cc8598f4https://doi.org/10.18653/v1/d19-1514
Ask AI
Helpful
Bookmark
Share
View Full Paper