PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
December 19, 2019217 citationsOpen Access

BERTje: A Dutch BERT Model

View Full Paper
WVWietse de VriesUniversity Medical Center GroningenACAndreas van CranenburghUniversity of GroningenABArianna BisazzaUniversity of Groningen

Key Points

Key points are not available for this paper at this time.

Abstract

The transformer-based pre-trained language model BERT has helped to improve state-of-the-art performance on many natural language processing (NLP) tasks. Using the same architecture and parameters, we developed and evaluated a monolingual Dutch BERT model called BERTje. Compared to the multilingual BERT model, which includes Dutch but is only based on Wikipedia text, BERTje is based on a large and diverse dataset of 2.4 billion tokens. BERTje consistently outperforms the equally-sized multilingual BERT model on downstream NLP tasks (part-of-speech tagging, named-entity recognition, semantic role labeling, and sentiment analysis). Our pre-trained Dutch BERT model is made available at https://github.com/wietsedv/bertje.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Vries et al. (2019) studied this question.

synapsesocial.com/papers/6a0ec24d950456576347bae4https://doi.org/10.48550/arxiv.1912.09582
Ask AI
Helpful
Bookmark
Share
View Full Paper