README.md 1.02 KB
Newer Older
1
2
3
4
5
6
7
---
language:
- multilingual
---

# bert-base-multilingual-cased-sentence

8
Sentence Multilingual BERT \(101 languages, cased, 12鈥憀ayer, 768鈥慼idden, 12鈥慼eads, 180M parameters\) is a representation鈥慴ased sentence encoder for 101 languages of Multilingual BERT. It is initialized with Multilingual BERT and then fine鈥憈uned on english MultiNLI\[1\] and on dev set of multilingual XNLI\[2\]. Sentence representations are mean pooled token embeddings in the same manner as in Sentence鈥態ERT\[3\].
9
10


11
\[1\]: Williams A., Nangia N. & Bowman S. \(2017\) A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference. arXiv preprint [arXiv:1704.05426](https://arxiv.org/abs/1704.05426)
12

13
\[2\]: Williams A., Bowman S. \(2018\) XNLI: Evaluating Cross-lingual Sentence Representations. arXiv preprint [arXiv:1809.05053](https://arxiv.org/abs/1809.05053)
14

15
\[3\]: N. Reimers, I. Gurevych \(2019\) Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. arXiv preprint [arXiv:1908.10084](https://arxiv.org/abs/1908.10084)