TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/XLM-E: Cross-lingual Language Model Pre-training via ELECTRA

XLM-E: Cross-lingual Language Model Pre-training via ELECTRA

Zewen Chi, Shaohan Huang, Li Dong, Shuming Ma, Bo Zheng, Saksham Singhal, Payal Bajaj, Xia Song, Xian-Ling Mao, Heyan Huang, Furu Wei

2021-06-30ACL 2022 5TranslationLanguage ModellingZero-Shot Cross-Lingual Transfer
PaperPDFCodeCode(official)Code

Abstract

In this paper, we introduce ELECTRA-style tasks to cross-lingual language model pre-training. Specifically, we present two pre-training tasks, namely multilingual replaced token detection, and translation replaced token detection. Besides, we pretrain the model, named as XLM-E, on both multilingual and parallel corpora. Our model outperforms the baseline models on various cross-lingual understanding tasks with much less computation cost. Moreover, analysis shows that XLM-E tends to obtain better cross-lingual transferability.

Results

TaskDatasetMetricValueModel
Cross-LingualXTREMEAvg85.5Turing ULR v6
Cross-LingualXTREMEQuestion Answering77.1Turing ULR v6
Cross-LingualXTREMESentence Retrieval94.4Turing ULR v6
Cross-LingualXTREMESentence-pair Classification91Turing ULR v6
Cross-LingualXTREMEStructured Prediction83.8Turing ULR v6
Cross-LingualXTREMEAvg84.5Turing ULR v5
Cross-LingualXTREMEQuestion Answering76.3Turing ULR v5
Cross-LingualXTREMESentence Retrieval93.7Turing ULR v5
Cross-LingualXTREMESentence-pair Classification90.3Turing ULR v5
Cross-LingualXTREMEStructured Prediction81.7Turing ULR v5
Cross-Lingual TransferXTREMEAvg85.5Turing ULR v6
Cross-Lingual TransferXTREMEQuestion Answering77.1Turing ULR v6
Cross-Lingual TransferXTREMESentence Retrieval94.4Turing ULR v6
Cross-Lingual TransferXTREMESentence-pair Classification91Turing ULR v6
Cross-Lingual TransferXTREMEStructured Prediction83.8Turing ULR v6
Cross-Lingual TransferXTREMEAvg84.5Turing ULR v5
Cross-Lingual TransferXTREMEQuestion Answering76.3Turing ULR v5
Cross-Lingual TransferXTREMESentence Retrieval93.7Turing ULR v5
Cross-Lingual TransferXTREMESentence-pair Classification90.3Turing ULR v5
Cross-Lingual TransferXTREMEStructured Prediction81.7Turing ULR v5

Related Papers

Visual-Language Model Knowledge Distillation Method for Image Quality Assessment2025-07-21A Translation of Probabilistic Event Calculus into Markov Decision Processes2025-07-17Making Language Model a Hierarchical Classifier and Generator2025-07-17VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning2025-07-17The Generative Energy Arena (GEA): Incorporating Energy Awareness in Large Language Model (LLM) Human Evaluations2025-07-17Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities2025-07-17Assay2Mol: large language model-based drug design using BioAssay context2025-07-16Describe Anything Model for Visual Question Answering on Text-rich Images2025-07-16