jhu-clsp / roberta-large-eng-ara-128k

huggingface.co
Total runs: 57
24-hour runs: 0
7-day runs: 5
30-day runs: 3
Model's Last Updated: June 25 2023
fill-mask

Introduction of roberta-large-eng-ara-128k

Model Details of roberta-large-eng-ara-128k

An English-Arabic Bilingual Encoder

from transformers import AutoModelForMaskedLM, AutoTokenizer
tokenizer = AutoTokenizer.from_pretrained("jhu-clsp/roberta-large-eng-ara-128k")
model = AutoModelForMaskedLM.from_pretrained("jhu-clsp/roberta-large-eng-ara-128k")

roberta-large-eng-ara-128k is an English�Arabic bilingual encoders of 24-layer Transformers (d_model= 1024), the same size as XLM-R large. We use the same Common Crawl corpus as XLM-R for pretraining. Additionally, we also use English and Arabic Wikipedia, Arabic Gigaword (Parker et al., 2011), Arabic OSCAR (Ortiz Su�rez et al., 2020), Arabic News Corpus (El-Khair, 2016), and Arabic OSIAN (Zeroual et al.,2019). In total, we train with 9.2B words of Arabic text and 26.8B words of English text, more than either XLM-R (2.9B words/23.6B words) or GigaBERT v4 (Lan et al., 2020) (4.3B words/6.1B words). We build an English�Arabic joint vocabulary using SentencePiece (Kudo and Richardson, 2018) with size of 128K. We additionally enforce coverage of all Arabic characters after normalization.

Pretraining Detail

We pretrain each encoder with a batch size of 2048 sequences and 512 sequence length for 250K steps from scratch roughly 1/24 the amount of pretraining compute of XLM-R. Training takes 8 RTX6000 GPUs roughly three weeks. We follow the pretraining recipe of RoBERTa (Liu et al., 2019) and XLM-R. We omit the next sentence prediction task and use a learning rate of 2e-4, Adam optimizer, and linear warmup of 10K steps then decay linearly to 0, multilingual sampling alpha of 0.3, and the fairseq (Ott et al., 2019) implementation.

Citation

Please cite this paper for reference:

@inproceedings{yarmohammadi-etal-2021-everything,
    title = "Everything Is All It Takes: A Multipronged Strategy for Zero-Shot Cross-Lingual Information Extraction",
    author = "Yarmohammadi, Mahsa  and
      Wu, Shijie  and
      Marone, Marc  and
      Xu, Haoran  and
      Ebner, Seth  and
      Qin, Guanghui  and
      Chen, Yunmo and
      Guo, Jialiang and
      Harman, Craig  and
      Murray, Kenton and
      White, Aaron Steven  and
      Dredze, Mark and
      Van Durme, Benjamin",
    booktitle = "Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing",
    year = "2021",
}

Runs of jhu-clsp roberta-large-eng-ara-128k on huggingface.co

57
Total runs
0
24-hour runs
0
3-day runs
5
7-day runs
3
30-day runs

More Information About roberta-large-eng-ara-128k huggingface.co Model

More roberta-large-eng-ara-128k license Visit here:

https://choosealicense.com/licenses/mit

roberta-large-eng-ara-128k huggingface.co

roberta-large-eng-ara-128k huggingface.co is an AI model on huggingface.co that provides roberta-large-eng-ara-128k's model effect (), which can be used instantly with this jhu-clsp roberta-large-eng-ara-128k model. huggingface.co supports a free trial of the roberta-large-eng-ara-128k model, and also provides paid use of the roberta-large-eng-ara-128k. Support call roberta-large-eng-ara-128k model through api, including Node.js, Python, http.

roberta-large-eng-ara-128k huggingface.co Url

https://huggingface.co/jhu-clsp/roberta-large-eng-ara-128k

jhu-clsp roberta-large-eng-ara-128k online free

roberta-large-eng-ara-128k huggingface.co is an online trial and call api platform, which integrates roberta-large-eng-ara-128k's modeling effects, including api services, and provides a free online trial of roberta-large-eng-ara-128k, you can try roberta-large-eng-ara-128k online for free by clicking the link below.

jhu-clsp roberta-large-eng-ara-128k online free url in huggingface.co:

https://huggingface.co/jhu-clsp/roberta-large-eng-ara-128k

roberta-large-eng-ara-128k install

roberta-large-eng-ara-128k is an open source model from GitHub that offers a free installation service, and any user can find roberta-large-eng-ara-128k on GitHub to install. At the same time, huggingface.co provides the effect of roberta-large-eng-ara-128k install, users can directly use roberta-large-eng-ara-128k installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

roberta-large-eng-ara-128k install url in huggingface.co:

https://huggingface.co/jhu-clsp/roberta-large-eng-ara-128k

Url of roberta-large-eng-ara-128k

roberta-large-eng-ara-128k huggingface.co Url

Provider of roberta-large-eng-ara-128k huggingface.co

jhu-clsp
ORGANIZATIONS

Other API from jhu-clsp

huggingface.co

Total runs: 400.5K
Run Growth: 63.7K
Growth Rate: 15.69%
Updated:October 07 2025
huggingface.co

Total runs: 79.5K
Run Growth: 32.6K
Growth Rate: 43.45%
Updated:October 18 2025
huggingface.co

Total runs: 1.5K
Run Growth: -12.6K
Growth Rate: -856.91%
Updated:April 09 2025
huggingface.co

Total runs: 958
Run Growth: 355
Growth Rate: 36.60%
Updated:April 05 2023
huggingface.co

Total runs: 602
Run Growth: 90
Growth Rate: 14.49%
Updated:April 30 2024
huggingface.co

Total runs: 442
Run Growth: 227
Growth Rate: 46.04%
Updated:April 09 2025
huggingface.co

Total runs: 352
Run Growth: 283
Growth Rate: 82.27%
Updated:June 02 2024
huggingface.co

Total runs: 196
Run Growth: -200
Growth Rate: -101.01%
Updated:April 09 2025
huggingface.co

Total runs: 90
Run Growth: -31
Growth Rate: -34.07%
Updated:August 25 2022
huggingface.co

Total runs: 69
Run Growth: 52
Growth Rate: 74.29%
Updated:April 09 2025
huggingface.co

Total runs: 52
Run Growth: 14
Growth Rate: 28.57%
Updated:September 18 2023
huggingface.co

Total runs: 42
Run Growth: 12
Growth Rate: 27.91%
Updated:April 09 2025
huggingface.co

Total runs: 29
Run Growth: 4
Growth Rate: 13.33%
Updated:April 09 2025