Marco-LLM is a series of advanced multilingual language models designed to bridge the performance gap between high-resource languages and low-resource languages. This repository contains the Marco-LLM base language model with 7 billion parameters.
The model has undergone extensive multilingual continual pretraining on a diverse dataset containing over 5 trillion tokens, with a particular focus on enhancing performance in low-resource languages while maintaining strong capabilities in high-resource languages like English and Chinese.
Compared to state-of-the-art open-source language models, Marco-LLM demonstrates significant improvements in multilingual tasks, including machine translation, question answering, and reasoning across multiple languages.
For more details, please refer to our
Hugging Face page
.
Model Details
Marco-LLM includes a 7B parameter model based on the Transformer architecture. The key features of Marco-LLM are:
Multilingual Training: The model is trained on a large-scale multilingual dataset covering 29 languages, including both high-resource languages (e.g., English, Chinese) and low-resource languages (e.g., Kazakh, Nepali).
Enhanced Tokenizer: An improved tokenizer is used to better handle multilingual data, ensuring higher efficiency and accuracy in tokenization.
Post-Training: Marco-LLM supports various post-training methods, such as Supervised Fine-tuning (SFT) and Direct Preference Optimization (DPO), to further enhance performance for specific tasks and languages.
Usage
It is not advised to use the base language models for direct text generation tasks. Instead, it is recommended to apply post-training methods such as Supervised Fine-tuning (SFT), Reinforcement Learning with Human Feedback (RLHF), or continued pretraining to adapt the models for specific use cases.
Citation
If you find our work helpful, please give us a citation.
@article{unique_identifier,
title={Marco-LLM: Bridging Languages via Massive Multilingual Training for Cross-Lingual Enhancement},
journal={arXiv},
volume={},
number={2412.04003},
year={2024},
url={https://arxiv.org/abs/2412.04003}
}
Runs of AIDC-AI Marco-LLM-GLO on huggingface.co
59
Total runs
0
24-hour runs
-5
3-day runs
-21
7-day runs
-898
30-day runs
More Information About Marco-LLM-GLO huggingface.co Model
Marco-LLM-GLO huggingface.co is an AI model on huggingface.co that provides Marco-LLM-GLO's model effect (), which can be used instantly with this AIDC-AI Marco-LLM-GLO model. huggingface.co supports a free trial of the Marco-LLM-GLO model, and also provides paid use of the Marco-LLM-GLO. Support call Marco-LLM-GLO model through api, including Node.js, Python, http.
Marco-LLM-GLO huggingface.co is an online trial and call api platform, which integrates Marco-LLM-GLO's modeling effects, including api services, and provides a free online trial of Marco-LLM-GLO, you can try Marco-LLM-GLO online for free by clicking the link below.
AIDC-AI Marco-LLM-GLO online free url in huggingface.co:
Marco-LLM-GLO is an open source model from GitHub that offers a free installation service, and any user can find Marco-LLM-GLO on GitHub to install. At the same time, huggingface.co provides the effect of Marco-LLM-GLO install, users can directly use Marco-LLM-GLO installed effect in huggingface.co for debugging and trial. It also supports api for free installation.