If you use the model, please reference this work in your paper
:
@inproceedings{tedeschi-etal-2021-wikineural-combined,
title = "{W}iki{NE}u{R}al: {C}ombined Neural and Knowledge-based Silver Data Creation for Multilingual {NER}",
author = "Tedeschi, Simone and
Maiorca, Valentino and
Campolungo, Niccol{\`o} and
Cecconi, Francesco and
Navigli, Roberto",
booktitle = "Findings of the Association for Computational Linguistics: EMNLP 2021",
month = nov,
year = "2021",
address = "Punta Cana, Dominican Republic",
publisher = "Association for Computational Linguistics",
url = "https://aclanthology.org/2021.findings-emnlp.215",
pages = "2521--2533",
abstract = "Multilingual Named Entity Recognition (NER) is a key intermediate task which is needed in many areas of NLP. In this paper, we address the well-known issue of data scarcity in NER, especially relevant when moving to a multilingual scenario, and go beyond current approaches to the creation of multilingual silver data for the task. We exploit the texts of Wikipedia and introduce a new methodology based on the effective combination of knowledge-based approaches and neural models, together with a novel domain adaptation technique, to produce high-quality training corpora for NER. We evaluate our datasets extensively on standard benchmarks for NER, yielding substantial improvements up to 6 span-based F1-score points over previous state-of-the-art systems for data creation.",
}
You can use this model with Transformers
pipeline
for NER.
from transformers import AutoTokenizer, AutoModelForTokenClassification
from transformers import pipeline
tokenizer = AutoTokenizer.from_pretrained("Babelscape/wikineural-multilingual-ner")
model = AutoModelForTokenClassification.from_pretrained("Babelscape/wikineural-multilingual-ner")
nlp = pipeline("ner", model=model, tokenizer=tokenizer, grouped_entities=True)
example = "My name is Wolfgang and I live in Berlin"
ner_results = nlp(example)
print(ner_results)
Limitations and bias
This model is trained on WikiNEuRal, a state-of-the-art dataset for Multilingual NER automatically derived from Wikipedia. Therefore, it might not generalize well to all textual genres (e.g. news). On the other hand, models trained only on news articles (e.g. only on CoNLL03) have been proven to obtain much lower scores on encyclopedic articles. To obtain more robust systems, we encourage you to train a system on the combination of WikiNEuRal with other datasets (e.g. WikiNEuRal + CoNLL).
wikineural-multilingual-ner huggingface.co is an AI model on huggingface.co that provides wikineural-multilingual-ner's model effect (), which can be used instantly with this Babelscape wikineural-multilingual-ner model. huggingface.co supports a free trial of the wikineural-multilingual-ner model, and also provides paid use of the wikineural-multilingual-ner. Support call wikineural-multilingual-ner model through api, including Node.js, Python, http.
wikineural-multilingual-ner huggingface.co is an online trial and call api platform, which integrates wikineural-multilingual-ner's modeling effects, including api services, and provides a free online trial of wikineural-multilingual-ner, you can try wikineural-multilingual-ner online for free by clicking the link below.
Babelscape wikineural-multilingual-ner online free url in huggingface.co:
wikineural-multilingual-ner is an open source model from GitHub that offers a free installation service, and any user can find wikineural-multilingual-ner on GitHub to install. At the same time, huggingface.co provides the effect of wikineural-multilingual-ner install, users can directly use wikineural-multilingual-ner installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
wikineural-multilingual-ner install url in huggingface.co: