AliFartout / Roberta-fa-en-ner

huggingface.co
Total runs: 45.9K
24-hour runs: -1.8K
7-day runs: 3.4K
30-day runs: 44.8K
Model's Last Updated: September 02 2023
token-classification

Introduction of Roberta-fa-en-ner

Model Details of Roberta-fa-en-ner

NER Model using Roberta

This markdown presents a Robustly Optimized BERT Pretraining Approach (RoBERTa) model trained on a combination of two diverse datasets for two languages: English and Persian. The English dataset used is CoNLL 2003 , while the Persian dataset is PEYMA-ARMAN-Mixed , a fusion of the "PEYAM" and "ARMAN" datasets, both popular for Named Entity Recognition (NER) tasks.

The model training pipeline involves the following steps:

Data Preparation: Cleaning, aligning, and mixing data from the two datasets. Data Loading: Loading the prepared data for subsequent processing. Tokenization: Utilizing tokenization to prepare the text data for model input. Token Splitting: Handling token splitting (e.g., "jack" becomes "_ja _ck") and using "-100" for optimization and ignoring certain tokens. Model Reconstruction: Adapting the RoBERTa model for token classification in NER tasks. Model Training: Training the reconstructed model on the combined dataset and evaluating its performance. The model's performance, as shown in the table below, demonstrates promising results:

Epoch Training Loss Validation Loss F1 Recall Precision Accuracy
1 0.072600 0.038918 89.5% 0.906680 0.883703 0.987799
2 0.027600 0.030184 92.3% 0.933840 0.915573 0.991334
3 0.013500 0.030962 94% 0.946840 0.933740 0.992702
4 0.006600 0.029897 94.8% 0.955207 0.941990 0.993574

The model achieves an impressive F1-score of almost 95%.

To use the model, the following Python code snippet can be employed:

from transformers import AutoConfig, AutoTokenizer, AutoModel

config = AutoConfig.from_pretrained("AliFartout/Roberta-fa-en-ner")
tokenizer = AutoTokenizer.from_pretrained("AliFartout/Roberta-fa-en-ner")
model = AutoModel.from_pretrained("AliFartout/Roberta-fa-en-ner")

By following this approach, you can seamlessly access and incorporate the trained multilingual NER model into various Natural Language Processing tasks.

Runs of AliFartout Roberta-fa-en-ner on huggingface.co

45.9K
Total runs
-1.8K
24-hour runs
-770
3-day runs
3.4K
7-day runs
44.8K
30-day runs

More Information About Roberta-fa-en-ner huggingface.co Model

More Roberta-fa-en-ner license Visit here:

https://choosealicense.com/licenses/apache-2.0

Roberta-fa-en-ner huggingface.co

Roberta-fa-en-ner huggingface.co is an AI model on huggingface.co that provides Roberta-fa-en-ner's model effect (), which can be used instantly with this AliFartout Roberta-fa-en-ner model. huggingface.co supports a free trial of the Roberta-fa-en-ner model, and also provides paid use of the Roberta-fa-en-ner. Support call Roberta-fa-en-ner model through api, including Node.js, Python, http.

Roberta-fa-en-ner huggingface.co Url

https://huggingface.co/AliFartout/Roberta-fa-en-ner

AliFartout Roberta-fa-en-ner online free

Roberta-fa-en-ner huggingface.co is an online trial and call api platform, which integrates Roberta-fa-en-ner's modeling effects, including api services, and provides a free online trial of Roberta-fa-en-ner, you can try Roberta-fa-en-ner online for free by clicking the link below.

AliFartout Roberta-fa-en-ner online free url in huggingface.co:

https://huggingface.co/AliFartout/Roberta-fa-en-ner

Roberta-fa-en-ner install

Roberta-fa-en-ner is an open source model from GitHub that offers a free installation service, and any user can find Roberta-fa-en-ner on GitHub to install. At the same time, huggingface.co provides the effect of Roberta-fa-en-ner install, users can directly use Roberta-fa-en-ner installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Roberta-fa-en-ner install url in huggingface.co:

https://huggingface.co/AliFartout/Roberta-fa-en-ner

Url of Roberta-fa-en-ner

Roberta-fa-en-ner huggingface.co Url

Provider of Roberta-fa-en-ner huggingface.co

AliFartout
ORGANIZATIONS

Other API from AliFartout