Kansallisarkisto / PaddleOCR_training

huggingface.co
Total runs: 0
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: December 03 2024
image-text-to-text

Introduction of PaddleOCR_training

Model Details of PaddleOCR_training

Training model from AIDA-project

This repository contains the model trained in AIDA-project in PaddleOCR training format. The model is trained on top of PaddleOCR-v3 latin model. It is trained on roughly 40 000 line images and 120 000 synthetic line images. The training data is mainly in Finnish, but contains some Swedish and English text and little French and German lines.

This repository is for finetuning our trained model, but you can find the inference model and more information about the model here https://github.com/project-AIDA/ . Additionally, the training data used for training this model can be found here https://huggingface.co/datasets/Kansallisarkisto/AIDA_ocr_training_data .

Model Training

In case you want to finetune the trained model, you should refer to the PaddleOCR docs here https://paddlepaddle.github.io/PaddleOCR/en/ppocr/model_train/recognition.html . This repository contains our checkpoints for our best performing model on Finnish language. Additionally it contains a config file for training the model. The necessary codes training can be found here https://github.com/PaddlePaddle/PaddleOCR/blob/main/README_en.md . You should download the codes and follow the installation instructions there

First you need prepare your data. You need textline and transcription combinations and you need to arrange them in PaddleOCR format. After that you can download the the model from this repository (include all best_accuracy.* files) and place them in a separate folder in your system. Then the last thing before training is to change the paths in the config file to correspond to your paths. I.e arguments called save_model_dir , pretrained_model and then in Train and Eval dataset data_dir and label_file_list .

After all that is done, you can start training. When run on the main folder of the downloaded github repository, the following command starts the training based on the configurations in the config file.

python3 tools/train.py -c configs/rec/PP-OCRv3/en_PP-OCRv3_rec.yml

Runs of Kansallisarkisto PaddleOCR_training on huggingface.co

0
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About PaddleOCR_training huggingface.co Model

More PaddleOCR_training license Visit here:

https://choosealicense.com/licenses/apache-2.0

PaddleOCR_training huggingface.co

PaddleOCR_training huggingface.co is an AI model on huggingface.co that provides PaddleOCR_training's model effect (), which can be used instantly with this Kansallisarkisto PaddleOCR_training model. huggingface.co supports a free trial of the PaddleOCR_training model, and also provides paid use of the PaddleOCR_training. Support call PaddleOCR_training model through api, including Node.js, Python, http.

Kansallisarkisto PaddleOCR_training online free

PaddleOCR_training huggingface.co is an online trial and call api platform, which integrates PaddleOCR_training's modeling effects, including api services, and provides a free online trial of PaddleOCR_training, you can try PaddleOCR_training online for free by clicking the link below.

Kansallisarkisto PaddleOCR_training online free url in huggingface.co:

https://huggingface.co/Kansallisarkisto/PaddleOCR_training

PaddleOCR_training install

PaddleOCR_training is an open source model from GitHub that offers a free installation service, and any user can find PaddleOCR_training on GitHub to install. At the same time, huggingface.co provides the effect of PaddleOCR_training install, users can directly use PaddleOCR_training installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

PaddleOCR_training install url in huggingface.co:

https://huggingface.co/Kansallisarkisto/PaddleOCR_training

Url of PaddleOCR_training

Provider of PaddleOCR_training huggingface.co

Kansallisarkisto
ORGANIZATIONS

Other API from Kansallisarkisto