The model can be used directly (without a language model) as follows:
from sentence_transformers import SentenceTransformer
model = SentenceTransformer("Lajavaness/sentence-flaubert-base")
sentences = ["Un avion est en train de décoller.",
"Un homme joue d'une grande flûte.",
"Un homme étale du fromage râpé sur une pizza.",
"Une personne jette un chat au plafond.",
"Une personne est en train de plier un morceau de papier.",
]
embeddings = model.encode(sentences)
Evaluation
The model can be evaluated as follows on the French test data of stsb.
from sentence_transformers import SentenceTransformer
from sentence_transformers.readers import InputExample
from sentence_transformers.evaluation import EmbeddingSimilarityEvaluator
from datasets import load_dataset
defconvert_dataset(dataset):
dataset_samples=[]
for df in dataset:
score = float(df['similarity_score'])/5.0# Normalize score to range 0 ... 1
inp_example = InputExample(texts=[df['sentence1'],
df['sentence2']], label=score)
dataset_samples.append(inp_example)
return dataset_samples
# Loading the dataset for evaluation
df_dev = load_dataset("stsb_multi_mt", name="fr", split="dev")
df_test = load_dataset("stsb_multi_mt", name="fr", split="test")
# Convert the dataset for evaluation# For Dev set:
dev_samples = convert_dataset(df_dev)
val_evaluator = EmbeddingSimilarityEvaluator.from_input_examples(dev_samples, name='sts-dev')
val_evaluator(model, output_path="./")
# For Test set:
test_samples = convert_dataset(df_test)
test_evaluator = EmbeddingSimilarityEvaluator.from_input_examples(test_samples, name='sts-test')
test_evaluator(model, output_path="./")
Test Result
:
The performance is measured using Pearson and Spearman correlation on the sts-benchmark:
@article{reimers2019sentence,
title={Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks},
author={Nils Reimers, Iryna Gurevych},
journal={https://arxiv.org/abs/1908.10084},
year={2019}
}
@article{martin2020camembert,
title={CamemBERT: a Tasty French Language Mode},
author={Martin, Louis and Muller, Benjamin and Su{\'a}rez, Pedro Javier Ortiz and Dupont, Yoann and Romary, Laurent and de la Clergerie, {\'E}ric Villemonte and Seddah, Djam{\'e} and Sagot, Beno{\^\i}t},
journal={Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics},
year={2020}
}
@article{thakur2020augmented,
title={Augmented SBERT: Data Augmentation Method for Improving Bi-Encoders for Pairwise Sentence Scoring Tasks},
author={Thakur, Nandan and Reimers, Nils and Daxenberger, Johannes and Gurevych, Iryna},
journal={arXiv e-prints},
pages={arXiv--2010},
year={2020}
Runs of Lajavaness sentence-flaubert-base on huggingface.co
36.2K
Total runs
0
24-hour runs
0
3-day runs
3.8K
7-day runs
35.1K
30-day runs
More Information About sentence-flaubert-base huggingface.co Model
sentence-flaubert-base huggingface.co is an AI model on huggingface.co that provides sentence-flaubert-base's model effect (), which can be used instantly with this Lajavaness sentence-flaubert-base model. huggingface.co supports a free trial of the sentence-flaubert-base model, and also provides paid use of the sentence-flaubert-base. Support call sentence-flaubert-base model through api, including Node.js, Python, http.
sentence-flaubert-base huggingface.co is an online trial and call api platform, which integrates sentence-flaubert-base's modeling effects, including api services, and provides a free online trial of sentence-flaubert-base, you can try sentence-flaubert-base online for free by clicking the link below.
Lajavaness sentence-flaubert-base online free url in huggingface.co:
sentence-flaubert-base is an open source model from GitHub that offers a free installation service, and any user can find sentence-flaubert-base on GitHub to install. At the same time, huggingface.co provides the effect of sentence-flaubert-base install, users can directly use sentence-flaubert-base installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
sentence-flaubert-base install url in huggingface.co: