lingtrain / labse-chuvash-3

huggingface.co
Total runs: 12
24-hour runs: 0
7-day runs: -1
30-day runs: -8
Model's Last Updated: August 25 2025
sentence-similarity

Introduction of labse-chuvash-3

Model Details of labse-chuvash-3

SentenceTransformer based on sentence-transformers/LaBSE

This is a sentence-transformers model finetuned from sentence-transformers/LaBSE . It maps sentences & paragraphs to a 768-dimensional dense vector space and can be used for semantic textual similarity, semantic search, paraphrase mining, text classification, clustering, and more.

Model Details
Model Description
  • Model Type: Sentence Transformer
  • Base model: sentence-transformers/LaBSE
  • Maximum Sequence Length: 256 tokens
  • Output Dimensionality: 768 dimensions
  • Similarity Function: Cosine Similarity
Model Sources
Full Model Architecture
SentenceTransformer(
  (0): Transformer({'max_seq_length': 256, 'do_lower_case': False}) with Transformer model: BertModel 
  (1): Pooling({'word_embedding_dimension': 768, 'pooling_mode_cls_token': True, 'pooling_mode_mean_tokens': False, 'pooling_mode_max_tokens': False, 'pooling_mode_mean_sqrt_len_tokens': False, 'pooling_mode_weightedmean_tokens': False, 'pooling_mode_lasttoken': False, 'include_prompt': True})
  (2): Dense({'in_features': 768, 'out_features': 768, 'bias': True, 'activation_function': 'torch.nn.modules.activation.Tanh'})
  (3): Normalize()
)
Usage
Direct Usage (Sentence Transformers)

First install the Sentence Transformers library:

pip install -U sentence-transformers

Then you can load this model and run inference.

from sentence_transformers import SentenceTransformer

# Download from the 🤗 Hub
model = SentenceTransformer("sentence_transformers_model_id")
# Run inference
sentences = [
    'Вӑл пӗлет: ҫак карапӑн командирӗ ҫамрӑк моряк, ӗлӗк артековец пулнӑскер, хӑйне вӗрентсе ӳстернӗ лагере асра тытса халӗ те тав туса саламлать.',
    'Он уже знал, что кораблем этим командует молодой моряк-командир, сам когда-то бывший артековец и поныне хранящий благодарную память о лагере.',
    'И разведчики это поняли.',
]
embeddings = model.encode(sentences)
print(embeddings.shape)
# [3, 768]

# Get the similarity scores for the embeddings
similarities = model.similarity(embeddings, embeddings)
print(similarities.shape)
# [3, 3]
Training Details
Training Dataset
Unnamed Dataset
  • Size: 1,455,347 training samples
  • Columns: sentence_0 , sentence_1 , and label
  • Approximate statistics based on the first 1000 samples:
    sentence_0 sentence_1 label
    type string string float
    details
    • min: 3 tokens
    • mean: 22.57 tokens
    • max: 190 tokens
    • min: 3 tokens
    • mean: 22.28 tokens
    • max: 207 tokens
    • min: 1.0
    • mean: 1.0
    • max: 1.0
  • Samples:
    sentence_0 sentence_1 label
    Каяссипе каяс марри ҫинчен шухӑшланӑ ҫӗртех Петян каймалла пулнӑ, мӗншӗн тесен ачасем чылай малалла утнӑ ӗнтӗ. Так что, когда в страшной борьбе с совестью победа осталась все-таки на стороне Пети, а совесть была окончательно раздавлена, оказалось, что мальчики зашли уже довольно далеко. 1.0
    — Чавсаран? — тӗлӗнчӗ Ван-Конет. — Локоть? — удивился Ван-Конет. 1.0
    Юлашкинчен пирӗн гаубицӑсем те ӗҫе тытӑнчӗҫ. Наконец открыли огонь и наши гаубицы. 1.0
  • Loss: MultipleNegativesRankingLoss with these parameters:
    {
        "scale": 20.0,
        "similarity_fct": "cos_sim"
    }
    
Training Hyperparameters
Non-Default Hyperparameters
  • eval_strategy : steps
  • per_device_train_batch_size : 20
  • per_device_eval_batch_size : 20
  • num_train_epochs : 1
  • fp16 : True
  • multi_dataset_batch_sampler : round_robin
All Hyperparameters
Click to expand
  • overwrite_output_dir : False
  • do_predict : False
  • eval_strategy : steps
  • prediction_loss_only : True
  • per_device_train_batch_size : 20
  • per_device_eval_batch_size : 20
  • per_gpu_train_batch_size : None
  • per_gpu_eval_batch_size : None
  • gradient_accumulation_steps : 1
  • eval_accumulation_steps : None
  • torch_empty_cache_steps : None
  • learning_rate : 5e-05
  • weight_decay : 0.0
  • adam_beta1 : 0.9
  • adam_beta2 : 0.999
  • adam_epsilon : 1e-08
  • max_grad_norm : 1
  • num_train_epochs : 1
  • max_steps : -1
  • lr_scheduler_type : linear
  • lr_scheduler_kwargs : {}
  • warmup_ratio : 0.0
  • warmup_steps : 0
  • log_level : passive
  • log_level_replica : warning
  • log_on_each_node : True
  • logging_nan_inf_filter : True
  • save_safetensors : True
  • save_on_each_node : False
  • save_only_model : False
  • restore_callback_states_from_checkpoint : False
  • no_cuda : False
  • use_cpu : False
  • use_mps_device : False
  • seed : 42
  • data_seed : None
  • jit_mode_eval : False
  • use_ipex : False
  • bf16 : False
  • fp16 : True
  • fp16_opt_level : O1
  • half_precision_backend : auto
  • bf16_full_eval : False
  • fp16_full_eval : False
  • tf32 : None
  • local_rank : 0
  • ddp_backend : None
  • tpu_num_cores : None
  • tpu_metrics_debug : False
  • debug : []
  • dataloader_drop_last : False
  • dataloader_num_workers : 0
  • dataloader_prefetch_factor : None
  • past_index : -1
  • disable_tqdm : False
  • remove_unused_columns : True
  • label_names : None
  • load_best_model_at_end : False
  • ignore_data_skip : False
  • fsdp : []
  • fsdp_min_num_params : 0
  • fsdp_config : {'min_num_params': 0, 'xla': False, 'xla_fsdp_v2': False, 'xla_fsdp_grad_ckpt': False}
  • tp_size : 0
  • fsdp_transformer_layer_cls_to_wrap : None
  • accelerator_config : {'split_batches': False, 'dispatch_batches': None, 'even_batches': True, 'use_seedable_sampler': True, 'non_blocking': False, 'gradient_accumulation_kwargs': None}
  • deepspeed : None
  • label_smoothing_factor : 0.0
  • optim : adamw_torch
  • optim_args : None
  • adafactor : False
  • group_by_length : False
  • length_column_name : length
  • ddp_find_unused_parameters : None
  • ddp_bucket_cap_mb : None
  • ddp_broadcast_buffers : False
  • dataloader_pin_memory : True
  • dataloader_persistent_workers : False
  • skip_memory_metrics : True
  • use_legacy_prediction_loop : False
  • push_to_hub : False
  • resume_from_checkpoint : None
  • hub_model_id : None
  • hub_strategy : every_save
  • hub_private_repo : None
  • hub_always_push : False
  • gradient_checkpointing : False
  • gradient_checkpointing_kwargs : None
  • include_inputs_for_metrics : False
  • include_for_metrics : []
  • eval_do_concat_batches : True
  • fp16_backend : auto
  • push_to_hub_model_id : None
  • push_to_hub_organization : None
  • mp_parameters :
  • auto_find_batch_size : False
  • full_determinism : False
  • torchdynamo : None
  • ray_scope : last
  • ddp_timeout : 1800
  • torch_compile : False
  • torch_compile_backend : None
  • torch_compile_mode : None
  • include_tokens_per_second : False
  • include_num_input_tokens_seen : False
  • neftune_noise_alpha : None
  • optim_target_modules : None
  • batch_eval_metrics : False
  • eval_on_start : False
  • use_liger_kernel : False
  • eval_use_gather_object : False
  • average_tokens_across_devices : False
  • prompts : None
  • batch_sampler : batch_sampler
  • multi_dataset_batch_sampler : round_robin
Training Logs
Epoch Step Training Loss
0.0069 500 0.6741
0.0137 1000 0.4247
0.0206 1500 0.3538
0.0275 2000 0.334
0.0344 2500 0.3155
0.0412 3000 0.2833
0.0481 3500 0.2689
0.0550 4000 0.2633
0.0618 4500 0.2577
0.0687 5000 0.2642
0.0756 5500 0.2484
0.0825 6000 0.237
0.0893 6500 0.2225
0.0962 7000 0.2359
0.1031 7500 0.2266
0.1099 8000 0.2222
0.1168 8500 0.2136
0.1237 9000 0.2236
0.1306 9500 0.2149
0.1374 10000 0.2199
0.1443 10500 0.206
0.1512 11000 0.216
0.1580 11500 0.2069
0.1649 12000 0.1903
0.1718 12500 0.1958
0.1786 13000 0.2076
0.1855 13500 0.2033
0.1924 14000 0.1893
0.1993 14500 0.2024
0.2061 15000 0.1873
0.2130 15500 0.1788
0.2199 16000 0.1959
0.2267 16500 0.1996
0.2336 17000 0.183
0.2405 17500 0.185
0.2474 18000 0.1752
0.2542 18500 0.1856
0.2611 19000 0.1948
0.2680 19500 0.1826
0.2748 20000 0.1672
0.2817 20500 0.1746
0.2886 21000 0.1801
0.2955 21500 0.1847
0.3023 22000 0.1673
0.3092 22500 0.1788
0.3161 23000 0.1667
0.3229 23500 0.1746
Framework Versions
  • Python: 3.12.10
  • Sentence Transformers: 4.1.0
  • Transformers: 4.51.3
  • PyTorch: 2.6.0+cu124
  • Accelerate: 1.8.1
  • Datasets: 3.6.0
  • Tokenizers: 0.21.1
Citation
BibTeX
Sentence Transformers
@inproceedings{reimers-2019-sentence-bert,
    title = "Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks",
    author = "Reimers, Nils and Gurevych, Iryna",
    booktitle = "Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing",
    month = "11",
    year = "2019",
    publisher = "Association for Computational Linguistics",
    url = "https://arxiv.org/abs/1908.10084",
}
MultipleNegativesRankingLoss
@misc{henderson2017efficient,
    title={Efficient Natural Language Response Suggestion for Smart Reply},
    author={Matthew Henderson and Rami Al-Rfou and Brian Strope and Yun-hsuan Sung and Laszlo Lukacs and Ruiqi Guo and Sanjiv Kumar and Balint Miklos and Ray Kurzweil},
    year={2017},
    eprint={1705.00652},
    archivePrefix={arXiv},
    primaryClass={cs.CL}
}

Runs of lingtrain labse-chuvash-3 on huggingface.co

12
Total runs
0
24-hour runs
1
3-day runs
-1
7-day runs
-8
30-day runs

More Information About labse-chuvash-3 huggingface.co Model

labse-chuvash-3 huggingface.co

labse-chuvash-3 huggingface.co is an AI model on huggingface.co that provides labse-chuvash-3's model effect (), which can be used instantly with this lingtrain labse-chuvash-3 model. huggingface.co supports a free trial of the labse-chuvash-3 model, and also provides paid use of the labse-chuvash-3. Support call labse-chuvash-3 model through api, including Node.js, Python, http.

labse-chuvash-3 huggingface.co Url

https://huggingface.co/lingtrain/labse-chuvash-3

lingtrain labse-chuvash-3 online free

labse-chuvash-3 huggingface.co is an online trial and call api platform, which integrates labse-chuvash-3's modeling effects, including api services, and provides a free online trial of labse-chuvash-3, you can try labse-chuvash-3 online for free by clicking the link below.

lingtrain labse-chuvash-3 online free url in huggingface.co:

https://huggingface.co/lingtrain/labse-chuvash-3

labse-chuvash-3 install

labse-chuvash-3 is an open source model from GitHub that offers a free installation service, and any user can find labse-chuvash-3 on GitHub to install. At the same time, huggingface.co provides the effect of labse-chuvash-3 install, users can directly use labse-chuvash-3 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

labse-chuvash-3 install url in huggingface.co:

https://huggingface.co/lingtrain/labse-chuvash-3

Url of labse-chuvash-3

labse-chuvash-3 huggingface.co Url

Provider of labse-chuvash-3 huggingface.co

lingtrain
ORGANIZATIONS

Other API from lingtrain

huggingface.co

Total runs: 10
Run Growth: 4
Growth Rate: 40.00%
Updated:February 04 2024