swardiantara / bert-tiny-sst5-k3-adaptive-euclidean

huggingface.co
Total runs: 81
24-hour runs: -1
7-day runs: -2
30-day runs: 8
Model's Last Updated: June 12 2026
sentence-similarity

Introduction of bert-tiny-sst5-k3-adaptive-euclidean

Model Details of bert-tiny-sst5-k3-adaptive-euclidean

SentenceTransformer based on google/bert_uncased_L-2_H-128_A-2

This is a sentence-transformers model finetuned from google/bert_uncased_L-2_H-128_A-2 . It maps sentences & paragraphs to a 128-dimensional dense vector space and can be used for retrieval.

Model Details
Model Description
  • Model Type: Sentence Transformer
  • Base model: google/bert_uncased_L-2_H-128_A-2
  • Maximum Sequence Length: 128 tokens
  • Output Dimensionality: 128 dimensions
  • Similarity Function: Cosine Similarity
  • Supported Modality: Text
Model Sources
Full Model Architecture
SentenceTransformer(
  (0): Transformer({'transformer_task': 'feature-extraction', 'modality_config': {'text': {'method': 'forward', 'method_output_name': 'last_hidden_state'}}, 'module_output_name': 'token_embeddings', 'architecture': 'BertModel'})
  (1): Pooling({'embedding_dimension': 128, 'pooling_mode': 'mean', 'include_prompt': True})
)
Usage
Direct Usage (Sentence Transformers)

First install the Sentence Transformers library:

pip install -U sentence-transformers

Then you can load this model and run inference.

from sentence_transformers import SentenceTransformer

# Download from the 🤗 Hub
model = SentenceTransformer("swardiantara/bert-tiny-sst5-k3-adaptive-euclidean")
# Run inference
sentences = [
    'a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films',
    "attal 's hang-ups surrounding infidelity are so old-fashioned and , dare i say , outdated , it 's a wonder that he could n't have brought something fresher to the proceedings simply by accident .",
    '... a ho-hum affair , always watchable yet hardly memorable .',
]
embeddings = model.encode(sentences)
print(embeddings.shape)
# [3, 128]

# Get the similarity scores for the embeddings
similarities = model.similarity(embeddings, embeddings)
print(similarities)
# tensor([[1.0000, 0.9332, 0.9227],
#         [0.9332, 1.0000, 0.9513],
#         [0.9227, 0.9513, 1.0000]])
Training Details
Training Dataset
Unnamed Dataset
  • Size: 111,102 training samples
  • Columns: text_a , text_b , and label
  • Approximate statistics based on the first 100 samples:
    text_a text_b label
    type string string list
    modality text text
    details
    • min: 7 tokens
    • mean: 23.0 tokens
    • max: 48 tokens
    • min: 15 tokens
    • mean: 35.4 tokens
    • max: 56 tokens
    • size: 2 elements
  • Samples:
    text_a text_b label
    a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films tully is in many ways the perfect festival film : a calm , self-assured portrait of small town regret , love , duty and friendship that appeals to the storytelling instincts of a slightly more literate filmgoing audience . [1.0, 0.0]
    a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films ... a complete shambles of a movie so sloppy , so uneven , so damn unpleasant that i ca n't believe any viewer , young or old , would have a good time here . [0.0, 1.0]
    a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films a semi-autobiographical film that 's so sloppily written and cast that you can not believe anyone more central to the creation of bugsy than the caterer had anything to do with it . [0.0, 1.0]
  • Loss: main .OrdinalProxyContrastiveLoss
Training Hyperparameters
Non-Default Hyperparameters
  • per_device_train_batch_size : 1024
  • learning_rate : 1e-05
  • load_best_model_at_end : True
All Hyperparameters
Click to expand
  • per_device_train_batch_size : 1024
  • num_train_epochs : 3
  • max_steps : -1
  • learning_rate : 1e-05
  • lr_scheduler_type : linear
  • lr_scheduler_kwargs : None
  • warmup_steps : 0
  • optim : adamw_torch
  • optim_args : None
  • weight_decay : 0.0
  • adam_beta1 : 0.9
  • adam_beta2 : 0.999
  • adam_epsilon : 1e-08
  • optim_target_modules : None
  • gradient_accumulation_steps : 1
  • average_tokens_across_devices : True
  • max_grad_norm : 1.0
  • label_smoothing_factor : 0.0
  • bf16 : False
  • fp16 : False
  • bf16_full_eval : False
  • fp16_full_eval : False
  • tf32 : None
  • gradient_checkpointing : False
  • gradient_checkpointing_kwargs : None
  • torch_compile : False
  • torch_compile_backend : None
  • torch_compile_mode : None
  • use_liger_kernel : False
  • liger_kernel_config : None
  • use_cache : False
  • neftune_noise_alpha : None
  • torch_empty_cache_steps : None
  • auto_find_batch_size : False
  • log_on_each_node : True
  • logging_nan_inf_filter : True
  • include_num_input_tokens_seen : no
  • log_level : passive
  • log_level_replica : warning
  • disable_tqdm : False
  • project : huggingface
  • trackio_space_id : None
  • trackio_bucket_id : None
  • trackio_static_space_id : None
  • per_device_eval_batch_size : 8
  • prediction_loss_only : True
  • eval_on_start : False
  • eval_do_concat_batches : True
  • eval_use_gather_object : False
  • eval_accumulation_steps : None
  • include_for_metrics : []
  • batch_eval_metrics : False
  • save_only_model : False
  • save_on_each_node : False
  • enable_jit_checkpoint : False
  • push_to_hub : False
  • hub_private_repo : None
  • hub_model_id : None
  • hub_strategy : every_save
  • hub_always_push : False
  • hub_revision : None
  • load_best_model_at_end : True
  • ignore_data_skip : False
  • restore_callback_states_from_checkpoint : False
  • full_determinism : False
  • seed : 42
  • data_seed : None
  • use_cpu : False
  • accelerator_config : {'split_batches': False, 'dispatch_batches': None, 'even_batches': True, 'use_seedable_sampler': True, 'non_blocking': False, 'gradient_accumulation_kwargs': None}
  • parallelism_config : None
  • dataloader_drop_last : False
  • dataloader_num_workers : 0
  • dataloader_pin_memory : True
  • dataloader_persistent_workers : False
  • dataloader_prefetch_factor : None
  • remove_unused_columns : True
  • label_names : None
  • train_sampling_strategy : random
  • length_column_name : length
  • ddp_find_unused_parameters : None
  • ddp_bucket_cap_mb : None
  • ddp_broadcast_buffers : False
  • ddp_static_graph : None
  • ddp_backend : None
  • ddp_timeout : 1800
  • fsdp : None
  • fsdp_config : None
  • deepspeed : None
  • debug : []
  • skip_memory_metrics : True
  • do_predict : False
  • resume_from_checkpoint : None
  • warmup_ratio : None
  • local_rank : -1
  • prompts : None
  • batch_sampler : batch_sampler
  • multi_dataset_batch_sampler : proportional
  • router_mapping : {}
  • learning_rate_mapping : {}
Training Logs
Epoch Step
1.0 109
2.0 218
3.0 327
  • The bold row denotes the saved checkpoint.
Training Time
  • Training : 42.7 seconds
  • Evaluation : 0.2 seconds
  • Total : 43.0 seconds
Framework Versions
  • Python: 3.12.4
  • Sentence Transformers: 5.5.1
  • Transformers: 5.11.0
  • PyTorch: 2.5.1+cu121
  • Accelerate: 1.13.0
  • Datasets: 2.21.0
  • Tokenizers: 0.22.2
Citation
BibTeX
Sentence Transformers
@inproceedings{reimers-2019-sentence-bert,
    title = "Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks",
    author = "Reimers, Nils and Gurevych, Iryna",
    booktitle = "Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing",
    month = "11",
    year = "2019",
    publisher = "Association for Computational Linguistics",
    url = "https://arxiv.org/abs/1908.10084",
}

Runs of swardiantara bert-tiny-sst5-k3-adaptive-euclidean on huggingface.co

81
Total runs
-1
24-hour runs
0
3-day runs
-2
7-day runs
8
30-day runs

More Information About bert-tiny-sst5-k3-adaptive-euclidean huggingface.co Model

bert-tiny-sst5-k3-adaptive-euclidean huggingface.co

bert-tiny-sst5-k3-adaptive-euclidean huggingface.co is an AI model on huggingface.co that provides bert-tiny-sst5-k3-adaptive-euclidean's model effect (), which can be used instantly with this swardiantara bert-tiny-sst5-k3-adaptive-euclidean model. huggingface.co supports a free trial of the bert-tiny-sst5-k3-adaptive-euclidean model, and also provides paid use of the bert-tiny-sst5-k3-adaptive-euclidean. Support call bert-tiny-sst5-k3-adaptive-euclidean model through api, including Node.js, Python, http.

bert-tiny-sst5-k3-adaptive-euclidean huggingface.co Url

https://huggingface.co/swardiantara/bert-tiny-sst5-k3-adaptive-euclidean

swardiantara bert-tiny-sst5-k3-adaptive-euclidean online free

bert-tiny-sst5-k3-adaptive-euclidean huggingface.co is an online trial and call api platform, which integrates bert-tiny-sst5-k3-adaptive-euclidean's modeling effects, including api services, and provides a free online trial of bert-tiny-sst5-k3-adaptive-euclidean, you can try bert-tiny-sst5-k3-adaptive-euclidean online for free by clicking the link below.

swardiantara bert-tiny-sst5-k3-adaptive-euclidean online free url in huggingface.co:

https://huggingface.co/swardiantara/bert-tiny-sst5-k3-adaptive-euclidean

bert-tiny-sst5-k3-adaptive-euclidean install

bert-tiny-sst5-k3-adaptive-euclidean is an open source model from GitHub that offers a free installation service, and any user can find bert-tiny-sst5-k3-adaptive-euclidean on GitHub to install. At the same time, huggingface.co provides the effect of bert-tiny-sst5-k3-adaptive-euclidean install, users can directly use bert-tiny-sst5-k3-adaptive-euclidean installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

bert-tiny-sst5-k3-adaptive-euclidean install url in huggingface.co:

https://huggingface.co/swardiantara/bert-tiny-sst5-k3-adaptive-euclidean

Url of bert-tiny-sst5-k3-adaptive-euclidean

bert-tiny-sst5-k3-adaptive-euclidean huggingface.co Url

Provider of bert-tiny-sst5-k3-adaptive-euclidean huggingface.co

swardiantara
ORGANIZATIONS

Other API from swardiantara