swardiantara / bert-tiny-sst5-k1-adaptive-euclidean

huggingface.co
Total runs: 74
24-hour runs: -1
7-day runs: -6
30-day runs: -12
Model's Last Updated: June 13 2026
sentence-similarity

Introduction of bert-tiny-sst5-k1-adaptive-euclidean

Model Details of bert-tiny-sst5-k1-adaptive-euclidean

SentenceTransformer based on google/bert_uncased_L-2_H-128_A-2

This is a sentence-transformers model finetuned from google/bert_uncased_L-2_H-128_A-2 . It maps sentences & paragraphs to a 128-dimensional dense vector space and can be used for retrieval.

Model Details
Model Description
  • Model Type: Sentence Transformer
  • Base model: google/bert_uncased_L-2_H-128_A-2
  • Maximum Sequence Length: 128 tokens
  • Output Dimensionality: 128 dimensions
  • Similarity Function: Cosine Similarity
  • Supported Modality: Text
Model Sources
Full Model Architecture
SentenceTransformer(
  (0): Transformer({'transformer_task': 'feature-extraction', 'modality_config': {'text': {'method': 'forward', 'method_output_name': 'last_hidden_state'}}, 'module_output_name': 'token_embeddings', 'architecture': 'BertModel'})
  (1): Pooling({'embedding_dimension': 128, 'pooling_mode': 'mean', 'include_prompt': True})
)
Usage
Direct Usage (Sentence Transformers)

First install the Sentence Transformers library:

pip install -U sentence-transformers

Then you can load this model and run inference.

from sentence_transformers import SentenceTransformer

# Download from the 🤗 Hub
model = SentenceTransformer("swardiantara/bert-tiny-sst5-k1-adaptive-euclidean")
# Run inference
sentences = [
    'a metaphor for a modern-day urban china searching for its identity .',
    '-lrb- janey -rrb- forgets about her other obligations , leading to a tragedy which is somehow guessable from the first few minutes , maybe because it echoes the by now intolerable morbidity of so many recent movies .',
    "a semi-autobiographical film that 's so sloppily written and cast that you can not believe anyone more central to the creation of bugsy than the caterer had anything to do with it .",
]
embeddings = model.encode(sentences)
print(embeddings.shape)
# [3, 128]

# Get the similarity scores for the embeddings
similarities = model.similarity(embeddings, embeddings)
print(similarities)
# tensor([[1.0000, 0.8648, 0.8707],
#         [0.8648, 1.0000, 0.9650],
#         [0.8707, 0.9650, 1.0000]])
Training Details
Training Dataset
Unnamed Dataset
  • Size: 42,720 training samples
  • Columns: text_a , text_b , and label
  • Approximate statistics based on the first 100 samples:
    text_a text_b label
    type string string list
    modality text text
    details
    • min: 7 tokens
    • mean: 22.86 tokens
    • max: 48 tokens
    • min: 33 tokens
    • mean: 43.28 tokens
    • max: 53 tokens
    • size: 2 elements
  • Samples:
    text_a text_b label
    a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films director lee has a true cinematic knack , but it 's also nice to see a movie with its heart so thoroughly , unabashedly on its sleeve . [1.0, 0.0]
    a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films a semi-autobiographical film that 's so sloppily written and cast that you can not believe anyone more central to the creation of bugsy than the caterer had anything to do with it . [0.0, 1.0]
    a stirring , funny and finally transporting re-imagining of beauty and the beast and 1930s horror films -lrb- janey -rrb- forgets about her other obligations , leading to a tragedy which is somehow guessable from the first few minutes , maybe because it echoes the by now intolerable morbidity of so many recent movies . [0.0, 0.75]
  • Loss: main .OrdinalProxyContrastiveLoss
Training Hyperparameters
Non-Default Hyperparameters
  • per_device_train_batch_size : 1024
  • learning_rate : 1e-05
  • load_best_model_at_end : True
All Hyperparameters
Click to expand
  • per_device_train_batch_size : 1024
  • num_train_epochs : 3
  • max_steps : -1
  • learning_rate : 1e-05
  • lr_scheduler_type : linear
  • lr_scheduler_kwargs : None
  • warmup_steps : 0
  • optim : adamw_torch
  • optim_args : None
  • weight_decay : 0.0
  • adam_beta1 : 0.9
  • adam_beta2 : 0.999
  • adam_epsilon : 1e-08
  • optim_target_modules : None
  • gradient_accumulation_steps : 1
  • average_tokens_across_devices : True
  • max_grad_norm : 1.0
  • label_smoothing_factor : 0.0
  • bf16 : False
  • fp16 : False
  • bf16_full_eval : False
  • fp16_full_eval : False
  • tf32 : None
  • gradient_checkpointing : False
  • gradient_checkpointing_kwargs : None
  • torch_compile : False
  • torch_compile_backend : None
  • torch_compile_mode : None
  • use_liger_kernel : False
  • liger_kernel_config : None
  • use_cache : False
  • neftune_noise_alpha : None
  • torch_empty_cache_steps : None
  • auto_find_batch_size : False
  • log_on_each_node : True
  • logging_nan_inf_filter : True
  • include_num_input_tokens_seen : no
  • log_level : passive
  • log_level_replica : warning
  • disable_tqdm : False
  • project : huggingface
  • trackio_space_id : None
  • trackio_bucket_id : None
  • trackio_static_space_id : None
  • per_device_eval_batch_size : 8
  • prediction_loss_only : True
  • eval_on_start : False
  • eval_do_concat_batches : True
  • eval_use_gather_object : False
  • eval_accumulation_steps : None
  • include_for_metrics : []
  • batch_eval_metrics : False
  • save_only_model : False
  • save_on_each_node : False
  • enable_jit_checkpoint : False
  • push_to_hub : False
  • hub_private_repo : None
  • hub_model_id : None
  • hub_strategy : every_save
  • hub_always_push : False
  • hub_revision : None
  • load_best_model_at_end : True
  • ignore_data_skip : False
  • restore_callback_states_from_checkpoint : False
  • full_determinism : False
  • seed : 42
  • data_seed : None
  • use_cpu : False
  • accelerator_config : {'split_batches': False, 'dispatch_batches': None, 'even_batches': True, 'use_seedable_sampler': True, 'non_blocking': False, 'gradient_accumulation_kwargs': None}
  • parallelism_config : None
  • dataloader_drop_last : False
  • dataloader_num_workers : 0
  • dataloader_pin_memory : True
  • dataloader_persistent_workers : False
  • dataloader_prefetch_factor : None
  • remove_unused_columns : True
  • label_names : None
  • train_sampling_strategy : random
  • length_column_name : length
  • ddp_find_unused_parameters : None
  • ddp_bucket_cap_mb : None
  • ddp_broadcast_buffers : False
  • ddp_static_graph : None
  • ddp_backend : None
  • ddp_timeout : 1800
  • fsdp : None
  • fsdp_config : None
  • deepspeed : None
  • debug : []
  • skip_memory_metrics : True
  • do_predict : False
  • resume_from_checkpoint : None
  • warmup_ratio : None
  • local_rank : -1
  • prompts : None
  • batch_sampler : batch_sampler
  • multi_dataset_batch_sampler : proportional
  • router_mapping : {}
  • learning_rate_mapping : {}
Training Logs
Epoch Step
1.0 42
2.0 84
3.0 126
  • The bold row denotes the saved checkpoint.
Training Time
  • Training : 17.1 seconds
  • Evaluation : 0.3 seconds
  • Total : 17.4 seconds
Framework Versions
  • Python: 3.12.4
  • Sentence Transformers: 5.5.1
  • Transformers: 5.11.0
  • PyTorch: 2.5.1+cu121
  • Accelerate: 1.13.0
  • Datasets: 2.21.0
  • Tokenizers: 0.22.2
Citation
BibTeX
Sentence Transformers
@inproceedings{reimers-2019-sentence-bert,
    title = "Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks",
    author = "Reimers, Nils and Gurevych, Iryna",
    booktitle = "Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing",
    month = "11",
    year = "2019",
    publisher = "Association for Computational Linguistics",
    url = "https://arxiv.org/abs/1908.10084",
}

Runs of swardiantara bert-tiny-sst5-k1-adaptive-euclidean on huggingface.co

74
Total runs
-1
24-hour runs
-15
3-day runs
-6
7-day runs
-12
30-day runs

More Information About bert-tiny-sst5-k1-adaptive-euclidean huggingface.co Model

bert-tiny-sst5-k1-adaptive-euclidean huggingface.co

bert-tiny-sst5-k1-adaptive-euclidean huggingface.co is an AI model on huggingface.co that provides bert-tiny-sst5-k1-adaptive-euclidean's model effect (), which can be used instantly with this swardiantara bert-tiny-sst5-k1-adaptive-euclidean model. huggingface.co supports a free trial of the bert-tiny-sst5-k1-adaptive-euclidean model, and also provides paid use of the bert-tiny-sst5-k1-adaptive-euclidean. Support call bert-tiny-sst5-k1-adaptive-euclidean model through api, including Node.js, Python, http.

bert-tiny-sst5-k1-adaptive-euclidean huggingface.co Url

https://huggingface.co/swardiantara/bert-tiny-sst5-k1-adaptive-euclidean

swardiantara bert-tiny-sst5-k1-adaptive-euclidean online free

bert-tiny-sst5-k1-adaptive-euclidean huggingface.co is an online trial and call api platform, which integrates bert-tiny-sst5-k1-adaptive-euclidean's modeling effects, including api services, and provides a free online trial of bert-tiny-sst5-k1-adaptive-euclidean, you can try bert-tiny-sst5-k1-adaptive-euclidean online for free by clicking the link below.

swardiantara bert-tiny-sst5-k1-adaptive-euclidean online free url in huggingface.co:

https://huggingface.co/swardiantara/bert-tiny-sst5-k1-adaptive-euclidean

bert-tiny-sst5-k1-adaptive-euclidean install

bert-tiny-sst5-k1-adaptive-euclidean is an open source model from GitHub that offers a free installation service, and any user can find bert-tiny-sst5-k1-adaptive-euclidean on GitHub to install. At the same time, huggingface.co provides the effect of bert-tiny-sst5-k1-adaptive-euclidean install, users can directly use bert-tiny-sst5-k1-adaptive-euclidean installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

bert-tiny-sst5-k1-adaptive-euclidean install url in huggingface.co:

https://huggingface.co/swardiantara/bert-tiny-sst5-k1-adaptive-euclidean

Url of bert-tiny-sst5-k1-adaptive-euclidean

bert-tiny-sst5-k1-adaptive-euclidean huggingface.co Url

Provider of bert-tiny-sst5-k1-adaptive-euclidean huggingface.co

swardiantara
ORGANIZATIONS

Other API from swardiantara