Bochkov / emergent-semantics-model-16-float-269m

huggingface.co
Total runs: 7
24-hour runs: 0
7-day runs: 2
30-day runs: 3
Model's Last Updated: January 08 2026
text-generation

Introduction of emergent-semantics-model-16-float-269m

Model Details of emergent-semantics-model-16-float-269m

Emergent Semantics — Model_16_FLOAT (269M)

This repository provides Model_16_FLOAT (269M) — an ablation model from the paper:

📚 Paper (Emergent Semantics Beyond Token Embeddings: Transformer LMs with Frozen Visual Unicode Representations) -

📚 Paper (Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate) -

This checkpoint is designed to study the effect of normalization / PCA-style processing in a minimal frozen embedding setting.

Unlike Model_UNI_GLYPH , this model does not use glyph-based embeddings. Instead, it uses a frozen 16-dimensional float embedding per token.


Key idea (what this ablation tests)

This model isolates the impact of having float frozen embeddings (with PCA + normalization ) versus the strictly binary token-ID variant ( Model_16_BIT ):

  • n_embed = 16 per token ( float components , not binary)
  • Embedding vectors are precomputed (PCA + L2 normalization) and then frozen
  • The embedding layer is never updated ( requires_grad=False )
  • To match the Transformer hidden size, the 16-dim embedding is expanded to 1024 via a non-trainable repetition : repeat_interleave(64) → 16 * 64 = 1024

This lets you test whether the model’s behavior changes when the frozen token “identifier” is:

  • discrete + purely ID-like ( 16-bit ), vs
  • continuous + normalized ( 16-float )

Important: parameter count difference (vs 335M models)

This checkpoint has ~269M parameters , while models with a standard n_embed=1024 embedding table (e.g. UNI_GLYPH / unfrozen baselines ) are ~335M .

This difference is expected and comes primarily from the embedding matrix size:

  • Standard embedding params: vocab_size * 1024 = 65536 * 1024 ≈ 67.1M
  • This model’s embedding params: vocab_size * 16 = 65536 * 16 ≈ 1.0M

So the Transformer backbone is the same (layers/heads/d_model), but the embedding table is much smaller, reducing total parameters.


Model summary
  • Architecture: decoder-only Transformer (GPT-like)
  • Hidden size ( d_model ): 1024
  • Layers: 16
  • Heads: 32
  • Positional encoding: rotary embeddings
  • Activation: GELU
  • Tokenizer / vocab size: 65,536 (bvv241-2-3 compatible)
  • Input embeddings: frozen , n_embed=16 ( float , PCA + L2 normalized), expanded to 1024 by repetition (non-trainable)
  • Output head: not tied to the input embeddings (trained separately)

Tokenizer

The intended tokenizer is bvv241-2-3 (same vocab size and indexing):

You may load the tokenizer either from this model repo (if included) or from the standalone tokenizer repo. The key requirement is exact vocab alignment .


How to use (Transformers)

import torch
from transformers import AutoTokenizer, AutoModelForCausalLM

tokenizer = AutoTokenizer.from_pretrained("Bochkov/emergent-semantics-model-16-float-269m")
model = AutoModelForCausalLM.from_pretrained("Bochkov/emergent-semantics-model-16-float-269m", trust_remote_code=True).to('cuda')

inputs = torch.tensor([tokenizer.encode("Question: What is the capital of Japan?\nAnswer:")], dtype=torch.long, device='cuda')

outputs = model.generate(
    inputs, 
    max_new_tokens=10,
    do_sample=False
)
print(tokenizer.decode(outputs[0].tolist()))

#Question: What is the capital of Japan?
#Answer:A temperature in

Intended use

Research only, especially for:

  • Comparing Model_16_FLOAT vs Model_16_BIT (effect of continuous normalized vectors vs binary ID)
  • Comparing Model_16_FLOAT vs Model_UNI_GLYPH (effect of glyph-derived structure vs minimal vectors)
  • Studying emergent semantics when embeddings are frozen and non-semantic

Not intended for production deployment.


Related links

🧑‍🔬 Citation & Concept

If you use this model or the underlying concepts in your research, please cite our work:

@article{
      bochkov2025emergent,
      title={Emergent Semantics Beyond Token Embeddings: Transformer {LM}s with Frozen Visual Unicode Representations},
      author={Andrey Bochkov},
      journal={Transactions on Machine Learning Research},
      issn={2835-8856},
      year={2025},
      url={https://openreview.net/forum?id=Odh8IynO1o},
      note={}
}
@misc{bochkov2025growingtransformersmodularcomposition,
      title={Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate}, 
      author={A. Bochkov},
      year={2025},
      eprint={2507.07129},
      archivePrefix={arXiv},
      primaryClass={cs.LG},
      url={https://arxiv.org/abs/2507.07129}, 
}

Runs of Bochkov emergent-semantics-model-16-float-269m on huggingface.co

7
Total runs
0
24-hour runs
1
3-day runs
2
7-day runs
3
30-day runs

More Information About emergent-semantics-model-16-float-269m huggingface.co Model

More emergent-semantics-model-16-float-269m license Visit here:

https://choosealicense.com/licenses/apache-2.0

emergent-semantics-model-16-float-269m huggingface.co

emergent-semantics-model-16-float-269m huggingface.co is an AI model on huggingface.co that provides emergent-semantics-model-16-float-269m's model effect (), which can be used instantly with this Bochkov emergent-semantics-model-16-float-269m model. huggingface.co supports a free trial of the emergent-semantics-model-16-float-269m model, and also provides paid use of the emergent-semantics-model-16-float-269m. Support call emergent-semantics-model-16-float-269m model through api, including Node.js, Python, http.

emergent-semantics-model-16-float-269m huggingface.co Url

https://huggingface.co/Bochkov/emergent-semantics-model-16-float-269m

Bochkov emergent-semantics-model-16-float-269m online free

emergent-semantics-model-16-float-269m huggingface.co is an online trial and call api platform, which integrates emergent-semantics-model-16-float-269m's modeling effects, including api services, and provides a free online trial of emergent-semantics-model-16-float-269m, you can try emergent-semantics-model-16-float-269m online for free by clicking the link below.

Bochkov emergent-semantics-model-16-float-269m online free url in huggingface.co:

https://huggingface.co/Bochkov/emergent-semantics-model-16-float-269m

emergent-semantics-model-16-float-269m install

emergent-semantics-model-16-float-269m is an open source model from GitHub that offers a free installation service, and any user can find emergent-semantics-model-16-float-269m on GitHub to install. At the same time, huggingface.co provides the effect of emergent-semantics-model-16-float-269m install, users can directly use emergent-semantics-model-16-float-269m installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

emergent-semantics-model-16-float-269m install url in huggingface.co:

https://huggingface.co/Bochkov/emergent-semantics-model-16-float-269m

Url of emergent-semantics-model-16-float-269m

emergent-semantics-model-16-float-269m huggingface.co Url

Provider of emergent-semantics-model-16-float-269m huggingface.co

Bochkov
ORGANIZATIONS

Other API from Bochkov

huggingface.co

Total runs: 19
Run Growth: 11
Growth Rate: 57.89%
Updated:January 07 2026
huggingface.co

Total runs: 17
Run Growth: 1
Growth Rate: 5.88%
Updated:July 15 2025
huggingface.co

Total runs: 17
Run Growth: 0
Growth Rate: 0.00%
Updated:July 15 2025
huggingface.co

Total runs: 17
Run Growth: 0
Growth Rate: 0.00%
Updated:July 15 2025
huggingface.co

Total runs: 16
Run Growth: -1
Growth Rate: -6.25%
Updated:July 15 2025
huggingface.co

Total runs: 16
Run Growth: 0
Growth Rate: 0.00%
Updated:July 15 2025
huggingface.co

Total runs: 16
Run Growth: 0
Growth Rate: 0.00%
Updated:July 15 2025
huggingface.co

Total runs: 16
Run Growth: 0
Growth Rate: 0.00%
Updated:July 15 2025
huggingface.co

Total runs: 15
Run Growth: 8
Growth Rate: 53.33%
Updated:October 21 2025
huggingface.co

Total runs: 15
Run Growth: -3
Growth Rate: -20.00%
Updated:July 15 2025
huggingface.co

Total runs: 15
Run Growth: 6
Growth Rate: 40.00%
Updated:October 21 2025
huggingface.co

Total runs: 13
Run Growth: 3
Growth Rate: 23.08%
Updated:October 21 2025
huggingface.co

Total runs: 12
Run Growth: -3
Growth Rate: -25.00%
Updated:January 01 2026
huggingface.co

Total runs: 11
Run Growth: 7
Growth Rate: 63.64%
Updated:October 21 2025
huggingface.co

Total runs: 10
Run Growth: 2
Growth Rate: 20.00%
Updated:October 21 2025
huggingface.co

Total runs: 10
Run Growth: 2
Growth Rate: 20.00%
Updated:October 21 2025
huggingface.co

Total runs: 8
Run Growth: 7
Growth Rate: 87.50%
Updated:January 01 2026
huggingface.co

Total runs: 7
Run Growth: 4
Growth Rate: 57.14%
Updated:October 21 2025
huggingface.co

Total runs: 7
Run Growth: 4
Growth Rate: 57.14%
Updated:October 21 2025
huggingface.co

Total runs: 7
Run Growth: 3
Growth Rate: 42.86%
Updated:October 21 2025
huggingface.co

Total runs: 7
Run Growth: 3
Growth Rate: 42.86%
Updated:October 21 2025
huggingface.co

Total runs: 5
Run Growth: -1
Growth Rate: -20.00%
Updated:October 21 2025
huggingface.co

Total runs: 5
Run Growth: -1
Growth Rate: -20.00%
Updated:October 21 2025
huggingface.co

Total runs: 4
Run Growth: -1
Growth Rate: -25.00%
Updated:January 01 2026