The
i3-80M Model
is a novel hybrid architecture combining convolutional/recurrent layers with full attention layers for efficient language modeling. This architecture uniquely blends RWKV-style time-mixing with Mamba state-space dynamics in the early layers, followed by standard multi-head attention in deeper layers.
This is the second model in the i3 series, scaling up from the original
i3-22M
with improved architecture and multi-dataset training.
Model Statistics
Total Parameters
: ~82.77M (82,765,160)
Architecture
: 10 Hybrid (RWKV-Mamba) + 6 Full Attention Layers = 16 Total Layers
Vocabulary Size
: 35,560 tokens (variable-length chunks with token)
i3-80m huggingface.co is an AI model on huggingface.co that provides i3-80m's model effect (), which can be used instantly with this FlameF0X i3-80m model. huggingface.co supports a free trial of the i3-80m model, and also provides paid use of the i3-80m. Support call i3-80m model through api, including Node.js, Python, http.
i3-80m huggingface.co is an online trial and call api platform, which integrates i3-80m's modeling effects, including api services, and provides a free online trial of i3-80m, you can try i3-80m online for free by clicking the link below.
FlameF0X i3-80m online free url in huggingface.co:
i3-80m is an open source model from GitHub that offers a free installation service, and any user can find i3-80m on GitHub to install. At the same time, huggingface.co provides the effect of i3-80m install, users can directly use i3-80m installed effect in huggingface.co for debugging and trial. It also supports api for free installation.