BharatGen
introduces the early checkpoint of SFT (Supervised Fine-Tuned) for
Param 1
, a bilingual language model trained from scratch in English and Hindi. With 2.9 billion parameters, this checkpoint builds upon the pretraining phase and serves as a foundation for more downstream tasks, safety testing, and customization.
📁 Folder Structure
Param1/
├── model.nemo # .nemo packaged model file
├── nemo_inference.sh # Shell script for running inference
└── README.md # model documentation file
Pre-Training Details
Dataset
: 7.5 Trillion tokens
Data Quality
: Highly curated with standard filtering and multiple processing steps.
Important Guidelines for Early Checkpoint Release of Param-1-2.9B-Instruct
Early Development Status
This model is in the initial phase of Param-1 Instruct Model.
It is yet to undergo full-supervised fine-tuning, safety alignment, or rigorous evaluation.
The release is intended to showcase progress, gather feedback, and encourage research and experimentation.
Outputs may at times be incoherent, irrelevant, or of suboptimal quality.
Data Sources and Potential Artifacts
To preserve the Model's understanding on the global front, part of the training Data also includes data crawled from the Internet hence it may contain inherited artifacts;
Due to the increased prevalence of AI-generated content online in the current times, the model may occasionally mimic such statements and incorrectly identify itself.
These artifacts are natural consequences of using publicly available data found on the internet although critical but important since such data is important for the model to build a global know-how and we will be addressing issues like this in future iterations of the current Model.
Lack of Alignment and Guardrails
A preliminary-level alignment or safety mechanisms have been implemented at this stage.
The model is yet to under go full-scale instruction tuning, supervised fine-tuning, or reinforcement learning from human feedback (RLHF).
As a result, it may occasionally:
Generate biased, offensive, or unsafe content
Be susceptible to misuse or prompt injection (jailbreaking)
Respond to harmful or unethical prompts without refusal
This model must not be deployed in any production without reading Intent use section.
Intended Use
This release is provided exclusively for research, experimentation and contribution to the open source community.
Suggested use cases include:
Assessing early-stage LLM behavior
Debugging model training pipelines and configurations
Benchmarking or custom fine-tuning by the community
Access to early-checkpoint should embibe a sense of motivation and enthusiasm among the open source community to take such early-stage check point and build India-Specific Innovative use cases on top of it. This should also help foster innovation among the Community.
Licensing and Responsibility
Released under an open license with responsible usage guidelines.
License: MIT
Users are expected to:
Adhere to ethical usage practices and legal regulations
Avoid malicious or unsafe deployment
Credit the authors as per the licensing terms
Acknowledgement of Origin
A home-grown effort initiated in India with limited resources.
This work represents a bottom-up initiative to develop LLMs from scratch within India.
It reflects our humble, resource-constrained journey to contribute meaningfully to the open-source AI ecosystem.
We hope to foster collaboration and growth within the broader community.
Transparency & Community Collaboration
We welcome contributions and open dialogue.
We encourage the community to share feedback, report issues, and collaborate.
Future versions will introduce better alignment, improved training scale, and more curated datasets.
Together, we aim to evolve toward safer and more capable AI systems.
📜 License
This SFT checkpoint is released under the
BharatGen non-commercial license
.
Please refer to the
LICENSE
for terms and conditions.
Runs of bharatgenai Param-1-2.9B-Instruct on huggingface.co
2.3K
Total runs
-104
24-hour runs
-67
3-day runs
135
7-day runs
1.5K
30-day runs
More Information About Param-1-2.9B-Instruct huggingface.co Model
Param-1-2.9B-Instruct huggingface.co is an AI model on huggingface.co that provides Param-1-2.9B-Instruct's model effect (), which can be used instantly with this bharatgenai Param-1-2.9B-Instruct model. huggingface.co supports a free trial of the Param-1-2.9B-Instruct model, and also provides paid use of the Param-1-2.9B-Instruct. Support call Param-1-2.9B-Instruct model through api, including Node.js, Python, http.
Param-1-2.9B-Instruct huggingface.co is an online trial and call api platform, which integrates Param-1-2.9B-Instruct's modeling effects, including api services, and provides a free online trial of Param-1-2.9B-Instruct, you can try Param-1-2.9B-Instruct online for free by clicking the link below.
bharatgenai Param-1-2.9B-Instruct online free url in huggingface.co:
Param-1-2.9B-Instruct is an open source model from GitHub that offers a free installation service, and any user can find Param-1-2.9B-Instruct on GitHub to install. At the same time, huggingface.co provides the effect of Param-1-2.9B-Instruct install, users can directly use Param-1-2.9B-Instruct installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
Param-1-2.9B-Instruct install url in huggingface.co: