CausalLM / 35b-beta-long

huggingface.co
Total runs: 112
24-hour runs: 2
7-day runs: 8
30-day runs: 79
Model's Last Updated: February 11 2025
text-generation

Introduction of 35b-beta-long

Model Details of 35b-beta-long

Sorry, it's no longer available on Hugging Face. Please reach out to those who have already downloaded it. If you have a copy, please refrain from re-uploading it to Hugging Face.

Due to repeated conflicts with HF and what we perceive as their repeated misuse of the "Contributor Covenant Code of Conduct," we have lost confidence in the platform and decided to temporarily suspend all new download access requests. It appears to us that HF's original intention has been abandoned in pursuit of commercialization, and they no longer prioritize the well-being of the community.

Demo:

35b-beta-long

This release, CausalLM/35b-beta-long, represents the culmination of our experience and accumulated training data in fine-tuning large language models. We are open-sourcing these weights to foster development within the open-source community.

We chose Cohere's multilingual, 35B-parameter with long context [CohereForAI/c4ai-command-r-v01] MHA model as our base. In our evaluation, it proved to be the most responsive to the quality of training data throughout the Supervised Fine-Tuning process, outperforming other open-source LLMs. Although its initial SFT/RL focuses on specific tasks and comes with a non-commercial license, we believe it's currently the best foundation for personal and internal use cases.

Utilizing extensive factual content from web crawls, we synthesized over 30 million multi-turn dialogue data entries, grounded in multiple web-pages or documents. This process involved substantial human oversight and a data pipeline designed to ensure high quality. The model was then trained on this data in full 128K context using BF16 precision. We also incorporated widely-used open-source dialogue datasets to enhance general conversational fluency.

Our data synthesis approach addressed crucial limitations in typical LLM training corpora. LLMs often struggle to extract thematic summaries, key information, or perform comparisons at the paragraph or document level. Therefore, we focused on generating fact-based data using multiple documents within a long context setting. This involved leveraging existing SOTA LLMs with human guidance to synthesize information through thematic summarization, information extraction, and comparison of source materials.

This approach yielded significant improvements in model performance during fine-tuning. We observed reductions in hallucinations, enhanced long-context capabilities, and improvements in general abilities such as math, coding, and knowledge recall. The training process incorporated both the original source material and the synthesized outputs, further reinforcing the model's ability to recall and utilize abstract concepts embedded within the pre-training data. Our analysis revealed that this combination of original and synthesized data was crucial for achieving a more balanced performance profile. Intermediate checkpoints and models trained solely on synthesized data are also released for research purposes.

Compared to the original task-specific model, our further fine-tuned model demonstrates more robust recall in long-context scenarios without requiring specific document formatting or prompt engineering. This fine-tuned model also exhibits performance comparable to models twice its size in quantifiable benchmarks.

As this model has only undergone SFT, it may still exhibit biases or generate undesirable content. We implemented basic safety measures using open-source refusal datasets to mitigate outputs related to illegal activities, NSFW content, and violence. However, further Reinforcement Learning is necessary for robust alignment with human values.

Please note

Tokenizer is different from cohere - and chat template is ChatML .

Pressure Testing from: https://github.com/LeonEricsson/llmcontext

image/png

Runs of CausalLM 35b-beta-long on huggingface.co

112
Total runs
2
24-hour runs
6
3-day runs
8
7-day runs
79
30-day runs

More Information About 35b-beta-long huggingface.co Model

More 35b-beta-long license Visit here:

https://choosealicense.com/licenses/wtfpl

35b-beta-long huggingface.co

35b-beta-long huggingface.co is an AI model on huggingface.co that provides 35b-beta-long's model effect (), which can be used instantly with this CausalLM 35b-beta-long model. huggingface.co supports a free trial of the 35b-beta-long model, and also provides paid use of the 35b-beta-long. Support call 35b-beta-long model through api, including Node.js, Python, http.

35b-beta-long huggingface.co Url

https://huggingface.co/CausalLM/35b-beta-long

CausalLM 35b-beta-long online free

35b-beta-long huggingface.co is an online trial and call api platform, which integrates 35b-beta-long's modeling effects, including api services, and provides a free online trial of 35b-beta-long, you can try 35b-beta-long online for free by clicking the link below.

CausalLM 35b-beta-long online free url in huggingface.co:

https://huggingface.co/CausalLM/35b-beta-long

35b-beta-long install

35b-beta-long is an open source model from GitHub that offers a free installation service, and any user can find 35b-beta-long on GitHub to install. At the same time, huggingface.co provides the effect of 35b-beta-long install, users can directly use 35b-beta-long installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

35b-beta-long install url in huggingface.co:

https://huggingface.co/CausalLM/35b-beta-long

Url of 35b-beta-long

35b-beta-long huggingface.co Url

Provider of 35b-beta-long huggingface.co

CausalLM
ORGANIZATIONS

Other API from CausalLM

huggingface.co

Total runs: 8.6K
Run Growth: 125
Growth Rate: 1.45%
Updated:May 25 2024
huggingface.co

Total runs: 704
Run Growth: 459
Growth Rate: 66.33%
Updated:February 11 2025
huggingface.co

Total runs: 292
Run Growth: 72
Growth Rate: 24.66%
Updated:December 10 2023
huggingface.co

Total runs: 261
Run Growth: -2.4K
Growth Rate: -931.50%
Updated:February 11 2025
huggingface.co

Total runs: 228
Run Growth: 90
Growth Rate: 39.47%
Updated:February 11 2025
huggingface.co

Total runs: 124
Run Growth: 78
Growth Rate: 62.90%
Updated:February 15 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 28 2024