merve / peft-copy-test

huggingface.co
Total runs: 15
24-hour runs: 0
7-day runs: 9
30-day runs: 12
Model's Last Updated: June 14 2023
text-generation

Introduction of peft-copy-test

Model Details of peft-copy-test

pull_figure

Llama-se-rl-peft

Adapter weights of a Reinforcement Learning fine-tuned model based on the LLaMA model (see Meta's LLaMA release for the original LLaMA model). The model is designed to generate human-like responses to questions in Stack Exchange domains of programming, mathematics, physics, and more. For more info check out the blog post and github example .

Model Details
Model Description

Developed by: Hugging Face

Model type: An auto-regressive language model based on the transformer architecture, and fine-tuned with Stack Exchange datasets .

Languages: Predominantly English, with additional data from languages with the following ISO codes:

bg ca cs da de es fr hr hu it nl pl pt ro ru sl sr sv uk

License: bigscience-openrail-m

Finetuned from: LLaMA

Model Sources

Repository: https://huggingface.co/trl-lib/llama-7b-se-rl-peft/tree/main

Base Model Repository: https://github.com/facebookresearch/llama

Demo: https://huggingface.co/spaces/trl-lib/stack-llama

Uses
Direct Use
  • Long-form question-answering on topics of programming, mathematics, and physics
  • Demonstrating a Large Language Model's ability to follow target behavior of generating answers to a question that would be highly rated on Stack Exchange .
Out of Scope Use
  • Replacing human expertise
Bias, Risks, and Limitations
Recommendations
  • Answers should be validated through the use of external sources.
  • Disparities between the data contributors and the direct and indirect users of the technology should inform developers in assessing what constitutes an appropriate use case.
  • Further research is needed to attribute model generations to sources in the training data, especially in cases where the model copies answers from the training data.
Training Details
Training Data

Original datasets are described in the LLaMA Model Card . Fine-tuning datasets for this model are based on Stack Exchange Paired , which consists of questions and answers from various domains in Stack Exchange, such as programming, mathematics, physics, and more. Specifically:

Traditional Fine-tuning: https://huggingface.co/datasets/lvwerra/stack-exchange-paired/tree/main/data/finetune

RL Fine-tuning: https://huggingface.co/datasets/lvwerra/stack-exchange-paired/tree/main/data/rl

Reward Model: https://huggingface.co/trl-lib/llama-7b-se-rm-peft

Training Procedure

The model was first fine-tuned on the Stack Exchange question and answer pairs and then RL fine-tuned using a Stack Exchange Reward Model. It is trained to respond to prompts with the following template:

Question: <Query> 

Answer: <Response>
Citation

BibTeX:

@misc {beeching2023stackllama,
    author       = { Edward Beeching and
                     Younes Belkada and
                     Kashif Rasul and
                     Lewis Tunstall and
                     Leandro von Werra and
                     Nazneen Rajani and
                     Nathan Lambert
                   },
    title        = { StackLLaMa: An RL Fine-tuned LLaMa Model for Stack Exchange Question and Answering },
    year         = 2023,
    url          = { https://huggingface.co/trl-lib/llama-7b-se-rl-peft },
    doi          = { 10.57967/hf/0513 },
    publisher    = { Hugging Face Blog }
}
Model Card Authors

Nathan Lambert , Leandro von Werra , Edward Beeching , Kashif Rasul , Younes Belkada , Margaret Mitchell

Runs of merve peft-copy-test on huggingface.co

15
Total runs
0
24-hour runs
4
3-day runs
9
7-day runs
12
30-day runs

More Information About peft-copy-test huggingface.co Model

peft-copy-test huggingface.co

peft-copy-test huggingface.co is an AI model on huggingface.co that provides peft-copy-test's model effect (), which can be used instantly with this merve peft-copy-test model. huggingface.co supports a free trial of the peft-copy-test model, and also provides paid use of the peft-copy-test. Support call peft-copy-test model through api, including Node.js, Python, http.

peft-copy-test huggingface.co Url

https://huggingface.co/merve/peft-copy-test

merve peft-copy-test online free

peft-copy-test huggingface.co is an online trial and call api platform, which integrates peft-copy-test's modeling effects, including api services, and provides a free online trial of peft-copy-test, you can try peft-copy-test online for free by clicking the link below.

merve peft-copy-test online free url in huggingface.co:

https://huggingface.co/merve/peft-copy-test

peft-copy-test install

peft-copy-test is an open source model from GitHub that offers a free installation service, and any user can find peft-copy-test on GitHub to install. At the same time, huggingface.co provides the effect of peft-copy-test install, users can directly use peft-copy-test installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

peft-copy-test install url in huggingface.co:

https://huggingface.co/merve/peft-copy-test

Url of peft-copy-test

peft-copy-test huggingface.co Url

Provider of peft-copy-test huggingface.co

merve
ORGANIZATIONS

Other API from merve

huggingface.co

Total runs: 53
Run Growth: -8
Growth Rate: -15.09%
Updated:October 12 2023
huggingface.co

Total runs: 37
Run Growth: -2
Growth Rate: -5.13%
Updated:February 26 2024
huggingface.co

Total runs: 34
Run Growth: -24
Growth Rate: -70.59%
Updated:January 06 2024
huggingface.co

Total runs: 29
Run Growth: 7
Growth Rate: 24.14%
Updated:January 28 2026
huggingface.co

Total runs: 26
Run Growth: 14
Growth Rate: 53.85%
Updated:September 19 2025
huggingface.co

Total runs: 16
Run Growth: 9
Growth Rate: 56.25%
Updated:October 04 2023
huggingface.co

Total runs: 14
Run Growth: 7
Growth Rate: 50.00%
Updated:July 18 2024
huggingface.co

Total runs: 13
Run Growth: 7
Growth Rate: 43.75%
Updated:February 22 2024
huggingface.co

Total runs: 11
Run Growth: -2
Growth Rate: -18.18%
Updated:September 11 2025
huggingface.co

Total runs: 10
Run Growth: 9
Growth Rate: 90.00%
Updated:November 25 2023
huggingface.co

Total runs: 9
Run Growth: 6
Growth Rate: 66.67%
Updated:February 10 2023
huggingface.co

Total runs: 8
Run Growth: -22
Growth Rate: -275.00%
Updated:December 18 2024
huggingface.co

Total runs: 8
Run Growth: 2
Growth Rate: 25.00%
Updated:March 26 2024
huggingface.co

Total runs: 8
Run Growth: 4
Growth Rate: 50.00%
Updated:July 11 2023
huggingface.co

Total runs: 6
Run Growth: 1
Growth Rate: 16.67%
Updated:March 26 2024
huggingface.co

Total runs: 5
Run Growth: 3
Growth Rate: 60.00%
Updated:April 02 2023
huggingface.co

Total runs: 5
Run Growth: 3
Growth Rate: 60.00%
Updated:April 02 2023
huggingface.co

Total runs: 4
Run Growth: 0
Growth Rate: 0.00%
Updated:June 14 2023
huggingface.co

Total runs: 1
Run Growth: -7
Growth Rate: -233.33%
Updated:January 28 2023
huggingface.co

Total runs: 1
Run Growth: 1
Growth Rate: 100.00%
Updated:December 20 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:April 25 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:May 22 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:August 30 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:May 30 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 04 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:October 22 2024