Sao mentioned this model had some fixes, so make sure you read their notes at the bottom of the model card please. I ran some tests through it to be safe, it seemed to perform well at a 16k context limit, but can't confirm anything beyond that. Unlike Stardust it has some of the evil characters giving you second chances, so you might need to prompt that out.
You probably need to be using a Min P of - at most .1 - with Nemo models. This is of course assuming you aren't using any other sliders to modify token selection probability (e.g. Top K). Going farther than that and it seems like you start to get broken words.
Anyway, another Lyra model. Hype indeed. Sao delivers once again.
[See Previous Models]
|
Lyra-v4a1
|
------------> Lyra-v4 [Seperate RL Step targeting Instruct and Coherency over Base Nemo instead of SFT First, Result is Merged with Lyra-v4a1, fixes most quant-based issues. Somehow.]
This uses ChatML, or any of its variants which were included in previous versions.
<|im_start|>system
This is the system prompt.<|im_end|>
<|im_start|>user
Instructions placed here.<|im_end|>
<|im_start|>assistant
The model's response will be here.<|im_end|>
--------------------------------------------------
[INST]system
This is another system prompt.[/INST]
[INST]user
Your instructions placed here.[/INST]
[INST]assistant
The model's response will be here.[/INST]
Recommended Samplers:
Temperature: 0.6 - 1 # Make sure min_p is set before Temperature in Sampler Orders
min_p: 0.1 - 0.2 # Crucial for NeMo
Recommended Stopping Strings:
<|im_end|>
</s>
[/INST]
Notes
- I think I fixed the extra token stuff some users seem to be facing, while retaining everything else? It's some error alright.
- If you're using XML tags, you may see weird malformed stopping strings. Just add them to your current list. and move on.
- Its pretty nice, imo. I've been messing around with it a lot.
- Make sure the ChatML template is correct, I think there's some issues with the one used in SillyTavern which might cause improper replies?
Runs of Statuo Lyra-V4-EXL2-8bpw on huggingface.co
1
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
1
30-day runs
More Information About Lyra-V4-EXL2-8bpw huggingface.co Model
Lyra-V4-EXL2-8bpw huggingface.co is an AI model on huggingface.co that provides Lyra-V4-EXL2-8bpw's model effect (), which can be used instantly with this Statuo Lyra-V4-EXL2-8bpw model. huggingface.co supports a free trial of the Lyra-V4-EXL2-8bpw model, and also provides paid use of the Lyra-V4-EXL2-8bpw. Support call Lyra-V4-EXL2-8bpw model through api, including Node.js, Python, http.
Lyra-V4-EXL2-8bpw huggingface.co is an online trial and call api platform, which integrates Lyra-V4-EXL2-8bpw's modeling effects, including api services, and provides a free online trial of Lyra-V4-EXL2-8bpw, you can try Lyra-V4-EXL2-8bpw online for free by clicking the link below.
Statuo Lyra-V4-EXL2-8bpw online free url in huggingface.co:
Lyra-V4-EXL2-8bpw is an open source model from GitHub that offers a free installation service, and any user can find Lyra-V4-EXL2-8bpw on GitHub to install. At the same time, huggingface.co provides the effect of Lyra-V4-EXL2-8bpw install, users can directly use Lyra-V4-EXL2-8bpw installed effect in huggingface.co for debugging and trial. It also supports api for free installation.