SpoomplesMaxx SFT is a
supervised fine-tuning
adapter trained on top of
aimeri/spoomplesmaxx-base-gemma3-27b-4500
— the end-run candidate checkpoint from the SpoomplesMaxx CPT run on Gemma 3 27B PT.
This stage teaches
system-prompt following
,
persona consistency
, and
structured narrative
reasoning
for collaborative fiction and character roleplay. It is implemented as a LoRA adapter
and requires the CPT base to run.
This model was trained on checkpoint 4500 out of ~5700 steps. A future run of the completed CPT model will follow.
Part of the
SpoomplesMaxx
project — a hobbyist ML research effort
focused on creative writing and roleplay capability in open base models.
CPT (spoomplesmaxx-base-gemma3-27b-4500)
↓
SFT ← you are here
↓
DPO (planned)
Uses
Direct Use
This model is suitable for:
Character roleplay and collaborative fiction with persistent persona
System-prompt-driven character following (multi-character mode-switching)
Structured narrative generation with explicit scene/continuity state tracking
Multilingual creative writing in English and Brazilian Portuguese
Load with a chat-capable inference frontend (e.g. SillyTavern, Oobabooga,
llama.cpp with a compatible template).
Downstream Use
Intended as an intermediate checkpoint in the full pipeline. A DPO alignment stage targeting
preference learning over character voice and narrative consistency is planned.
Can also serve as a base for further LoRA fine-tuning on custom character datasets or
domain-specific creative writing corpora.
Out-of-Scope Use
General-purpose instruction following or assistant tasks — the model is optimized for
character roleplay and creative writing, not factual Q&A or task completion
Production deployment without further alignment — this is a hobbyist research artifact
Safety-critical applications — the base CPT model is uncensored; this SFT stage does not
add safety alignment
Training Details
Training Data
Two datasets were used:
Dataset
Rows
Format
Description
spoomplesmaxx-olivia-sft
~78,000
DanChat-2
Persona dataset teaching system-prompt following and character consistency. Derived from a paraphrased and sycophancy-stripped source. Features Olivia Costa — a Brazilian zoologist-turned-ML-hobbyist — as the anchor persona, alongside a broad library of character cards for mode-switching.
spoomplesmaxx-rp-reasoning
~9,600
ShareGPT
Structured RP reasoning traces in SCENE / CHARACTERS / CONTINUITY / THREADS format. Teaches explicit narrative state tracking and continuity management across multi-turn fiction.
Olivia Costa
— the primary persona in the Olivia dataset — is warm but direct,
anti-sycophantic, and trained to present as human. She serves as both the default persona and
the calibration signal for tone and voice across the character library.
Training Procedure
SFT was performed on the CPT checkpoint using LoRA, with vision components frozen (inherited
from base). Training targets text layers only.
Training Hyperparameters
Training regime:
bf16 mixed precision
Adapter type:
LoRA
Training framework:
TRL + Unsloth
Evaluation
No formal benchmarks have been run on this adapter. Evaluation is qualitative — persona
consistency across multi-turn conversations, system-prompt adherence across character cards,
and narrative coherence in the rp-reasoning traces.
If you run evaluations or have qualitative findings, please open a discussion.
Earlier CPT runs on SmolLM3 3B, GLM-4-32B, and Qwen3-14B are archived or available
separately. Gemma 3 27B was selected for this stage based on superior out-of-the-box creative
writing quality and multilingual coverage.
Model Card Authors
aimeri
Model Card Contact
Open a discussion on the repository page.
Runs of aimeri spoomplesmaxx-27b-4500 on huggingface.co
19
Total runs
4
24-hour runs
7
3-day runs
6
7-day runs
9
30-day runs
More Information About spoomplesmaxx-27b-4500 huggingface.co Model
spoomplesmaxx-27b-4500 huggingface.co
spoomplesmaxx-27b-4500 huggingface.co is an AI model on huggingface.co that provides spoomplesmaxx-27b-4500's model effect (), which can be used instantly with this aimeri spoomplesmaxx-27b-4500 model. huggingface.co supports a free trial of the spoomplesmaxx-27b-4500 model, and also provides paid use of the spoomplesmaxx-27b-4500. Support call spoomplesmaxx-27b-4500 model through api, including Node.js, Python, http.
spoomplesmaxx-27b-4500 huggingface.co is an online trial and call api platform, which integrates spoomplesmaxx-27b-4500's modeling effects, including api services, and provides a free online trial of spoomplesmaxx-27b-4500, you can try spoomplesmaxx-27b-4500 online for free by clicking the link below.
aimeri spoomplesmaxx-27b-4500 online free url in huggingface.co:
spoomplesmaxx-27b-4500 is an open source model from GitHub that offers a free installation service, and any user can find spoomplesmaxx-27b-4500 on GitHub to install. At the same time, huggingface.co provides the effect of spoomplesmaxx-27b-4500 install, users can directly use spoomplesmaxx-27b-4500 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
spoomplesmaxx-27b-4500 install url in huggingface.co: