A reasoning-focused fine-tune of
Qwen 3.5 27B
, trained on a small, high-quality dataset curated through a custom
Drift Diffusion Modeling (DDM)
pipeline. Second iteration of the Ornstein series with improved segment-level reasoning modeling.
I'm a PhD student in visual neuroscience at the University of Toronto who also happens to spend way too much time fine-tuning, merging, and quantizing open-weight models on rented H100s and a local DGX Spark. All training compute is self-funded — balancing GPU costs against a student budget. If my uploads have been useful to you, consider buying a PhD student a coffee. It goes a long way toward keeping these experiments running.
Most reasoning fine-tunes throw large volumes of synthetic data at a base model and hope for the best. Ornstein takes the opposite approach: every single training example passed through a multi-stage quality pipeline that measures whether a reasoning trace is actually
reasoning
or just generating tokens that look like reasoning.
The core insight is that language models frequently produce
degenerate reasoning
— long chains of text that superficially resemble deep thought (hedging, restating the problem, circling without progress) but carry little actual signal. The DDM pipeline detects and separates these from genuine premium reasoning traces, producing a training mix that teaches the model what good thinking actually looks like.
What's New in v2
Ornstein v2 builds on the original with improved
segment-level reasoning modeling
— the DDM pipeline now does more to model individual segments of the reasoning process, leading to finer-grained quality separation and better training signal.
Training Data at a Glance
Drift Score Distribution:
The DDM pipeline assigns each reasoning trace a drift score. Premium traces cluster low, degenerate traces cluster high, with a fitted threshold cleanly separating the two pools.
Category Mix:
Math-heavy dataset with code, science, and logic rounding it out.
Reasoning Depth:
Premium traces average substantially deeper thinking than degenerate traces, which tend to be shallow repetition that inflates token count without substance.
Difficulty x Pool:
The degenerate pool skews toward hard problems where models are most likely to loop or stall.
DDM Curation Pipeline
Drift Diffusion Modeling works by decomposing each reasoning trace into uniform segments and tracking how "reasoning quality" evolves across the trace. Each segment is scored on multiple dimensions that capture whether the model is mimicking cognitive progress — things like introducing new ideas, self-correcting, verifying intermediate results, and exploring alternative approaches.
These per-segment scores are accumulated into a drift trajectory. Premium traces maintain healthy trajectories throughout. Degenerate traces accumulate deficit as the model loops, repeats itself, or pads without substance — and the drift score crosses a threshold fitted via ROC analysis.
Ornstein uses
<think>...</think>
blocks for extended reasoning:
<think>
Let me work through this step by step...
[multi-phase reasoning with self-correction and verification]
</think>
[Final answer]
Intended Use
Ornstein-27B-v2 is designed for tasks that benefit from structured, multi-step reasoning — math, logic, code analysis, scientific problems, and complex question answering. The DDM curation specifically optimizes for traces that exhibit genuine cognitive progress rather than verbose restating.
Limitations
Single epoch training means the model retains most of the base Qwen 3.5 27B behavior; the fine-tune primarily shapes reasoning style rather than injecting new knowledge
The DDM pipeline optimizes for English reasoning traces; performance on other languages reflects the base model
Extended thinking can still occasionally loop on adversarial or highly ambiguous prompts
License
Apache 2.0
Citation
If you use Ornstein-27B-v2 or the DDM curation methodology in your work:
@misc{ornstein27bv2,
author = {DJLougen},
title = {Ornstein-27B-v2: DDM-Curated Reasoning Fine-Tune of Qwen 3.5 27B},
year = {2026},
publisher = {Hugging Face},
url = {https://huggingface.co/DJLougen/Ornstein-27B-v2}
}
Runs of DJLougen Ornstein-27B-v2 on huggingface.co
21
Total runs
0
24-hour runs
-33
3-day runs
-54
7-day runs
-201
30-day runs
More Information About Ornstein-27B-v2 huggingface.co Model
Ornstein-27B-v2 huggingface.co is an AI model on huggingface.co that provides Ornstein-27B-v2's model effect (), which can be used instantly with this DJLougen Ornstein-27B-v2 model. huggingface.co supports a free trial of the Ornstein-27B-v2 model, and also provides paid use of the Ornstein-27B-v2. Support call Ornstein-27B-v2 model through api, including Node.js, Python, http.
Ornstein-27B-v2 huggingface.co is an online trial and call api platform, which integrates Ornstein-27B-v2's modeling effects, including api services, and provides a free online trial of Ornstein-27B-v2, you can try Ornstein-27B-v2 online for free by clicking the link below.
DJLougen Ornstein-27B-v2 online free url in huggingface.co:
Ornstein-27B-v2 is an open source model from GitHub that offers a free installation service, and any user can find Ornstein-27B-v2 on GitHub to install. At the same time, huggingface.co provides the effect of Ornstein-27B-v2 install, users can directly use Ornstein-27B-v2 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.