CultriX / Qwen2.5-14B-Hyperionv3

huggingface.co
Total runs: 22
24-hour runs: 0
7-day runs: -4
30-day runs: 3
Model's Last Updated: February 07 2025
text-generation

Introduction of Qwen2.5-14B-Hyperionv3

Model Details of Qwen2.5-14B-Hyperionv3

merge

This is a merge of pre-trained language models created using mergekit .

Merge Details
Merge Method

This model was merged using the della_linear merge method using CultriX/Qwen2.5-14B-Wernickev3 as a base.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:





merge_method: della_linear
base_model: CultriX/Qwen2.5-14B-Wernickev3
dtype: bfloat16
out_dtype: bfloat16
parameters:
  epsilon: 0.009        # Further reduced for ultra-fine parameter scaling.
  lambda: 1.6            # Increased to emphasize significant model contributions.
  normalize: true       # Balances the parameter integration for stability.
  rescale: true         # Enabled to align parameter scales across models.
  int8_mask: false      # Disabled to allow full-precision computations for enhanced accuracy.
  density: 0.90          # Balanced density for optimal generalization and performance.
  
adaptive_merge_parameters:
  task_weights:
    tinyArc: 1.6            # Prioritizes logical reasoning improvements.
    tinyHellaswag: 1.5      # Strengthened contextual understanding and consistency.
    tinyMMLU: 1.8           # Enhanced domain knowledge for multitask benchmarks.
    tinyTruthfulQA: 1.9     # Maximized for accurate factual reasoning and QA.
    tinyTruthfulQA_mc1: 1.75 # Increased focus for multiple-choice reasoning.
    tinyWinogrande: 1.75    # Advanced reasoning and contextual prediction improvement.
    IFEval: 2.15            # Enhanced instruction-following tasks boosted by multitask contributors.
    BBH: 1.95               # Further improved for complex reasoning tasks.
    MATH: 2.45              # Highest priority, focusing on mathematical excellence.
    GPQA: 2.1               # Boosted graduate-level QA capabilities.
    MUSR: 1.9               # Nuanced multi-step reasoning strengthened further.
    MMLU-PRO: 1.9           # Maximized domain multitask performance.
  smoothing_factor: 0.035   # Further reduced for precise task-specific blending.

gradient_clipping:
  CultriX/Qwen2.5-14B-Wernickev3: 0.88  # Increased for enhanced stability.
  CultriX/Qwenfinity-2.5-14B: 0.85      # Adjusted for consistent multitask integration.
  djuna/Q2.5-Veltha-14B-0.5: 0.91   # Maintained advanced reasoning contributions.
  CultriX/SeQwence-14B-EvolMerge: 0.88   # Generalist multitask support remains stable.
  qingy2024/Fusion4-14B-Instruct: 0.93  # Mathematically focused tasks maximized.
  CultriX/Qwen2.5-14B-Emerged: 0.88      # Increased for logical reasoning enhancements.
  sometimesanotion/Lamarck-14B-v0.6: 0.89 # Balanced multi-step reasoning contributions.
  allknowingroger/QwenSlerp5-14B: 0.87   # Contextual and logical reasoning integration refined.

models:
  - model: CultriX/Qwen2.5-14B-Wernickev3
    parameters:
      weight: 0.30       # Core backbone for multitask reasoning.
      density: 0.75      # Further increased to preserve critical reasoning parameters.
  
  - model: CultriX/Qwenfinity-2.5-14B
    parameters:
      weight: 0.25       # Comprehensive multitask performer.
      density: 0.65
  
  - model: djuna/Q2.5-Veltha-14B-0.5
    parameters:
      weight: 0.25       # Advanced reasoning support for GPQA and MUSR.
      density: 0.74
  
  - model: CultriX/SeQwence-14B-EvolMerge
    parameters:
      weight: 0.20       # Enhanced contributions to BBH and MUSR.
      density: 0.55
  
  - model: qingy2024/Fusion4-14B-Instruct
    parameters:
      weight: 0.19       # Mathematical reasoning priority.
      density: 0.77
  
  - model: CultriX/Qwen2.5-14B-Emerged
    parameters:
      weight: 0.21       # Maintains overall task performance with balanced strengths.
      density: 0.72      # Increased for better integration.
  
  - model: CultriX/Qwen2.5-14B-Broca
    parameters:
      weight: 0.16       # Logical reasoning and factual QA enhancements.
      density: 0.68      # Increased to better support Broca's specialized tasks.
  
  - model: sometimesanotion/Lamarck-14B-v0.6
    parameters:
      weight: 0.15       # Multi-step reasoning tasks contributor.
      density: 0.63      # Slight increase for better integration.
  
  - model: allknowingroger/QwenSlerp5-14B
    parameters:
      weight: 0.16       # Contextual reasoning improvements.
      density: 0.64      # Increased for enhanced performance.


Runs of CultriX Qwen2.5-14B-Hyperionv3 on huggingface.co

22
Total runs
0
24-hour runs
2
3-day runs
-4
7-day runs
3
30-day runs

More Information About Qwen2.5-14B-Hyperionv3 huggingface.co Model

Qwen2.5-14B-Hyperionv3 huggingface.co

Qwen2.5-14B-Hyperionv3 huggingface.co is an AI model on huggingface.co that provides Qwen2.5-14B-Hyperionv3's model effect (), which can be used instantly with this CultriX Qwen2.5-14B-Hyperionv3 model. huggingface.co supports a free trial of the Qwen2.5-14B-Hyperionv3 model, and also provides paid use of the Qwen2.5-14B-Hyperionv3. Support call Qwen2.5-14B-Hyperionv3 model through api, including Node.js, Python, http.

Qwen2.5-14B-Hyperionv3 huggingface.co Url

https://huggingface.co/CultriX/Qwen2.5-14B-Hyperionv3

CultriX Qwen2.5-14B-Hyperionv3 online free

Qwen2.5-14B-Hyperionv3 huggingface.co is an online trial and call api platform, which integrates Qwen2.5-14B-Hyperionv3's modeling effects, including api services, and provides a free online trial of Qwen2.5-14B-Hyperionv3, you can try Qwen2.5-14B-Hyperionv3 online for free by clicking the link below.

CultriX Qwen2.5-14B-Hyperionv3 online free url in huggingface.co:

https://huggingface.co/CultriX/Qwen2.5-14B-Hyperionv3

Qwen2.5-14B-Hyperionv3 install

Qwen2.5-14B-Hyperionv3 is an open source model from GitHub that offers a free installation service, and any user can find Qwen2.5-14B-Hyperionv3 on GitHub to install. At the same time, huggingface.co provides the effect of Qwen2.5-14B-Hyperionv3 install, users can directly use Qwen2.5-14B-Hyperionv3 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Qwen2.5-14B-Hyperionv3 install url in huggingface.co:

https://huggingface.co/CultriX/Qwen2.5-14B-Hyperionv3

Url of Qwen2.5-14B-Hyperionv3

Qwen2.5-14B-Hyperionv3 huggingface.co Url

Provider of Qwen2.5-14B-Hyperionv3 huggingface.co

CultriX
ORGANIZATIONS

Other API from CultriX

huggingface.co

Total runs: 70
Run Growth: 1
Growth Rate: 1.43%
Updated:February 21 2024
huggingface.co

Total runs: 56
Run Growth: 0
Growth Rate: 0.00%
Updated:January 27 2024
huggingface.co

Total runs: 50
Run Growth: -8
Growth Rate: -16.00%
Updated:January 27 2024
huggingface.co

Total runs: 35
Run Growth: 10
Growth Rate: 28.57%
Updated:November 27 2024
huggingface.co

Total runs: 20
Run Growth: 0
Growth Rate: 0.00%
Updated:February 21 2024
huggingface.co

Total runs: 20
Run Growth: 0
Growth Rate: 0.00%
Updated:November 27 2024
huggingface.co

Total runs: 14
Run Growth: 0
Growth Rate: 0.00%
Updated:February 08 2024