Median Total Time
28.90s
Median TTFT
6.56s
Median Prefill TPS
1058.80
Median Gen TPS
6.25
Context Size
262144
Quantization
r64 on INT8
Engine
vllm
Creation Method
Merge
Model Type
Gemma31B
Chat Template
Gemma4
Reasoning
Yes
Vision
Yes
Parameters
31B
Added At
9/26/2026
license: apache-2.0 tags:
"When the stories bleed out"
Novelist end up to be raw and experimental. So I crank it up with more models, infusing creativity and worked on stability. Also added lm_head on top with StyleTune.
Below is the exact mergekit_config.yml recipe used to synthesize this model:
I took Novelist idea about description focus and remade it, so at the end less slop survive.
merge_method: dare_ties
base_model: F:\AI\Merge\Gemma-4-it
tokenizer_source: union
parameters:
lambda: 1.0
dtype: bfloat16
models:
- model: F:\AI\Merge\G4-Gutenberg
parameters:
density: [0.50, 0.50, 0.50, 0.40, 0.45]
weight: [0.40, 0.40, 0.40, 0.40, 0.40]
- model: F:\AI\Merge\Melinoe
parameters:
density: [0.30, 0.30, 0.30, 0.30, 0.35]
weight: [0.30, 0.30, 0.30, 0.20, 0.40]
- model: F:\AI\Merge\Glimmer
parameters:
density: [0.20, 0.20, 0.20, 0.30, 0.20]
weight: [0.30, 0.30, 0.30, 0.40, 0.20]
The first stage was fine, but lacks of consistency, so I glued it with using two smart models.
models:
- model: F:\AI\Merge\NovelistX
- model: F:\AI\Merge\GarnetV2
- model: F:\AI\Merge\Gemopus
merge_method: model_stock
base_model: F:\AI\Merge\Gemma-4-it
dtype: bfloat16
tokenizer_source: base