Median Total Time
12.55s
Median TTFT
7.17s
Median Prefill TPS
1005.65
Median Gen TPS
24.73
Context Size
262144
Quantization
r64 on INT8
Engine
vllm
Creation Method
Merge
Model Type
Gemma31B
Chat Template
Gemma4
Reasoning
Yes
Vision
Yes
Parameters
31B
Added At
7/13/2026
license: apache-2.0 base_model: google/gemma-4-31B-it tags:
This repository contains merged Hugging Face weights for an RP-enhanced version of Gemma 4 31B it.
The model is based on the uncensored Heretic SDFT line, then further tuned for roleplay behavior: stronger scene continuation, better character and world continuity, more active narrative pressure, and reduced generic LLM phrasing.
This model is provided for research and evaluation purposes only.
The underlying Heretic SDFT base intentionally has weakened or removed guardrails compared with standard aligned chat models. Users are responsible for how they run, prompt, deploy, and distribute outputs from this model.
Do not use this model to generate harmful, illegal, abusive, exploitative, or otherwise unsafe content.
google/gemma-4-31B-it.Compared with the previous Heretic SDFT RP checkpoints, this version showed statistically meaningful improvements in roleplay-oriented evaluation:
This model is intended for:
This is an experimental release. Outputs may still be unstable, repetitive, unsafe, or stylistically uneven depending on sampling settings and prompt structure.
Recommended starting inference settings:
temperature: 0.8-0.95
top_p: 0.95
top_k: 64
Special thanks to the people who helped make this release possible: