Median Total Time
22.01s
Median TTFT
1.25s
Median Prefill TPS
1789.55
Median Gen TPS
34.43
Context Size
262144
Quantization
r128 on INT8
Engine
vllm
Creation Method
Finetune
Model Type
Qwen38
Chat Template
Qwen3.5
Reasoning
Yes
Vision
Yes
Parameters
27B
Added At
9/28/2026
thumbnail: >- https://cdn-uploads.huggingface.co/production/uploads/634262af8d8089ebaefd410e/LzXxeS6Rkun66x_UEoDKi.jpeg license: apache-2.0 language:
The newest member of your roleplay stack
A roleplay-focused finetune for improved prose and more creative reasoning.
An RP-focused finetune of Qwen 3.8 27B for improved prose and better (or at least more creative) reasoning.
Here's some links to support Fizz, the member who trained this model:
Support Fizz on Ko-fiAs well as her Monero address:
The original chat template has had its defaults adjusted. If you notice anything, it was intentional.
Chat format: standard Qwen3.8-style ChatML. A chat template mismatch is the one thing that visibly disappoints her.
Samplers: the author ran her comfortably at temperature 1.0 to 1.25 with either min_p 0.1 or top_p 0.95. Others found success at 0.7 temp and nothing else. Iunno!