Gemma-4-31B-Glistening-Gem-v1.0

Creative model

View on Hugging FaceBack to Models

Hourly Usage

Performance Metrics

Median Total Time

42.04s

Median TTFT

8.96s

Median Prefill TPS

1416.32

Median Gen TPS

19.82

Model Information

Context Size

262144

Quantization

r64

Engine

vllm

Creation Method

LoRA Finetune

Model Type

Gemma31B

Chat Template

Gemma4

Reasoning

Yes

Vision

Yes

Parameters

31B

Added At

7/25/2026


base_model:

  • zerofata/G4-MeroMero-31B
  • BeaverAI/Artemis-31B-v1h-GGUF
    llmfan46/gemma-4-Ortenzya-The-Creative-Wordsmith-31B-it-uncensored-heretic-GGUF library_name: transformers tags:
  • mergekit
  • merge
  • not-for-all-audiences license: apache-2.0 language:
  • en

Glistening-Gem-31B-v1.0

An Experimental Merge  ·  31B  ·  Apache 2.0

This is an experimental merge of BeaverAI/Artemis-31B-v1h-GGUF, zerofata/G4-MeroMero-31B, and llmfan46/gemma-4-Ortenzya-The-Creative-Wordsmith-31B-it-uncensored-heretic.

TheDrummer (creator of Artemis) has not released the full precision FP16 weights for the Artemis v1h model, so I had to get creative with a Q8_0 GGUF file to produce this merge. There is some precision loss in doing that, but Q8_0 is already pretty close to FP16 in practice, and I think the noise effectively washes out during the merge process.

This merge recipe grew out of my initial experiment with sophosympatheia/Mero-Artemis-31B-v0.3.1 which was positively received by the community.

Glimmer has more diversity in its prose and style compared to Mero-Artemis thanks to Ortenzya's influence. However, Mero-Artemis is overall more stable.

Known Issues

This model requires conservative sampler settings to maintain coherence as the conversation grows in context length. It will produce erroneous outputs if you run it with "aggressive" sampler settings for creativity, so dial it back if you encounter misspellings, grammatical errors, or illogical outputs.

I'd say this model is right on the cusp of instability, but that puts it in an interesting position.

IF YOU GET WEIRD OUTPUTS, TURN DOWN THE HEAT

Sampler Tips

You can use the master import JSON in this repo (Glimmer_SillyTavern_Master_Import.json) to deploy the conservative sampler settings below, which are likely to be compatible with more backend/frontend combos. I recommend using these values as a starting point for your own experiments. It's not like the model falls apart if you deviate from these settings, but they should be a reliable starting point for most creative tasks.

Conservative Settings

Run these settings as a starting point. Raise temperature and lower Min-P if you want more creativity.

Temp 0.7
Min-P 0.2
Adaptive-P Target 0.6
Adaptive-P Decay 0.5
DRY Mult. 0.8
DRY Base 1.8

Prompting Tips

You can download the Glistening-Gem_SillyTavern_Master_Import.json file from this repo and import it directly into SillyTavern to get system prompt, chat template, and sampler settings all in one go.

Donations

Donations

If you feel like saying thanks with a donation, I'm on Ko-Fi

Quantizations

Please see the sidebar of the model card where a link to quantizations can be found, or click here for the list of them.

License

Apache 2.0, inherited down from Gemma.

Merge Details

This is a merge of pre-trained language models created using mergekit.

Merge Method

This model was merged using the DELLA merge method using gemma-4-31B-it as a base.

Models Merged

The following models were included in the merge: