Boulesis-26B-A4B

Composite Gemma 4 RP model (QK task arithmetic + fused LoRA).

The idea was to retain the model's core intelligence and knowledge, while diversifying its prose and making it more decisive. I also wanted to sharpen its attention to context so it could dig deeper into the character card, organically pulling lore and facts into the roleplay rather than just mirroring the user.

As a result, Boulesis moves beyond passive reactivity to genuinely advance the narrative, all while suffering zero catastrophic forgetting of its base intelligence.

YOU WILL GET THE BEST RESULTS WITH THINKING ON!

🏆 Best model for RP, ERP, and Dark RP in the 24–26B category, according to CaliperBench results.

Thinking mode, 7 sep. 2026

How it's built

Most RP merges blend whole models. This one edits one part of attention and leaves the rest alone.

Part Source
Body coder3101/gemma-4-26B-A4B-it-heretic (ARA-abliterated, layers 10–30)
lm_head copied from Gryphe/Gemma-4-26B-A4B-StyleTune-V2 (one tensor out of 659)
q_proj, k_proj base + α·(Pantheon-Reasoning-1.1 − unsloth/gemma-4-26B-A4B-it), α = 0.6
v_proj, o_proj LoRA r=32, alpha=64, 55 projections, baked at effective scale 0.26
Everything else untouched

q_proj and k_proj decide where the model looks. v_proj and o_proj decide what it carries back. That split comes from mechanistic interpretability work.

My guess was that RP reactivity is mostly an attention routing problem: the model over-weights the user's latest turn at the expense of the character card and prior context. To test this, I grafted the Q/K projections from Pantheon-Reasoning.

Using it

YOU WILL GET THE BEST RESULTS WITH THINKING ON!

Recommended settings:

Parameter Value
Temperature 1.0
Top-K 64
Repetition Penalty 1.05-1.1
Thanks to DifficultyThin8462
If the reasoning doesn't work when connecting GGUF ver. in KoboldCPP and SillyTavern

You need to force this in KoboldCPP. Go to the Content tab and enable these options.

image

For SillyTavern, it is recommended to set the template as shown:

image

Credits

Thanks to coder3101 and Gryphe for the fine-tunes, and the entire 26B-Suite team for their intellectual support. Speсial thanks for Naphula and redaihf. You guys are awesome!

Downloads last month
128
Safetensors
Model size
27B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for SubMaroon/Boulesis-26B-A4B