Nimbz/Split-31B
Nimbz/Split-31B is an experimental 31 billion parameter language model based on the Gemma 4 architecture, created by Nimbz through a two-phase della_linear merge process. This model is specifically optimized for strong instruction adherence, persona tracking, and maintaining world-state stability over long sessions, making it suitable for structured output requirements. It offers a context length of 32768 tokens and is designed for roleplay scenarios requiring consistent character and format adherence.
Loading preview...
Nimbz/Split-31B: A Dual-Personality Gemma 4 Merge
Nimbz/Split-31B is an experimental 31 billion parameter language model built upon the Gemma 4 architecture, developed by Nimbz using a two-phase della_linear merge method. This model is one of two distinct 'heads' from the same core, each optimized for different output characteristics.
Key Capabilities & Design Philosophy
Split-31B is engineered for strong instruction adherence, persona tracking, and maintaining world-state stability throughout extended conversational sessions. Its development involved carefully adjusting specific layers of the underlying Gemma 4 models:
- Phase 1 (Comprehension): Focused on modifying
q_proj,k_proj(for attention mechanism) andgate_proj,up_proj(for knowledge routing) from three base models:Melinoe-VL-heretic(for persona tracking and world-state),Glistening-Gem v2.1(for format adherence and in-character responses), andMeroMero v2(for output variety). - Phase 2 (Voice): Adjusted
v_proj(for expression) ando_proj,mlp.down_proj(for output writing) usingscotoma-2(to counter Gemma's inherent 'tics'),Dark-Thoughts V2(for a 'darker lean' and less sanitization),Glistening-Gem v2.0(for self-repetition and coherence), andMelinoe-VL-heretic(for 'willingness').
When to Use Split-31B
Choose Nimbz/Split-31B if your application requires:
- Strict instruction following: Ideal for tasks demanding precise formatting, such as trackers or specific header/footer blocks.
- Consistent character and persona: Excels in roleplay scenarios where maintaining a stable persona and world-state is crucial.
- Long-session coherence: Designed to stay on track and provide structured output over extended interactions.
For use cases requiring more creativity, a 'darker lean,' or 'lewdness' at the expense of some prompt adherence, the companion model, Split-Untied-31B, is recommended. The model supports a context length of 32768 tokens.