Pantheon-Reasoning-27B: A Dense Reasoning RP Model from Gryphe

Gryphe has released Pantheon-Reasoning-27B, a fine-tuned reasoning model for roleplay built on llmfan46/Qwen3.6-27B-uncensored-heretic-v2-Native-MTP-Preserved. The model aims to bring structured reasoning to character work — weighing tone, planning narrative beats, and considering how a character would actually respond before generating a line.
The training data composition (all with full reasoning traces):
- Pantheon data (~28%) — core roleplay corpus with back-generated reasoning traces
- Opus-4.6-Reasoning-24k (~21%) — cleaned Claude Opus 4.6 reasoning traces for STEM, coding, and instruction-following
- WorldSim data (~16%) — long-form Opus 4.6 narrative roleplay with native reasoning, mainly third-person present tense
- Text adventure data (~16%) — interactive fiction and text adventure content with back-generated reasoning
- General roleplay data (~16%) — varied roleplay transcripts with back-generated reasoning
- Tiamat data (~3%) — character/RP dataset from Tiamat-24B-Magistral with multi-step improvement pipeline, reasoning back-generated per exchange
The model was trained with preserve_thinking: true, so thinking tags remain active across all assistant turns in multi-turn conversations — not just the first.
GGUF quants are available for local inference. The base model choice (Qwen 3.6 27B) was intentional for refusal reduction and writing capability. Gryphe notes they considered Gemma 4 31B but found it “an absolute pain to train” due to architectural quirks.
📖 Read the full source: r/LocalLLaMA
👀 See Also

OpenClaw 2026.4.29 Broken – Downgrade to 2026.2.6
OpenClaw version 2026.4.29 is broken with random errors, slow CLI, double replies. Downgrade to 2026.2.6 to fix.

Claude Code v2.1.116: Performance improvements, terminal fixes, and security updates
Claude Code v2.1.116 delivers significant performance improvements including up to 67% faster /resume on 40MB+ sessions, smoother terminal scrolling, and faster MCP startup. The release also fixes terminal rendering issues, adds security protections for dangerous path operations, and resolves multiple bugs affecting slash commands and plugin management.

David Silver's Ineffable Intelligence Raises $1.1B for RL-Based Superlearner Without Human Data
Ineffable Intelligence, founded by DeepMind alum David Silver, raised $1.1B at a $5.1B valuation to build a reinforcement learning-based 'superlearner' that discovers knowledge without human data.

Google donates Agent Payments Protocol (AP2) to FIDO Alliance, releases v0.2 with 'Human Not Present' payments
Google is donating the Agent Payments Protocol (AP2) to the FIDO Alliance, and releasing v0.2 with support for autonomous 'Human Not Present' payments and a new Verifiable Intent standard co-developed with Mastercard.