Mistral Medium 3.5 128B Released: Dense Model with Configurable Reasoning and Vision

Mistral AI has released Mistral Medium 3.5 (128B), a dense transformer model that replaces Mistral Medium 3.1 and Magistral in Le Chat, and Devstral 2 in their coding agent Vibe. It's a single set of weights handling instruction-following, reasoning, and coding.
Key Features
- Dense 128B parameters — not Mixture of Experts.
- 256k context window for long inputs.
- Multimodal input: accepts text and images; outputs text only. Vision encoder trained from scratch to handle variable sizes and aspect ratios.
- Configurable reasoning effort: toggle per request between instant reply (
none) and deep reasoning (high). - Native function calling and JSON output for agentic workflows.
- Multilingual: supports English, French, Spanish, German, Italian, Portuguese, Dutch, Chinese, Japanese, Korean, Arabic, and others.
- Strong system prompt adherence.
Recommended Settings
- Reasoning effort:
nonefor quick replies;highfor complex prompts and agentic usage (e.g.,reasoning_effort="high"). - Temperature: 0.7 with
highreasoning; 0.0–0.7 withnonedepending on desired creativity.
License
Released under a Modified MIT License — open-source for commercial and non-commercial use, with exceptions for large revenue companies.
GGUF Quantizations Available
Unsloth has published a GGUF version on Hugging Face: unsloth/Mistral-Medium-3.5-128B-GGUF
This model is relevant for developers running local AI coding agents, particularly those needing high-quality instruction following, reasoning, and vision in a single dense model with a large context window.
📖 Read the full source: r/LocalLLaMA
👀 See Also

China Blocks Meta's Acquisition of AI Startup Manus
China's government blocked Meta's proposed acquisition of AI startup Manus, citing national security concerns. The deal was reportedly valued at over $1 billion.

OpenAI's $10B PE Joint Venture: What It Means for AI Deployment
OpenAI finalizes a $10 billion joint venture with private equity firms to scale AI infrastructure and enterprise deployment, as reported by Bloomberg.

Anthropic Removes Gmail Message Body Access from Claude Connector
Anthropic has removed the gmail_read_message and gmail_search_messages tools from the Gmail connector, replacing them with get_thread and search_threads that no longer return message bodies or attachment content.

IDP Leaderboard benchmark shows Claude Sonnet 4.6 matches Opus 4.6 for document AI tasks
The IDP Leaderboard tested 16 AI models on 9,000+ documents across OCR, table extraction, key extraction, visual QA, handwriting, and long documents. Claude Sonnet 4.6 scored 80.8 overall, essentially matching Opus 4.6 at 80.3, while Haiku 4.5 scored 69.6.