Mistral launches Medium 3.5 and cloud coding agents

Mistral has released Medium 3.5, a 128B open-weights model now in public preview and set as the default in Mistral Vibe and Le Chat. The company also introduced remote coding agents in Vibe and a new Work mode in Le Chat for multi-step tasks.

Mistral launches Medium 3.5 and cloud coding agents

Mistral has launched Medium 3.5, a new flagship model that combines instruction-following, reasoning, and coding in one 128B dense model. The company said the model is available in public preview, ships as open weights under a modified MIT license, and can be self-hosted on as few as four GPUs.

The new model is now the default in Mistral Vibe and Le Chat. Mistral said Medium 3.5 is built for long-running coding and productivity tasks, and that it can adjust its reasoning effort per request so the same model can handle a quick chat response or a longer agentic run.

⚡ New to this?

This matters because Mistral is pushing more of the work from a local laptop into cloud-based agents that can keep going after the user walks away. A coding agent is software that can read tasks, call tools, and make changes with some autonomy, while a model is the AI system doing the reasoning behind it.

For non-experts, the big change is that the assistant is no longer just answering questions. It can now work through longer jobs across code, email, calendars, and internal tools, with approvals built in for sensitive steps.

🦞 OpenClaw angle

If you run self-hosted agent workflows, treat Medium 3.5 as a model to benchmark for long-running, tool-heavy jobs rather than only chat. Test sandboxing, approval checkpoints, and audit visibility before letting an agent touch GitHub, Jira, Slack, or databases.

If you build automation around coding tasks, design for parallel remote runs and resumable sessions, since Mistral is explicitly carrying task state and approvals across local-to-cloud handoff. Also check whether your own workflow can emit structured diffs, tool-call logs, and PRs, since that is the output style Mistral says this model handles well.

Mistral also said the model is powering two new agent features. In Vibe, coding agents can now run remotely in the cloud, start from either the Vibe CLI or Le Chat, and keep working while the user steps away.

The company said remote sessions can run in parallel, which means users are no longer the bottleneck on every step the agent takes. Ongoing local CLI sessions can also be “teleported” to the cloud, with session history, task state, and approvals carried over.

According to Mistral, the remote runtime shows file diffs, tool calls, progress states, and questions while the agent runs. When a task finishes, the agent can open a pull request on GitHub and notify the user so they can review the result rather than each individual action.

Mistral said Vibe is designed for high-volume, well-defined engineering work such as module refactors, test generation, dependency upgrades, CI investigations, and bug fixes. The system plugs into GitHub, Linear, Jira, Sentry, Slack, and Teams, and each coding session runs in an isolated sandbox that allows broad edits and installs.

The company also introduced Work mode in Le Chat, which it described as a preview feature for complex multi-step tasks. Mistral said Work mode uses a new agent backed by Medium 3.5 and a new harness, allowing the assistant to read and write, use several tools at once, and keep working until the task is complete.

Mistral said Work mode is meant for cross-tool workflows such as catching up across email, messages, and calendar in one run, preparing meeting briefings, researching topics across the web and internal documents, and triaging inboxes or drafting replies. The company said connectors are on by default in Work mode so the agent can use documents, mailboxes, calendars, and other systems for context.

According to Mistral, every action the agent takes is visible, including tool calls and reasoning rationale. The company said Le Chat will ask for explicit approval, based on user permissions, before sensitive actions such as sending messages, writing documents, or modifying data.

Mistral said Medium 3.5 scored 77.6% on SWE-Bench Verified and 91.4 on τ³-Telecom. The company also said the model is available through its API at $1.5 per million input tokens and $7.5 per million output tokens, and that it is hosted for prototyping on NVIDIA GPU-accelerated endpoints on build.nvidia.com as well as via NVIDIA NIM.

The release is available today in Mistral Vibe and Le Chat on Pro, Team, and Enterprise plans.

Source: r/LocalLLaMA ↗

More from AI News