IT之家 · 智能时代T2·78 pts
JetBrains Releases Mellum2.1: RL Enhances Agentic Coding with Throughput Nearly Double Qwen3.5-9B
Original: JetBrains 编程 AI 模型 Mellum2.1 发布:高负载推理吞吐量近 Qwen3.5-9B 两倍
- Mellum2.1 retains the 12B MoE architecture (2.5B active parameters) under Apache 2.0 license.
- Key upgrade involves expanding RL from a final stage to a main training phase, improving root cause analysis and fix verification.
- High-load inference throughput is nearly double that of Qwen3.5-9B, with single-request speed up by ~1.6x.