Create

Sign in to ReadmeX

or

How we handle your data: Privacy Policy

ReadmeX
ReadmeX

A clearer picture, in a conversation.

Catch up on what matters, then ask a little deeper.

Your community and people briefings stay personal to you.

Story

JetBrains releases Mellum2.1 coding model

AI summary

JetBrains launched Mellum2.1, keeping Mellum2's 12B mixture-of-experts architecture with 2.5B active parameters and the Apache 2.0 license, and focusing the upgrade on agentic coding. JetBrains says reinforcement learning was expanded from a short post-training step into the main training stage, letting the model explore codebases, edit files, identify root causes of failing tests and draft and verify fixes. The model is available on Hugging Face, with GGUF builds for llama.cpp, Ollama and LM Studio and a vLLM MTP speculative-decoding component promised later.

Why it matters: It adds another open-weight option for coding agents that can run on local or self-hosted infrastructure, and the vendor-reported throughput edge could cut inference costs at scale.

JetBrainsMellum2.1Mellum2

Source textIT之家 AI · 3 min read
JetBrains 编程 AI 模型 Mellum2.1 发布:高负载推理吞吐量近 Qwen3.5-9B 两倍 - IT之家

首页

IT圈

最会买

设置

  • 日夜间

    随系统

    浅色

    深色

  • 主题色

    黑色

投稿

订阅

软媒应用

业界 手机 电脑 测评 视频 AI 苹果 iPhone 鸿蒙 软件

智车 数码 学院 游戏 直播 5G 微软 Win10 Win11 专题

首页 > 智能时代>人工智能

JetBrains 编程 AI 模型 Mellum2.1 发布:高负载推理吞吐量近 Qwen3.5-9B 两倍

2026/10/9 12:59:08 来源:IT之家 作者:故渊 责编:故渊

评论:

感谢IT之家网友 有鲫雪狐 的线索投递!

IT之家 10 月 9 日消息,JetBrains 昨日(10 月 8 日)发布博文,宣布上线 Mellum2.1 模型,重点增强智能体编程能力。该模型延续 Mellum2 的 12B 混合专家架构,拥有 2.5B 活跃参数,并继续采用 Apache 2.0 许可证发布。

Mellum2.1 主要升级预训练后的强化学习阶段,从短期收尾环节扩展为训练主体,并在数学、算法竞赛、科学、工具使用和软件工程等任务中补充训练数据。

JetBrains 为该模型搭建了内部强化学习环境基础设施,在训练期间启动数百万个沙盒,覆盖数千个环境。此外在训练前,团队还筛选了开放数据集,剔除测试缺陷、不可验证答案及难度失当的任务。

升级后,Mellum2.1 可探索代码库、编辑文件并检查自身修改。JetBrains 称,其最大改进出现在智能体编程领域,模型可识别失败测试的根因,起草修复方案并验证结果。

在性能方面,Mellum2.1 与 Mellum2 架构一致,速度保持不变,多 Token 预测(MTP)进一步提升响应速度。单请求场景下,MTP 使其速度提高约 1.6 倍。

Mellum2.1 compared with Mellum2, Qwen3.5-9B, and Gemma 4 E4B

JetBrains 将 Mellum2.1 与 Mellum2、Qwen3.5-9B 及 Gemma 4 E4B 在相同评估设置下比较。结果显示,高负载时 Mellum2.1 的推理吞吐 Token 数量接近 Qwen3.5-9B 的两倍,并在编程、算法竞赛、数学、工具调用和通用知识等维度均有提升。IT之家附上相关截图如下:

Output tokens per second on one H200 for Mellum2.1, Qwen3.5-9B, and Gemma 4 E4B

Mellum2.1 现已登陆 Hugging Face,支持本地或自有基础设施部署,让企业可将代码和数据保留在自身环境中。官方后续还将会推出面向 llama.cpp、Ollama 和 LM Studio 的 GGUF 版本,以及 vLLM 的 MTP 推测解码组件。

下载IT之家APP,签到赚金币兑豪礼

相关文章

关键词:JetBrains,AI

软媒旗下网站: IT之家 最会买 - 返利返现优惠券 iPhone之家 Win7之家 Win10之家 Win11之家

软媒旗下软件: 软媒手机APP应用 魔方 最会买 要知

Read the original →
Group chat

No comments yet. Start the conversation.