aifollow.news 搜索
返回 vLLM 发布记录
vLLM 发布记录· · 仅日期

v0.22.1

来源摘录

## Highlights This release features 8 commits from 6 contributors (1 new)! v0.22.1 is a patch release on top of v0.22.0 with targeted bug fixes plus a couple of additions: new model support for JetBrains' Mellum v2, zentorch-accelerated quantized linear inference on AMD Zen CPUs, and fixes for multi-node Ray data-parallel serving, DeepSeek-V4 initialization, and a few model-loading regressions. ### Model Support * New model: JetBrains' **Mellum v2**, an open-weights Mixture-of-Experts code-generation model (#43992). * **DeepSeek-V4**: resolve a CUTLASS `fmin` compatibility issue that broke initialization (0decac0d). * Fix `OlmoHybridForCausal

完整正文暂未获取。

前往原始出处阅读
发现内容有误?提交纠错